Senior On-Premise LLM Inference & GPU Systems Consultant
Summary
Senior engineer building and maintaining on-premise LLM inference systems on NVIDIA H200 clusters and OpenShift AI, deploying open-source LLMs like Llama for private GenAI environments.