Tech jobs
Job listings
Senior Site Reliability Engineer
Build and run large-scale, cloud-native systems that power Adobe’s AI features, including ML inference infrastructure, container orchestration, and automated patch management across AWS, Azure, and GCP.
Senior Engineer, Platform Engineering and Architecture
Lead the architecture and engineering of a bank-grade AI & Agentic Platform, designing agentic runtimes, LLM gateways, identity layers, and cloud-native infrastructure to support secure, scalable agent workflows across the enterprise.
Senior Artificial Intelligence/Machine Learning Engineer
Builds and optimizes AI/ML pipelines for automated code documentation, focusing on LLM inference, RAG, and NLI validation to process a 400K+ LOC industrial automation codebase with strict accuracy and latency targets.
AI Infrastructure Engineer
Optimize LLM inference performance on Intel GPUs by profiling bottlenecks, writing custom kernels, and contributing to open-source serving frameworks like vLLM and SGLang.
Full Stack Developer
Builds full-stack web apps in Python (Flask/Django) with React/TypeScript front ends, integrating LLM capabilities via RAG and processing large datasets with PostgreSQL, ArangoDB, and vector search.
Forward Deployed Engineer
Build and deploy AI inference systems with customers, writing production code and debugging across the full stack from requests to kernel dispatch.
Lead Software Engineer - LLM Ops Platform Reliability
Build and operate scalable LLM serving infrastructure using cloud and Kubernetes, ensuring reliability and performance for AI systems in production.
LLM Inference & GPU Systems Consultant
Optimize and deploy large-scale on-prem LLM inference systems using NVIDIA H200 GPUs, vLLM, TensorRT-LLM, and OpenShift AI for enterprise private GenAI environments.
Principal ML Ops Engineer (EMEA Remote)
Build and operate scalable ML inference platforms for an AI-native cloud startup, designing GPU-powered serving systems, deployment pipelines, and observability for real-time AI applications.
ML Ops Engineer (EMEA Remote)
Build and operate scalable ML inference platforms using vLLM/TGI/Triton to serve AI models with low latency and high GPU efficiency for a next-gen cloud startup.
Gen AI Engineer
Build and deploy generative AI agents and multi-agent systems using LangChain, LangGraph, MCP, and vector databases for a 12-month contract in Charlotte.
Principal Engineer, Agentic AI Systems
About Inflection AI Inflection AI is a Public Benefit Corporation empowering people with human-centered, emotionally intelligent AI. We’re shaping the future of AI by combining emotional intelligence (EQ) and raw…
AI Engineer (Managed Services)
Build and deploy enterprise LLM applications, RAG systems, and AI agents using open-source models (DeepSeek, Qwen, Kimi) and frameworks like LangChain and vLLM.
AI / LLM Inference Engineer
Engineer AI/LLM inference on GPU clusters: benchmark, tune, and optimize model serving with vLLM, Triton, or TensorRT-LLM to hit latency, throughput, and memory targets.
Staff AI Platform Engineer, Infrastructure Services
Build and scale SentinelOne’s AI Gateway infrastructure (Kong-based) to route, secure, and monitor AI coding assistant traffic, while operating self-hosted LLM stacks and driving reliability across Kubernetes and CI/CD systems.
Senior AI Platform Engineer, Infrastructure Services
Build and scale SentinelOne’s AI Gateway infrastructure (Kong-based) to route, secure, and monitor AI coding assistant traffic, while operating self-hosted LLM stacks and driving reliability across Kubernetes and CI/CD tooling.
Domain Architect AI Storage
Salary: $150 – $170 per hour About the Role The Domain Architect - AI Storage acts as the primary technical authority for the physical and logical lifecycle of high-performance data platforms across diverse client…
Senior ML Engineer (ИИ-агенты для бизнеса)
Разрабатываете NLP/мультимодальные пайплайны, RAG-системы и ИИ-агентов для бизнес-задач, интегрируете их в высоконагруженные сервисы банка.