Tech jobs
Job listings
AI Platform Engineer
Design and operate scalable AI inference platforms for production ML workloads, optimizing GPU utilization, LLM serving, and cloud-native infrastructure.
AI Engineer
Build and optimize production-grade AI systems using RAG pipelines, vector stores, and cloud-native tools to deliver secure, scalable solutions for federal environments.
Software Engineer Senior-Ai Engineer
Build and deploy production-grade LLM systems for a large bank, focusing on fine-tuning, optimization, RAG pipelines, and agentic workflows integrated with enterprise data.
Software Engineer Senior-Ai Engineer
Build and deploy production-grade LLM systems for a large bank, focusing on fine-tuning, RAG pipelines, and agentic workflows integrated with enterprise data.
AI Engineer / AI-разработчик
Обязанности: Разрабатывать системы анализа здоровья: наше ключевое направление AI, который помогает пользователю понять своё состояние и вовремя дойти до врача; Внедрять AI-фичи в продукт и в компанию: интеграция LLM…
Staff Machine Learning Engineer - LLM Quantization & Deployment
Build and deploy quantized large language models for in-vehicle AI, focusing on PTQ, QAT, and low-bit inference to ensure numerical consistency and performance on XPENG’s Turing AI chip.
Senior Machine Learning Engineer - LLM Quantization & Deployment
Build and deploy quantized large language models for in-vehicle AI, focusing on PTQ, QAT, and low-bit inference to optimize performance on XPENG’s Turing AI chip.
Senior AI Solution Architect
Designs and validates AI accelerator systems (e.g., Gaudi, GPUs) for large-scale ML workloads, debugging hardware, firmware, and software layers while leading cross-functional teams.
AI Software Application Developer
Develops and optimizes distributed AI infrastructure and software for LLMs, vision AI, and robotics using Intel Xeon/GPU hardware.
AI Framework Software Engineer, PyTorch
Develop and optimize PyTorch-based AI frameworks for Intel hardware, focusing on distributed algorithms, kernel fusion, and GPU enablement to boost model performance.
AI Framework Software Engineer
Develops and optimizes AI frameworks like SGLang, implementing distributed algorithms and performance tuning for deep learning models across hardware backends.
AI Framework Software Engineer
Develop and optimize AI frameworks like vLLM and PyTorch, implementing distributed algorithms and performance enhancements for deep learning models on Intel hardware.
AI Framework Engineer - QA & Benchmarking
Build and maintain automated QA flows and benchmarks for AI frameworks like PyTorch and vLLM, validating performance across hardware and integrating tests into CI/CD pipelines.
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference
Builds and optimizes inference stacks for large-scale Apple foundation models, enabling AI features across services like Siri and Photos with low latency.
Account Solutions Architect - Engaged
Design and deploy AI infrastructure for greenfield customers, partnering with sales and engineering to scale LLM workloads on CoreWeave’s GPU-first cloud.
Senior Machine Learning Engineer
Build and maintain an internal AIOps/ML/LLM platform, including Kubernetes infrastructure, to power security analytics and automation workflows.
Senior AI Engineer
Build and deploy small/large language models to optimize telecom RAN networks, focusing on RAG pipelines, fine-tuning, and hallucination mitigation in a private-cloud environment.
Senior AI Engineer
Senior AI engineer designs, builds, and operates enterprise AI systems—from inference engines and GPU-accelerated platforms up to RAG pipelines and agent workflows—while mentoring teams and advising clients on production AI outcomes.
Senior Software Engineer 3 - (Python, Argo CD, Kubernetes)
Build and maintain AI inference infrastructure using Python, Kubernetes, and Argo CD to deploy and manage LLM services for a government customer.
Cloud Service Security Platform DevOps & Maintenance
Build and maintain the CI/CD, MLOps, and Internal Developer Platform that lets AI teams deploy models and services without opening tickets, using Kubernetes, Terraform, and observability tools.