Tech jobs
Job listings
Research Scientist
About Company Founded in 2022, XG Tech is driving the future of smart vehicles. Its mission is to empower the digital transformation of automobiles, moving from distributed computing to a centralized, cross-domain…
AI Platform Engineer
Design and operate scalable AI inference platforms for production ML workloads, optimizing GPU utilization, LLM serving, and cloud-native infrastructure.
AI Engineer
Build and optimize production-grade AI systems using RAG pipelines, vector stores, and cloud-native tools to deliver secure, scalable solutions for federal environments.
Software Engineer Senior-Ai Engineer
Build and deploy production-grade LLM systems for a large bank, focusing on fine-tuning, optimization, RAG pipelines, and agentic workflows integrated with enterprise data.
Software Engineer Senior-Ai Engineer
Build and deploy production-grade LLM systems for a large bank, focusing on fine-tuning, RAG pipelines, and agentic workflows integrated with enterprise data.
Senior Manager, Technology Consulting
Are you passionate about squeezing every last drop of performance out of advanced hardware accelerators? Step into a role where you will shape the future of AI. In this role, you will drive the performance and…
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference
Builds and optimizes inference stacks for large-scale Apple foundation models, enabling AI features across services like Siri and Photos with low latency.
MLOps Engineer - AI Specialist
Designs and evaluates MLOps pipelines and GPU-accelerated training systems, focusing on JAX/PyTorch and custom kernel optimization for AI research labs.
Cloud Service Security Platform DevOps & Maintenance
Build and maintain the CI/CD, MLOps, and Internal Developer Platform that lets AI teams deploy models and services without opening tickets, using Kubernetes, Terraform, and observability tools.
Engineering Manager, AI Platform - Managed AI
Crusoe is on a mission to accelerate the abundance of energy and intelligence . As the only vertically integrated AI infrastructure company built from the ground up, we own and operate each layer of the stack — from…
Senior Software Engineer I, Inference
Senior engineer builds and optimizes CoreWeave’s Kubernetes-native AI inference platform, improving latency, throughput, and reliability to meet strict P99 SLAs while mentoring peers.
Senior Software Engineer, Inference
Senior engineer building and optimizing CoreWeave's Kubernetes-native AI inference platform to meet strict latency and reliability SLAs, using Python/Go, CUDA, and distributed systems.
Software Engineer, Inference AI/ML
Develops and optimizes AI model-serving systems on GPU infrastructure, focusing on latency, reliability, and cost while working with tools like Triton, vLLM, and Kubernetes.
Staff Software Engineer, Inference
Leads architecture and performance for CoreWeave’s Kubernetes-native AI inference platform, optimizing low-latency, high-throughput systems and GPU resource management.
Senior Software Engineer - GPU Kernel Authoring & Optimization
Write, profile, and optimize CUDA kernels for LLM inference to maximize throughput and minimize latency on NVIDIA GPUs, using DSLs like Triton or Mojo and benchmarking with MLPerf.
Compiler Engineer - Kernelize
Kernelize builds compiler backends for AI accelerators to enable and optimize Triton kernels for broad adoption for ML training and inference. We bridge the gap (disconnect) between AI optimizers and HW vendors'…
Runtime Engineer - Kernelize
Kernelize builds compiler backends for AI accelerators to enable and optimize Triton kernels for broad adoption for ML training and inference. We bridge the gap (disconnect) between AI optimizers and HW vendors'…
Senior ML Engineer
Build, fine-tune, and deploy deep-learning models end-to-end: train Transformers, optimise inference with NVIDIA’s stack, and package models into production-ready containers.
Старший Data Scientist, Возвраты ML
Senior ML engineer builds NLP and computer-vision models to automate returns, disputes and call-center reviews for a large marketplace.
Forward Deployed Engineer (AI & Infrastructure) [VT26-FDE] [Hong Kong] [Full Time]
On-site engineer deploying and optimizing AI systems for regulated clients using Kubernetes, Docker, and AI-driven automation tools like vibe coding and agent swarms.