Tech jobs
Job listings
AI/ML Architect
Design and deploy scalable, real-time AI systems including LLM inference pipelines, RAG, and vector databases using Python, TensorFlow/PyTorch, and Kubernetes.
Tech Lead Manager, Jockey Core
About * Who we are Video is 90% of the world's data. Most of it is invisible to machines. TwelveLabs builds the intelligence layer to change that. Our multimodal AI models understand video the way humans do —…
Senior Machine Learning Engineer, Jockey Core
Build and optimize the production serving stack for Jockey Core, TwelveLabs’ reasoning LLM that powers agentic video understanding across millions of hours of content.
Développeur Full stack
Build and maintain full-stack apps integrating AI modules (RAG, agents) using React, Node.js, and Python; design scalable APIs, manage data pipelines, and deploy on cloud.
Backend Engineer - API
Build and scale xAI’s high-throughput API infrastructure in Rust/C++ to serve LLM inference globally with low latency and high availability.
Backend Engineer - API
Build and scale SpaceXAI’s high-throughput API in Rust/C++ to serve AI models globally with low latency, handling billions of tokens per minute.
Principal Data Engineer
Principal Data Engineer builds and leads the AI data stack for Anaplan’s LLM and agentic systems, designing retrieval layers, vector/graph databases, and real-time GenAI features for enterprise planning workflows.
AI Engineer, Computer Vision & Video Analytics
Build and test lightweight AI models for video analytics using PyTorch/TensorFlow, then convert them for edge deployment and benchmark performance on target hardware.
Senior Machine Learning Engineer, AI Platform
Build and own the shared AI platform that trains and serves Adobe’s generative AI models at global scale, focusing on GPU fleet utilization, low-latency inference, and end-to-end model deployment pipelines.
AI Computing Software Development Engineer, TensorRT-LLM
Develops and optimizes inference software for large language models using TensorRT-LLM, focusing on performance and scalability across platforms.
Machine Learning Engineer 3 ( Firefly )
Build and optimize scalable generative AI inference pipelines and APIs for Adobe Firefly, integrating models into Photoshop, Illustrator, and other products while focusing on latency and performance.
Lead Engineer AI/ML - Onsite
Lead Engineer builds and deploys production-grade AI/ML models and inference pipelines for retail and operations use cases using Python, PyTorch/TensorFlow, and cloud ML tooling.
Senior Developer Relations Manager – GR00T End-to-End Workflow
Own developer and partner adoption of NVIDIA’s Isaac GR00T end-to-end robotics workflow, guiding teams from model training to real-time deployment on Jetson Thor.
Senior Solutions Architect, Physical AI Cloud
Design and scale Kubernetes-native environments for distributed robotics AI workloads, including simulation, synthetic data generation, and inference using NVIDIA frameworks like OSMO and Isaac Sim.

Computer Vision Engineer
Build the onboard perception system for an all-electric air taxi, fusing camera, lidar, and radar data to detect obstacles and enable precision landings in real time.
Senior Machine Learning Software Engineer
Build and optimize low-level compute kernels and inference pipelines for large language models running on custom ML hardware, integrating with frameworks like PyTorch and vLLM.
Machine Learning Software Engineer, Data Plane
Builds and optimizes low-level compute kernels and serving infrastructure for large language model inference on custom ML hardware, integrating frameworks like PyTorch and vLLM.
Senior Software Development Engineer, Stores Foundational AI
Build and scale foundational large language models for Amazon’s shopping experiences, focusing on ML infrastructure, post-training, and reinforcement learning to improve personalization and customer interactions.
Machine Learning Software Engineer, Data Plane
Build and optimize low-level compute kernels and serving integrations for large language model inference on custom ML hardware, spanning model execution, memory management, and distributed systems.