Tech jobs
Job listings
(Staff/ Sr. Staff) Machine Learning Engineer
Lead research and development of ML algorithms and toolchains for Edge AI, focusing on quantization, compression, and on-device learning to optimize models for constrained hardware.
AI Infrastructure Engineer
Optimize LLM inference performance on Intel GPUs by profiling bottlenecks, writing custom kernels, and contributing to open-source serving frameworks like vLLM and SGLang.
AI Platform Engineer (Google Cloud Platform)
Build and maintain an AI platform on Google Cloud for an insurance company, deploying GenAI/ML services, vector search, and agentic frameworks to power enterprise teams like Actuarial and Underwriting.
Lead Software Engineer - LLM Ops Platform Reliability
Build and operate scalable LLM serving infrastructure using cloud and Kubernetes, ensuring reliability and performance for AI systems in production.
AI Engineer (Managed Services)
Build and deploy enterprise LLM applications, RAG systems, and AI agents using open-source models (DeepSeek, Qwen, Kimi) and frameworks like LangChain and vLLM.
AI / LLM Inference Engineer
Engineer AI/LLM inference on GPU clusters: benchmark, tune, and optimize model serving with vLLM, Triton, or TensorRT-LLM to hit latency, throughput, and memory targets.
Computer Vision Engineer / Machine Learning Engineer
Research and develop on-device generative and 3D spatial AI models for mobile platforms, optimizing models for edge deployment and shipping to product.

Senior Forward Deployed Engineer I (AI Infra)
Senior engineer in Bengaluru who partners with AI-native companies to deploy, optimize, and scale production AI systems on DigitalOcean’s AI-Native Cloud, focusing on inference, agentic workloads, and platform tooling.
GPU Software Engineer (Graphics / ML)
Build and optimize real-time graphics and ML inference pipelines for interactive visual apps, integrating super-resolution and denoising models into DX12/Vulkan renderers.
Applied Data Scientist
Build and deploy realtime intent-drift detection models for AI coding agents running on developer machines, using small open models and local-first Python tooling.
Member of Technical Staff, Accelerator Systems
Optimize and port AI inference/training kernels (SGLang, Miles) across NVIDIA/AMD GPUs, TPUs, CPUs, and emerging accelerators to maximize performance on heterogeneous hardware.
System Software Engineer - Local AI
Build and optimize on-device AI inference software for NVIDIA GPUs, focusing on low-latency, memory efficiency, and deployment on RTX/DGX systems.
Machine Learning Engineer
Build and optimize ML models to enhance Altera’s FPGA compiler, focusing on timing closure, resource utilization, and power efficiency using PyTorch/TensorFlow.
Staff AI Software Engineer
Design and deploy cutting-edge AI/ML applications, including LLMs and computer vision, in a defense-focused R&D environment requiring Secret clearance.
Algorithm Engineer, Deep Learning & Vision (New Grad)
Develop and optimize deep learning models for autonomous truck perception, mapping, and planning using PyTorch, collaborating with cross-functional teams to integrate solutions into production pipelines.
Senior Security Machine Learning / AI Engineer
Build and fine-tune AI models for security alert triage and risk scoring using enterprise telemetry, then deploy a multi-model routing layer that keeps costs predictable while improving accuracy over time.
