Tech jobs
Job listings
AI / LLM Inference Engineer
Engineer AI/LLM inference on GPU clusters: benchmark, tune, and optimize model serving with vLLM, Triton, or TensorRT-LLM to hit latency, throughput, and memory targets.
Software Engineer (Computer Vision)
Builds and optimizes real-time computer-vision pipelines using Python, C/C++, OpenCV, and GStreamer for defense-tech systems.
Lead Software Engineer (Signal Processing)
Lead a team building real-time radar and electronic-warfare signal-processing software in C++/CUDA, architecting DSP pipelines for deployable systems and validating performance via simulation and field tests.
Computer Vision Engineer / Machine Learning Engineer
Research and develop on-device generative and 3D spatial AI models for mobile platforms, optimizing models for edge deployment and shipping to product.
Domain Architect - AI Compute
Salary: $150 – $170 per hour About the Role The Domain Architect - AI Compute acts as the primary technical authority for the physical and logical lifecycle of high-performance GPU compute fleets across diverse client…
Sr MLOps Engineer
Build and maintain the Kubernetes-based infrastructure and MLOps tooling that trains, validates, and deploys AI models for Intuitive Surgical’s medical devices, ensuring reproducible workflows and GPU cluster health.
Machine Learning Engineer - Spiking Neural Networks
Develop spiking neural networks and machine-vision software in C++ with CUDA/GPU acceleration, turning research into production-grade systems.
C++ Developer
Build high-performance C++ applications for real-time data processing, distributed systems, and backend integration in a security and communications-focused product company.
Solution Architect (ГигаЧат Бизнес)
Solution Architect designs enterprise-scale GenAI/LLM systems, including RAG pipelines, agent orchestration, and secure on-prem deployments, while guiding clients through AI transformation and regulatory compliance.
Software Engineer, GPU Cluster Infrastructure
Build and operate a unified GPU compute platform for AI training and inference, hiding cloud complexity behind Kubernetes operators and schedulers that manage multi-cloud NVIDIA clusters.

Senior Forward Deployed Engineer I (AI Infra)
Senior engineer in Bengaluru who partners with AI-native companies to deploy, optimize, and scale production AI systems on DigitalOcean’s AI-Native Cloud, focusing on inference, agentic workloads, and platform tooling.
GPU Software Engineer (Graphics / ML)
Build and optimize real-time graphics and ML inference pipelines for interactive visual apps, integrating super-resolution and denoising models into DX12/Vulkan renderers.
FPGA Software Entwickler für Video / Bildverarbeitung (m/w/d)
Develop FPGA-based video/image processing software using VHDL/Verilog, C/C++, Python, and tools like Vivado/Vitis for embedded systems.
Member of Technical Staff, Accelerator Systems
Optimize and port AI inference/training kernels (SGLang, Miles) across NVIDIA/AMD GPUs, TPUs, CPUs, and emerging accelerators to maximize performance on heterogeneous hardware.
Embedded Software Engineer
Develop embedded Linux software for small satellites, including ML and autonomy experiments, interfacing with custom hardware via protocols like I2C and UART.
CFD Software Engineer
Develop GPU-accelerated CFD simulation software with AI-assisted methods, designing solvers and APIs for modern engineering workflows.

Member of Technical Staff - Low Level & Kernels Capabilities
Build low-level reinforcement learning environments targeting hardware features like GPUs, FPGAs, and vector ISAs to stress-test AI models, using C/C++/CUDA and Python.
AI Engineer (AMK)
Build and maintain data pipelines for AI systems, integrate ML components, and optimize inference workflows in Python.
DL Performance Software Engineer - LLM Inference
Build and optimize GPU kernels and inference frameworks (e.g., vLLM) to accelerate large language model serving, integrating research into production-grade, open-source software.