Tech jobs
Job listings
Machine Learning Engineer
Build and deploy large-scale machine learning models and AI agents to optimize Micron’s semiconductor manufacturing workflows using distributed training and GPU optimization techniques.
HPC Customer Solutions Engineer
Support and optimize high-performance computing (HPC) systems for scientific and AI workloads, troubleshooting applications and collaborating with engineering to improve platform usability.
Senior Software Engineer
Build and integrate AI/ML features into enterprise Kubernetes environments using Python/Golang, vLLM, PyTorch, and OpenShift, while contributing to open source projects.
Tools Development Engineer, Machine Learning
Build AI-driven tools to automate gameplay testing and validate NVIDIA GPUs using deep learning, reinforcement learning, and generative models.
Deep Learning Software Engineer, TensorRT Performance (Remote)
Build and optimize NVIDIA’s deep-learning inference stack (TensorRT, TensorRT-LLM) by profiling models, writing GPU kernels, and integrating OSS frameworks to maximize GenAI performance across datacenter and edge GPUs.
Senior Deep Learning Frameworks CUDA Software Engineer (Remote)
Build and optimize CUDA features for AI frameworks like PyTorch and TRT-LLM, improving multi-GPU performance and distributed runtime for training and inference workloads.
Senior Deep Learning Algorithm Engineer (Remote)
Senior engineer building and optimizing NVIDIA’s Dynamo inference platform, integrating open-source frameworks like vLLM and TensorRT-LLM to maximize AI throughput and latency.
Deep Learning Performance Software Engineer, LLM Inference
Build and optimize GPU-powered inference systems for large language models, improving serving efficiency and performance through kernel tuning and novel algorithms.
Senior Infrastructure Software Engineer (Remote)
Build and maintain scalable CI/CD infrastructure for NVIDIA’s TensorRT Edge-LLM, automating builds, tests, and deployments across embedded and cloud platforms using tools like GitLab, Kubernetes, and Docker.
Senior Associate, Generative AI, Data and Analytics, Advisory
Design and optimize ML pipelines, LLM serving, and GPU architectures for generative AI solutions using frameworks like PyTorch, Hugging Face, and vLLM.
Sovereign Engineering Platform SRE - T Cloud Public (REF5740Q)
Build and run a secure Kubernetes-based platform that hosts AI engineering tools, GitOps workflows, and observability stacks for AI-assisted software development in sovereignty-sensitive environments.
AI Engineer - Tieto Banktech (m/f/d)
Build and deploy AI features for banking platforms, leading from prototype to production while ensuring regulatory compliance and cross-team adoption.
Senior DevOps/Cloud Platform Engineer (AWS| Kubernetes|AI Infrastructure)
Build and maintain secure, scalable AWS and Kubernetes platforms for AI workloads, including LLM inference, while automating CI/CD and ensuring SOC 2/HITRUST compliance.
AI Engineer
Build and integrate a GenAI platform using large language models, RAG pipelines, and self-hosted frontier models for a Polish tech firm.
AI Engineer
Build autonomous AI agents that plan, collaborate, and act using frameworks like LangChain and DSPy, integrating LLMs with tools and APIs for real-world execution.
Forward Deployed Engineer
Forward Deployed Engineer at VESSL AI: a hands-on role bridging customers and GPU cloud tech, designing PoCs, solving AI workload issues, and driving adoption of VESSL’s GPUaaS platform for training and inference.
Senior DevOps LLM (Инвестиционный бизнес)
Senior DevOps engineer deploying and optimizing open-source LLMs (DeepSeek, Qwen) on-premises, building high-throughput inference servers and observability for an investment platform.
Staff AI Engineer, Model Post-Training and Alignment
Lead post-training and alignment of large language models using DPO, GRPO, and RLAIF; design reward models and optimize inference with vLLM/SGLang for a leading crypto exchange.
Software Development Engineer
A junior Python engineer builds AI-powered document processing tools (PDFs) in a Dublin-based team, focusing on LLM-driven workflows, embeddings, and semantic search while learning from senior developers.
Data Scientist в команду LLM Core
Build and improve Avito’s in-house large language models by designing experiments, training pipelines, and optimizing inference for production use.