Tech jobs
Job listings
Software Engineer II — Agentic AI Foundations
Build a secure, vendor-agnostic agent platform for identity verification workflows, integrating LLMs, orchestration, and safety controls in production.
Senior Cloud Engineer (Toronto, ON, Canada)
Design and migrate AWS cloud infrastructure for scalable Django apps, lead Kubernetes clusters, and advise clients on secure, cost-efficient cloud solutions.
MLOps
Build and run a production-grade ML platform: design AI system architectures, deploy and optimize LLM inference servers, and maintain MLOps pipelines with GPU scheduling and monitoring.
ML Platform Engineer
ML Platform Engineer builds and scales AI model pipelines, deploys ML services on Kubernetes, and maintains FastAPI-based inference APIs for a large insurer’s internal AI ecosystem.
Senior AI Engineer
Build, deploy, and optimize production AI systems including LLMs and generative AI models, focusing on inference pipelines, performance, and cost efficiency.
Staff AI Engineer
Lead architecture and deployment of production AI systems, optimizing inference pipelines and mentoring engineers in a fast-growing company.
Senior Системный аналитик (Гиперперсонализация)
Designs and documents high-load lending services for Sberbank, integrating AI agents and LLM prompts to automate dynamic pricing and credit decisions.
Middle Системный аналитик (Гиперперсонализация)
System analyst at Sberbank designing and improving online lending services, integrating AI agents and LLM APIs, and modeling high-load distributed systems for millions of users.
Senior Principal Backend Development Engineer
Build and own the agentic AI platform for a crypto exchange, designing orchestration engines, MCP servers, and lifecycle toolchains to power conversational agents handling trading, compliance, and customer service.
Helix AI Engineer, Training Performance
Optimize distributed AI training for 100B+ parameter models across 100k+ GPUs, writing CUDA/Triton kernels and co-designing hardware-efficient training recipes.
AI Platform Engineer
Build and scale an AI model service platform from scratch, designing APIs, routing, and load balancing for high-volume inference across connected hardware.
Member of Technical Staff — Model Optimization and Inference (New Grad)
Optimize and deploy real-time AI avatars by accelerating LLM, audio, and diffusion model inference to sub-500ms latency using quantization, KV cache tuning, and custom kernels.
Member of Technical Staff — Model Optimization and Inference (Experienced)
Optimize and deploy real-time AI avatar models for sub-500ms latency, focusing on KV cache, quantization, and kernel-level acceleration across LLMs, audio, and diffusion components.
Principal Architect, Simulation Platform (Remote)
Lead the architecture of a GPU-powered simulation platform for AI data centers, enabling real-time power orchestration and validation of control systems across heterogeneous time domains.
Senior Machine Learning Research Engineer
Build and scale generative and predictive ML models for cellular behavior using PyTorch and distributed training, bridging research prototypes to production-grade systems in a TechBio company.
Research Scientist
About Company Founded in 2022, XG Tech is driving the future of smart vehicles. Its mission is to empower the digital transformation of automobiles, moving from distributed computing to a centralized, cross-domain…
AI Platform Engineer
Design and operate scalable AI inference platforms for production ML workloads, optimizing GPU utilization, LLM serving, and cloud-native infrastructure.