Tech jobs
Job listings
Principal Machine Learning Engineer
About A1 There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to…
Software Engineer - ML Infrastructure
Build and maintain the machine-learning infrastructure that trains and deploys AI models for medical imaging, including distributed training, data pipelines, and production serving.
Senior Site Reliability Engineer - Fleet
Build and operate AI cloud infrastructure, automating large-scale HPC clusters with Ansible/Terraform and monitoring GPU/fabric health to ensure high availability for AI workloads.
[인턴] [NetsPresso] R&D AI Research Engineer (XPU Enabler팀)
Research and develop quantization, pruning, and inference optimizations for LLM/VLM and MoE models to run efficiently on GPUs and NPUs.
Developer Relations Engineer
Build and deploy AI models on Vast.ai’s GPU marketplace, create public guides and demos, and engage with developers to solve real workloads and share solutions.
LLM Engineer (Data and Optimization)
Build, optimize, and deploy large language and multimodal models for industrial use, focusing on training, compression, RAG, and agent workflows.
LLM Research Intern
Research intern experimenting with open-weight LLMs, fine-tuning, RAG, and evaluation to adapt models for enterprise use cases alongside ProCogia’s AI teams.
Helix AI Engineer, Training Performance
Optimize distributed AI training for 100B+ parameter models across 100k+ GPUs, writing CUDA/Triton kernels and co-designing hardware-efficient training recipes.
Member of Technical Staff — Pretraining Infra (Experienced)
Build and scale distributed training infrastructure for large-scale AI avatar models, optimizing GPU clusters, parallelism, and multimodal data pipelines for real-time, full-duplex training.
LLM Pre-training & Distributed Engineer (AI Infrastructure)
Engineer large-scale LLM pre-training pipelines and distributed GPU clusters using PyTorch, DeepSpeed, or Megatron-LM, optimizing networking, memory, and fault tolerance for month-long runs.
AI SoC Architect Networking
Designs high-speed SoC networking engines and interconnect fabrics for custom AI servers, optimizing distributed AI workloads with RDMA, RoCEv2, and collective-operation acceleration.
AI SoC Architect Networking
Architect high-speed SoC networking engines and interconnect fabrics for custom AI servers, optimizing distributed AI workloads with RDMA, RoCEv2, and collective-operation acceleration.
Research Scientist
About Company Founded in 2022, XG Tech is driving the future of smart vehicles. Its mission is to empower the digital transformation of automobiles, moving from distributed computing to a centralized, cross-domain…
Staff Applied AI/ML Scientist
Build and fine-tune large language models to automate healthcare claims and conversational AI, using Python, PyTorch, and cloud ML platforms.
Senior Applied AI/ML Scientist
Build and fine-tune large language models for healthcare, creating conversational AI that automates claims and improves patient experiences using Python, PyTorch, and cloud ML platforms.
Staff Machine Learning Engineer, Foundation - Seattle
Build and scale AI/ML systems to personalize Qualtrics' SaaS platform, leading architecture, deployment, and optimization of large-scale machine learning models and infrastructure.