Tech jobs
Job listings
Machine Learning Platform Engineer
Build and operate the ML infrastructure powering a global AI assistant, including training, deployment, inference, and observability systems in Python and PyTorch/JAX.
Principal Machine Learning Engineer
About A1 There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to…
Software Engineer - ML Infrastructure
Build and maintain the machine-learning infrastructure that trains and deploys AI models for medical imaging, including distributed training, data pipelines, and production serving.
Senior AI/Computer Vision Engineer
Build and deploy AI/computer-vision systems end-to-end, from data pipelines and model training to C++ edge integration and production validation.
Principal Product Manager, Augmented Memory Grid (AMG)
Own the roadmap for WEKA’s Augmented Memory Grid, optimizing LLM inference performance by offloading KV-cache and integrating with engines like vLLM and NVIDIA Triton.
Senior Data Engineer - AI & MLOps
Build and optimize AI/ML and data pipelines for a sensor-fusion system, focusing on real-time processing, CI/CD, and MLOps to support an autonomous team in Sydney.
[인턴] [NetsPresso] R&D AI Engineer (Quantization팀)
Research and implement model quantization algorithms to optimize AI models for on-device deployment using PyTorch, ONNX, and related tools.
[인턴] [NetsPresso] R&D AI Engineer (Model Representation팀)
R&D intern on Nota’s AI team optimizing models for deployment on edge devices using PyTorch, ONNX, and frameworks like ExecuTorch/TensorRT.
[인턴] [NetsPresso] R&D AI Research Engineer (XPU Enabler팀)
Research and develop quantization, pruning, and inference optimizations for LLM/VLM and MoE models to run efficiently on GPUs and NPUs.
AI Software Developer
Develops and maintains AI models and data pipelines for mobile game publishing, integrating LLMs and transformer-based systems into scalable game and embedded solutions.
Senior Manager, Sales Engineering — AI / GPU Cloud (NeoCloud)
Why this role exists K0rdent AI is the orchestration layer that turns raw, disaggregated GPU infrastructure into a multi-tenant, production-ready AI cloud — without locking companies into a single hyperscaler or…
LLM Research Intern
Research intern experimenting with open-weight LLMs, fine-tuning, RAG, and evaluation to adapt models for enterprise use cases alongside ProCogia’s AI teams.
Software Engineer II — Agentic AI Foundations
Build a secure, vendor-agnostic agent platform for identity verification workflows, integrating LLMs, orchestration, and safety controls in production.
ML Infrastructure Engineer
Build and optimize GPU infrastructure for AI workloads, profiling performance across hardware and frameworks to guide platform decisions and hardware development.
ML Platform Engineer
ML Platform Engineer builds and scales AI model pipelines, deploys ML services on Kubernetes, and maintains FastAPI-based inference APIs for a large insurer’s internal AI ecosystem.
Machine Learning Engineer*
Build and deploy end-to-end ML systems for industrial sorting machines, focusing on computer vision and MLOps across cloud and edge.
Data Scientist / ML Engineer (Computer Vision), Junior+ / Middle
Build and ship computer-vision models: train, evaluate, wrap in services, and demo via simple web UIs. Core stack: Python, PyTorch, OpenCV, FastAPI/Flask, Docker.
Member of Technical Staff — Model Optimization and Inference (New Grad)
Optimize and deploy real-time AI avatars by accelerating LLM, audio, and diffusion model inference to sub-500ms latency using quantization, KV cache tuning, and custom kernels.
Member of Technical Staff — Model Optimization and Inference (Experienced)
Optimize and deploy real-time AI avatar models for sub-500ms latency, focusing on KV cache, quantization, and kernel-level acceleration across LLMs, audio, and diffusion components.