Tech jobs
Job listings
AI Engineer
Build and deploy production-grade AI systems for edge environments like offshore platforms, mines, and autonomous vehicles, optimizing models for latency, robustness, and security across disconnected settings.
AI Engineer (AMK)
Build and maintain AI data pipelines, ETL workflows, and integrate ML components into production systems, focusing on model inference and RAG optimization.
AI Engineer
Build and deploy AI models: clean datasets, train neural networks, convert models to APIs, and integrate them into apps and cloud systems.
Engineering Manager, ML Performance
Like Google's own ambitions, the work of a Software Engineer goes beyond just Search. Software Engineering Managers have not only the technical expertise to take on and provide technical leadership to major…
Principal Compiler Software Development Engineer - AI/ML
Principal engineer optimizing AI/ML compilers and runtime for AMD GPUs, focusing on MLIR/LLVM transformations and ONNX operators to accelerate training and inference workloads.
Senior AI Engineer
Senior AI Engineer builds and deploys PyTorch-based vision and multimodal models to power photo selection and editing for professional photographers, optimizing for cloud and edge deployment.
Python Developer
Develop and maintain scalable Python solutions for financial services clients, focusing on enterprise-grade architecture, performance optimization, and DevOps practices.
Principal Engineer - Backend
Principal Engineer designs and leads backend infrastructure for Encord’s AI data platform, building scalable distributed systems and APIs that power model training and evaluation at scale.
Senior Deep Learning Engineer
Senior Deep Learning Engineer optimizes neural networks to run efficiently on custom AI hardware, builds model optimization pipelines, and collaborates with hardware teams to deploy AI inference at scale.
AI Engineer — Vision AI & Edge Intelligence
Build and deploy real-time Vision AI and edge-inference systems using PyTorch/TensorFlow, ONNX/TensorRT, and NVIDIA stacks for object detection, video analytics, and multimodal models.
Machine Learning Engineer, Integrations
Build and integrate computer-vision models into client stacks using PyTorch, OpenCV, and Gradio; prototype cutting-edge algorithms and ship them in Datature’s MLOps platform.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Develop and optimize a high-performance GPU inference engine for large language models, focusing on operator fusion, compilation, and distributed parallelism to reduce latency and boost throughput.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Build and optimize ByteDance’s large-model inference engine, focusing on GPU performance, operator fusion, compilation optimizations, and distributed parallelism to reduce latency and boost throughput.
Backend Software Engineer - Recommendation Content Understanding Architecture
Build and optimize TikTok’s multi-modal content understanding pipeline, leveraging vector retrieval and large-scale ML models to improve recommendation relevance and system performance.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Optimize and architect ByteDance’s large-model inference engine for GPU/NPU, using CUDA, operator fusion, and distributed parallelism to boost throughput and cut latency.
Graduate Backend Inference Engineer — GPU & Optimization
Optimizes GPU/TPU-based inference engines for large models, focusing on performance, memory, and low-latency pipelines using C/C++, Python, and CUDA.
Backend Engineer - Model Training Infra (Singapore) Technology - Backend Singapore Regular
Build and scale the ML infrastructure powering ByteDance’s global ad, search, and e-commerce ranking systems using C/C++, CUDA, Python, and frameworks like TensorFlow/PyTorch.
Backend Engineer - Multi-Modal Content & Recommendations
Design and improve TikTok’s multi-modal content understanding pipeline for recommendations, ensuring high availability and scalable vector processing.
Backend Engineer Intern (ByteRec Recommendation Infrastructure) - 2026 Start (BS/MS)
Backend engineering intern building scalable recommendation infrastructure using multi-modal content processing, vector retrieval, and RAG systems in Python/C++.
Inference Systems Backend Engineer - ARK Large Model Platform (Singapore)
Build and optimize large-scale distributed inference systems for LLMs on VolcanoEngine, tuning GPU clusters and scheduling traffic to handle hundreds of billions of tokens daily.