Tech jobs
Job listings
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Optimize and architect ByteDance’s large-model inference engine for GPU/NPU, using CUDA, operator fusion, and distributed parallelism to boost throughput and cut latency.
Software Engineer, AI Infrastructure & Operations, AI Practice
Build and maintain secure, scalable AI infrastructure for Singapore’s government, automating ML pipelines, optimizing LLMs, and enforcing compliance standards.
Backend Engineer - AML Engine Orchestration Singapore Regular
Build and optimize distributed orchestration frameworks for large-scale ML training and inference in Kubernetes, focusing on resource efficiency and next-gen recommendation systems.
Inference Systems Backend Engineer - ARK Large Model Platform (Singapore)
Build and optimize large-scale distributed inference systems for LLMs on VolcanoEngine, tuning GPU clusters and scheduling traffic to handle hundreds of billions of tokens daily.
Senior Backend Engineer - Machine Learning Platform (R&D, CTR/VTR Predictor) - Ego team
Build and optimize low-latency, high-throughput ML inference services for CTR/CVR prediction and generative recommendation using LLMs, focusing on GPU acceleration and end-to-end pipeline optimization.
Senior Backend Engineer - Marketplace Intelligence & Data (Shared Components)
Builds scalable ML infrastructure and distributed systems for Shopee’s marketplace, focusing on vector search, indexing, and agent workflow orchestration.
Senior Backend Engineer - AML Engine Orchestration
Build and optimize distributed orchestration and training systems for large-scale AI models in recommendation and ads, using Kubernetes, Go, and Python.
Inference Systems Backend Engineer - ARK Large Model Platform (Singapore)
Build and optimize large-scale distributed inference systems for LLMs on VolcanoEngine, tuning GPU clusters and traffic scheduling to handle hundreds of billions of tokens daily.
Backend Engineer - AML Framework Development (Search, Ads, and Recommendation Direction)
Build and optimize a high-performance GPU inference engine for large AI models powering ads, search, and recommendation ranking systems using C/C++, Python, and CUDA.
Backend Engineer , AML Engine Orchestration
Build and optimize distributed orchestration systems for large-scale ML training and inference in ByteDance’s AML engine, using Kubernetes, Go, and Python.
Senior Devops Engineer
Build and automate CI/CD pipelines for deploying LLMs and AI agents on Kubernetes in a regulated banking environment using Docker, Terraform, and model serving platforms like vLLM.
Devops Engineer (Strong in Java & Python)
Build and automate CI/CD pipelines for deploying and managing AI models (LLMs, agents) in a secure banking environment using Kubernetes, Docker, and cloud platforms.
DevOps Engineer
Own and evolve Menlo Cloud, a hybrid Kubernetes platform powering robotics R&D, inference serving, and agentic coding—designing networking, CI/CD, observability, and GPU orchestration at scale.
Full Stack Machine Learning Engineer (Datacentre AI Engineering) - Riyadh, KSA
Build and deploy AI inference services, agentic workflows, and LLM runtimes for Qualcomm’s rack-scale data centers, integrating Kubernetes, Prometheus, and Terraform.
Senior DevOps Engineer
Build and run the infrastructure for Generative AI products: pipelines, platforms, and security controls on AWS, including GPU clusters, model serving, and CI/CD.
Platform Engineer (Data Platform, AI, 25-35K)
Build and maintain the AI platform and data lakehouse, automating operations and integrating services with APIs and cloud infrastructure using Python, Kubernetes, and MLOps tools.
AI Engineer (LLM/ Chatbot)
Build and optimize production-grade LLM systems, integrating commercial APIs and self-hosted models, and implementing RAG pipelines and end-to-end LLM workflows.
Insurance AI Architect / AI Engineer / AI Transformation Lead
Lead AI architecture and transformation for insurers, designing multi-agent workflows, document intelligence pipelines, and secure cloud-native AI systems.
AI Engineer (LLM/ Chatbot)
Build and deploy production-grade LLM chatbots and RAG pipelines using commercial APIs and self-hosted open-source models, optimizing for latency, cost, and reliability.
Senior AI/LLM Engineer
Lead training, alignment, and optimization of large language models using RLHF, SFT, and quantization; build reward models, red-team models, and optimize inference pipelines in Python/C++/Rust.