Tech jobs
Job listings
[LTA-TRO] SENIOR/ EXECUTIVE/ AI SOFTWARE DEVELOPER - TRANSPORT AI PROGRAMME
Build and deploy AI-powered transport applications using LLMs, computer vision, and agentic systems to optimize Singapore’s land transport network.
[LTA-TRO] SENIOR/ EXECUTIVE/ AI SOFTWARE DEVELOPER - TRANSPORT AI PROGRAMME
Build and deploy AI-powered transport applications for Singapore’s land transport network, integrating LLMs, computer vision, and agentic systems into production workflows with safety and reliability guardrails.
AI Scientist - LLM & AI Agent
Build and deploy modular AI agents using LangGraph, MCP, and RAG for fintech workflows, focusing on reasoning, memory, and multi-agent orchestration.
AI Engineer
Build and deploy scalable AI systems that turn physical-world data into enterprise intelligence, optimizing energy and operations for industries like climate tech.
AI Engineer (AI Products)
Build and fine-tune multilingual LLMs for underrepresented languages, engineer scalable data pipelines, and deploy AI-powered products at a national AI program.
Lead Data and AI Solution Engineer
Lead the design and deployment of GenAI solutions for insurance, including RAG pipelines, vector databases, and LLM fine-tuning on AWS to deliver context-aware AI capabilities.
Software Engineer (AI Systems)
Build and deploy AI-powered features for a fintech company, integrating LLMs, vector databases, and agent workflows into production systems using Python, Node.js, and cloud-native tools.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Develop and optimize a high-performance GPU inference engine for large language models, focusing on operator fusion, compilation, and distributed parallelism to reduce latency and boost throughput.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Build and optimize ByteDance’s large-model inference engine, focusing on GPU performance, operator fusion, compilation optimizations, and distributed parallelism to reduce latency and boost throughput.
Backend Inference Runtime Engineer Graduate (AML Inference) - 2027 Start
Optimize and architect ByteDance’s large-model inference engine for GPU/NPU, using CUDA, operator fusion, and distributed parallelism to boost throughput and cut latency.
Software Engineer, AI Infrastructure & Operations, AI Practice
Build and maintain secure, scalable AI infrastructure for Singapore’s government, automating ML pipelines, optimizing LLMs, and enforcing compliance standards.
Backend Engineer - AML Engine Orchestration Singapore Regular
Build and optimize distributed orchestration frameworks for large-scale ML training and inference in Kubernetes, focusing on resource efficiency and next-gen recommendation systems.
Inference Systems Backend Engineer - ARK Large Model Platform (Singapore)
Build and optimize large-scale distributed inference systems for LLMs on VolcanoEngine, tuning GPU clusters and scheduling traffic to handle hundreds of billions of tokens daily.
Senior Backend Engineer - Machine Learning Platform (R&D, CTR/VTR Predictor) - Ego team
Build and optimize low-latency, high-throughput ML inference services for CTR/CVR prediction and generative recommendation using LLMs, focusing on GPU acceleration and end-to-end pipeline optimization.
Senior Backend Engineer - Marketplace Intelligence & Data (Shared Components)
Builds scalable ML infrastructure and distributed systems for Shopee’s marketplace, focusing on vector search, indexing, and agent workflow orchestration.
Senior Backend Engineer - AML Engine Orchestration
Build and optimize distributed orchestration and training systems for large-scale AI models in recommendation and ads, using Kubernetes, Go, and Python.
Inference Systems Backend Engineer - ARK Large Model Platform (Singapore)
Build and optimize large-scale distributed inference systems for LLMs on VolcanoEngine, tuning GPU clusters and traffic scheduling to handle hundreds of billions of tokens daily.
Backend Engineer - AML Framework Development (Search, Ads, and Recommendation Direction)
Build and optimize a high-performance GPU inference engine for large AI models powering ads, search, and recommendation ranking systems using C/C++, Python, and CUDA.
Backend Engineer , AML Engine Orchestration
Build and optimize distributed orchestration systems for large-scale ML training and inference in ByteDance’s AML engine, using Kubernetes, Go, and Python.
Senior Devops Engineer
Build and automate CI/CD pipelines for deploying LLMs and AI agents on Kubernetes in a regulated banking environment using Docker, Terraform, and model serving platforms like vLLM.