Tech jobs
Job listings
Embedded AI Engineer Intern [IDA: 00051]
Intern designs and implements computer-vision models for embedded systems, optimizing object detection and AI inference on edge devices like NVIDIA Jetson.
DL Performance Software Engineer - LLM Inference
Build and optimize high-performance LLM inference systems using CUDA, vLLM, and GPU kernels to maximize serving efficiency for large-scale AI models.
Principal Machine Learning Engineer
Leads ML architecture for digital pathology, designing scalable systems in Python/C++ and deploying models on cloud platforms to optimize medical imaging algorithms.
Sr./ Staff AI/ML Engineer - Vision Systems
Lead the development and deployment of AI/ML computer-vision systems for industrial quality-assurance hardware, using Python, OpenCV, TensorFlow/PyTorch, and edge/GPU acceleration.
Senior Machine Learning Engineer
Build and scale a machine-learning platform that accelerates drug discovery, enabling the full ML lifecycle from development through deployment.
[Uni - Jan till Jun 2027] Robotics and Physical AI Engineer Intern
Build and simulate humanoid robots that mimic human movement, integrating AI reasoning with physical motor control using ROS 2, reinforcement learning, and NVIDIA Isaac Sim.
Lead Deep Learning/CUDA Engineer (GigaChat)
Lead a team optimizing CUDA and deep-learning pipelines for large language model inference, ensuring performance, stability, and cost efficiency in production.
Software Engineer Intern, Backend Development
Backend intern at Appier builds and optimizes scalable services for AdTech/MarTech products using Python, Go, or Java, with a hybrid schedule in Taipei.
Member of Technical Staff, ML Inference Engineering
About the Role Sanas is bringing real-time speech and language models on-premise — deployed at scale directly inside sovereign data centers, not served from behind a hosted cloud endpoint. It's one of the most…
Senior Manager, NCP and ISV Business Development
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re…
Computer Vision Engineer (Machine Learning)
Develops and optimizes computer-vision models for defect detection and automated inspection, integrating deep learning with robotics and manufacturing systems.
AI Solutions Architect - Presales
Pre-sales role designing end-to-end AI infrastructure (GPU clusters, networking, storage, MLOps) and translating customer needs into technical proposals to accelerate enterprise AI deals across APAC.
Senior Research Fellow (Quantum-HPC Systems Architect)
Design and lead HPC and distributed AI infrastructure for quantum-classical systems, integrating cloud-native APIs and Slurm GPU plugins while mentoring engineers.
ROS Architect
Design and optimize ROS-based Physical AI systems, tuning Linux kernels and DDS middleware for low-latency robot autonomy and edge compute.
Account Manager - AI Natives
Would you like to be a part of one of the most exciting companies in technology? NVIDIA, the world leader in Visual and Accelerated Computing, is seeking a Sales Account Manager with a proven track record in selling…
Senior Director, NCP and ISV Business Development
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re…
Software Engineer – Linux Infrastructure and Distributed Computing
Build and optimize Linux-based distributed systems and HPC workloads using C++, Kubernetes, and GPU acceleration for semiconductor inspection tools.
AI Research Engineer
Research and optimize deep learning models for edge AI platforms using quantization, model compression, and efficient inference techniques in PyTorch and ONNX Runtime.
Software Engineer, Parallel Scientific Computing
Develop and optimize parallel computational kernels for a custom Scientific Processing Unit (SPU) architecture, translating scientific algorithms into high-performance C++/Python code for HPC simulations and real hardware.