Tech jobs
Job listings
Embedded Software Engineer, Senior
Designs and deploys secure, rugged embedded systems integrating CPUs/GPUs/TPUs/FPGAs for edge computing, optimizing AI models for real-time inference and hardening against cyber/physical threats in mission-critical environments.
Software Engineer Senior-Ai Engineer
Build and deploy production-grade LLM systems for a large bank, focusing on fine-tuning, optimization, RAG pipelines, and agentic workflows integrated with enterprise data.
Staff Machine Learning Engineer, Generative AI Modeling and Inference
Build and optimize generative AI models and inference pipelines for Snapchat’s AR and creative tools, delivering scalable, on-device experiences for millions of users.
AI Engineer
Build and deploy AI models for edge devices: port vision/speech models to NPU-enabled SoCs, extend AI tooling for hardware-aware development, and create reference designs for smart-home and surveillance use cases.
AI/ML Engineer(LLM & Agentic AI)
Build, fine-tune, and deploy LLM-powered AI agents using Vertex AI, PaLM, and Gemini, focusing on prompt engineering, RAG, RLHF, and agent orchestration.
Staff Machine Learning Engineer, Generative AI Modeling and Inference
Build and optimize generative AI models for Snapchat, including LLMs, video generation, and AR, focusing on efficient inference and on-device deployment.
Software Engineer Senior-Ai Engineer
Build and deploy production-grade LLM systems for a large bank, focusing on fine-tuning, RAG pipelines, and agentic workflows integrated with enterprise data.
Staff Machine Learning Engineer - LLM Quantization & Deployment
Build and deploy quantized large language models for in-vehicle AI, focusing on PTQ, QAT, and low-bit inference to ensure numerical consistency and performance on XPENG’s Turing AI chip.
Senior Machine Learning Engineer - LLM Quantization & Deployment
Build and deploy quantized large language models for in-vehicle AI, focusing on PTQ, QAT, and low-bit inference to optimize performance on XPENG’s Turing AI chip.
Sr. Applied Intelligence Architect
Architect and deploy AI systems for semiconductor manufacturing, focusing on model selection, governance, and integration across engineering workflows and enterprise platforms.
Machine Learning Engineer
About Sunset At its core, Sunset was founded to help founders. We started by supporting startups through shutting down, but we have since expanded into unlocking a new revenue stream for all types of businesses. In…
Senior Software Engineer - GPU Kernel Authoring & Optimization
Write, profile, and optimize CUDA kernels for LLM inference to maximize throughput and minimize latency on NVIDIA GPUs, using DSLs like Triton or Mojo and benchmarking with MLPerf.
Applied AI Engineer, Inference
Applied AI Engineer, Inference at CoreWeave. Improve real-world performance of AI models via benchmarking, profiling, and optimization.
Forward Deployed Engineer (Inference & Post-Training) - Mandarin Speaking
Optimize and deploy large language models for enterprise customers, tuning inference engines and post-training pipelines to meet performance targets.
Senior Platform Engineer
Build and run a modern AI-powered platform on GCP, owning cloud-native infrastructure, CI/CD, security, observability, and production operations for scalable, reliable services.
Senior Data Scientist
Build and deploy ML/AI models end-to-end, from data exploration to production, using Python, PyTorch/TensorFlow, and cloud platforms like AWS/GCP/Azure.
Senior ML Engineer
Build, fine-tune, and deploy deep-learning models end-to-end: train Transformers, optimise inference with NVIDIA’s stack, and package models into production-ready containers.
Senior ML Engineer (Token Factory)
Build and optimize large-scale AI training and inference systems using Python and JAX, focusing on fine-tuning, distributed training, and low-precision computation for foundation models.