Tech jobs
Job listings
System Software Engineer, LLM Inference
Build and optimize open-source LLM inference engines for CXL-based memory offloading, integrating cryptographic acceleration and contributing upstream to AI infrastructure projects.
Senior Machine Learning Engineer, Physical AI
Owns the full ML lifecycle for physical AI, from sensor data pipelines to deploying optimized models on constrained hardware, collaborating with embedded teams to ensure reliability and performance in real-world devices.
Senior Software Engineer - Low-Level Systems, Compilers & Embedded AI
Build low-level compilers, runtimes, and AI/DSP optimizations for next-gen hardware in C++/Python, working closely with hardware teams.
Principal Machine Learning Engineer
About A1 There are over 5 billion users using basic applications today such email, notes, tasks that are not AI-native. Our mission is to build a proactive smart assistant for everyday users to bring intelligence to…
AI Engineer
Looking for an AI Engineer to Develop, deploy, and operate AI/LLM models across Clinets dual environment — GCP for public-cloud workloads, Humain sovereign cloud for classified data. Requirements Build and fine-tune…
Software Engineer - ML Infrastructure
Build and maintain the machine-learning infrastructure that trains and deploys AI models for medical imaging, including distributed training, data pipelines, and production serving.
Forward Deployed Engineer
Forward Deployed Engineer at DeepInfra designs and runs AI inference benchmarks, tunes deployments on cutting-edge hardware, and partners with sales to win enterprise deals.
Senior Principal FPGA & Digital Signals Processing Engineer
Lead FPGA firmware architecture for a portable SIGINT/EW platform, migrating GPU DSP algorithms to FPGA and optimizing real-time RF signal processing for defense missions.
Principal Product Manager, Augmented Memory Grid (AMG)
Own the roadmap for WEKA’s Augmented Memory Grid, optimizing LLM inference performance by offloading KV-cache and integrating with engines like vLLM and NVIDIA Triton.
Silicon Architect, AI Accelerator
Designs AI accelerator silicon and compute subsystems, defining ISA, connectivity, and memory while optimizing for performance, power, and area in GenAI hardware.
Software Engineer III - Data Analytics Platform
Builds and optimizes backend services for large language model inference at a global bank, focusing on performance, reliability, and scalability within a distributed systems environment.
[인턴] [NetsPresso] R&D AI Engineer (Quantization팀)
Research and implement model quantization algorithms to optimize AI models for on-device deployment using PyTorch, ONNX, and related tools.
[인턴] [NetsPresso] R&D AI Research Engineer (XPU Enabler팀)
Research and develop quantization, pruning, and inference optimizations for LLM/VLM and MoE models to run efficiently on GPUs and NPUs.
Embedded Software Engineers 3 years + C, C++ Development and Machine Learning
Develops embedded C/C++ firmware for sensor-based AI systems, integrating ML models (PyTorch/TensorFlow) with DSP and quantization for consumer electronics and automotive.
LLM Engineer (Data and Optimization)
Build, optimize, and deploy large language and multimodal models for industrial use, focusing on training, compression, RAG, and agent workflows.
LLM Engineer for AI Software Development Tools
Build and deploy large language models for code generation and reasoning using PyTorch/TensorFlow, distributed training, and GPU clusters.
LLM Research Intern
Research intern experimenting with open-weight LLMs, fine-tuning, RAG, and evaluation to adapt models for enterprise use cases alongside ProCogia’s AI teams.
Senior Machine Learning Engineer
Build and deploy AI models for public-safety products like cloud and device-based systems, focusing on computer vision, speech recognition, and NLP.
Lead AI Engineer
Lead AI Engineer builds agentic AI systems, RAG pipelines, and fine-tuned language models while ensuring security and compliance for a banking-focused AI product.