Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Build and optimize GPU kernels and inference frameworks (e.g., vLLM) to accelerate large language model serving, integrating research into production-grade, open-source software.
Build and optimize on-device AI inference software for NVIDIA GPUs, focusing on low-latency, memory efficiency, and deployment on RTX/DGX systems.
Designs and deploys large-scale AI infrastructure for NVIDIA’s Canadian cloud partners, advising on GPU/GPU-networking solutions for AI training and inference pipelines.
Builds AI-powered mapping systems for autonomous vehicles, integrating real-time sensor data with transformer-based models to improve navigation safety and efficiency.
Research engineer translating biometrics and AI research into scalable production software using ML/DL, computer vision, C++, and Python to build biometric recognition systems at Thales.
Build and optimize high-performance simulation systems in C++ to test and validate autonomous trucking software, integrating AI-driven scenario generation and reinforcement learning for scalable virtual environments.
Develop and optimize deep learning models for autonomous truck perception, mapping, and planning using PyTorch, collaborating with cross-functional teams to integrate solutions into production pipelines.
Sr. Manager, Future Computing Prototyping Studio Location: Shanghai, Shanghai, China Time Type: Full time Job Description Sr. Manager, Future Computing Prototyping Studio Description - The Opportunity HP is creating a…
Develops and deploys production AI systems for retail shelf intelligence, focusing on computer vision tasks like object detection, OCR, and product recognition to improve inventory accuracy and retail operations.
Optimize ML inference for edge accelerators and GPUs, focusing on transformer-based models for low-power in-vehicle compute by improving compilers, runtimes, and kernels.
Develops Unity-based immersive and interactive applications for museum exhibits, focusing on real-time 3D rendering, performance optimization, and cross-functional collaboration.
Lead a team building neural rendering and generative models to simulate camera, LiDAR, and radar data for autonomous truck perception training and validation.
Lead a team to design, build, and deploy production-grade ML systems and pipelines using Python, PyTorch, and cloud-native tools to improve United Airlines' operations and customer experiences.
Разработчик C++ создает и оптимизирует алгоритмы для распознавания объектов в высоконагруженных системах, используя данные с камер, лидаров и радаров для автономных автомобилей.
ML engineer builds and optimizes RAG pipelines, fine-tunes VLM for technical docs, and prepares on-prem LLM/VLM inference for an AI platform serving engineers.
Build and run the ML platform that trains, serves, monitors, and retires models in batch and real-time on CPU/GPU, using Kubernetes, Triton, MLflow, Prometheus and CI/CD.
Designs end-to-end AI/deep-learning architectures for geospatial platforms, customizing models like LLM/RAG and deploying on GPU/cloud infrastructure for defense, government, and commercial clients.
About the role: We’re looking for an experienced GPU Driver Developer. You will focus the design, development, implementation, and toolchain of a high-performance host-side driver stack and API for our proprietary…
Build and scale the infrastructure that powers a computer-vision engine detecting underground utilities, designing pipelines, data systems, and deployment workflows for AI models.
Build open-source AI solution blueprints and reference implementations for enterprise customers, focusing on production-grade architectures, MLOps, and secure deployments in regulated environments.
We couldn't check your fit for this role — add a CV to your profile to see it next time.