Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Optimize inference for local LLMs and internal models using vLLM/Triton, manage GPU resources, and monitor high-load AI infrastructure for a logistics-focused AI company.
Build and optimize large-scale generative language models (LLMs) for GigaChat, focusing on Russian-language performance, distributed training, and efficiency improvements.
Build and scale the cloud infrastructure for an AI startup, setting up core systems and tooling from the ground up.
Build and own cloud infrastructure for an AI-driven materials discovery platform, focusing on GPU compute, CI/CD, and reproducibility to accelerate scientific breakthroughs.
Build and test lightweight AI models for video analytics using PyTorch/TensorFlow, then convert them for edge deployment and benchmark performance on target hardware.
Lead AI research to build high-fidelity, multi-physics simulation models for engineering and manufacturing, using deep learning and probabilistic methods to optimize design and operations.
Build and optimize AI-driven simulation models for engineering and manufacturing, scaling deep learning and distributed training across cloud and on-premise systems.
Build and optimize ML frameworks, runtimes, and compute pipelines in C++, Rust, Python, and CUDA to accelerate research and production workloads.
Develops and optimizes inference software for large language models using TensorRT-LLM, focusing on performance and scalability across platforms.
Build and optimize scalable generative AI inference pipelines and APIs for Adobe Firefly, integrating models into Photoshop, Illustrator, and other products while focusing on latency and performance.
Build cloud profiling tools like Nsight Systems and Nsight Cloud, collecting and visualizing performance data for clusters and cloud environments using C++, Python, JavaScript, and Go.
Senior Embedded Software Engineer builds high-performance orchestration software that coordinates classical and quantum computing resources across CPUs, GPUs, FPGAs, and high-speed interconnects.
Own developer and partner adoption of NVIDIA’s Isaac GR00T end-to-end robotics workflow, guiding teams from model training to real-time deployment on Jetson Thor.
Build and run reliability tests for NVIDIA’s datacenter OS and server platforms, automate test suites with Python/CI/CD, and debug failures in Linux, firmware, and AI tooling stacks.
Lead developer adoption of NVIDIA’s Cosmos 3 world model for robotics, guiding partners through synthetic data generation, neural simulation, and Physical AI workflows using NVIDIA’s stack.
NVIDIA is looking for a hands-on Solutions Architect Manager to lead a team of GPU, networking & software solution architects and engineers. Do you want to build and lead a group that designs, debugs, and deploys…
Drive adoption of NVIDIA’s GPU-accelerated data processing stack by integrating libraries like cuDF and RAPIDS into third-party query engines and OLAP databases, shaping cross-ecosystem acceleration strategies.
Build and extend NVIDIA Warp, a Python framework for high-performance simulation and differentiable programming on GPUs, integrating CUDA features and partnering with internal teams to drive adoption.
Build reinforcement-learning environments to train frontier models, blending research and engineering with Python, PyTorch/JAX, and transformer internals.
Build reinforcement-learning environments and reward functions to train frontier AI models, blending research with engineering in a startup setting.
We couldn't check your fit for this role — add a CV to your profile to see it next time.