Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Build predictive world models using generative AI and deep learning to simulate how scenes evolve for autonomous driving and robotics, leveraging diffusion models and multimodal data.
AI Application Engineer (Part-Time) Location: Argentina, Brazil, Peru, Colombia, Costa Rica, Mexico Department: Workana Premium Workplace: remote Employment Type: full Description Client: medxprts.ai Location: Remote…
Research and develop quantization, pruning, and inference optimizations for LLM/VLM and MoE models to run efficiently on GPUs and NPUs.
Build, optimize, and deploy large language and multimodal models for industrial use, focusing on training, compression, RAG, and agent workflows.
Optimize distributed AI training for 100B+ parameter models across 100k+ GPUs, writing CUDA/Triton kernels and co-designing hardware-efficient training recipes.
Build and scale distributed training infrastructure for large-scale AI avatar models, optimizing GPU clusters, parallelism, and multimodal data pipelines for real-time, full-duplex training.
Engineer large-scale LLM pre-training pipelines and distributed GPU clusters using PyTorch, DeepSpeed, or Megatron-LM, optimizing networking, memory, and fault tolerance for month-long runs.
Designs high-speed SoC networking engines and interconnect fabrics for custom AI servers, optimizing distributed AI workloads with RDMA, RoCEv2, and collective-operation acceleration.
About Company Founded in 2022, XG Tech is driving the future of smart vehicles. Its mission is to empower the digital transformation of automobiles, moving from distributed computing to a centralized, cross-domain…
Build and fine-tune large language models to automate healthcare claims and conversational AI, using Python, PyTorch, and cloud ML platforms.
Build and scale AI/ML systems to personalize Qualtrics' SaaS platform, leading architecture, deployment, and optimization of large-scale machine learning models and infrastructure.
The Company: Faraday Future is a California-based technology company focused on the design, engineering, and development of intelligent, connected electric vehicles and related artificial intelligence–enabled…
ADVANCE YOUR CAREER. ADVANCE THE WORLD. At AMD, we believe technology can change lives for the better. It can heal us, entertain us, and make us more connected, productive, and understanding of the world around us. And…
Build and maintain an AI/MLOps platform to deploy large language models and AI agents, optimize distributed training, and manage GPU clusters on public clouds.
Lead the development of Grab’s proprietary foundation models and generative recommendation systems, scaling distributed training and deploying AI solutions for millions of users.
Build predictive world models using generative AI and deep learning to simulate how scenes evolve, training autonomous driving and robotics policies from multimodal sensor data.
Build and own the shared AI platform that trains and serves Adobe’s generative AI products at global scale, focusing on GPU fleet utilization, model serving, and distributed systems for low-latency inference.
Lead the design and optimization of large-scale distributed AI training systems, improving performance and scalability for advanced neural networks across GPU clusters.
Senior SRE builds and maintains Voleon’s AI research compute clusters, ensuring 24/7 uptime and performance for ML workloads across on-prem and cloud using IaC, observability stacks, and SRE practices.
Build and scale ML infrastructure for brain-computer interface R&D, including distributed training pipelines and large-scale data platforms to support neuroscientific modeling and neural decoding.
We couldn't check your fit for this role — add a CV to your profile to see it next time.