Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Lead a team of compiler engineers to design and optimize ML compiler solutions for AWS Neuron, driving hardware bring-up and influencing chip design for Trainium-based ML accelerators.
Build and optimize the Neuron compiler to transform ML models (PyTorch, TensorFlow, JAX) for AWS Inferentia/Trainium chips, solving compiler passes and performance tuning for large language and diffusion models.
Optimize AWS Neuron’s ML software stack for performance on custom Inferentia/Trainium accelerators, writing high-performance kernels and collaborating across compiler, frameworks, and hardware teams.
Build and optimize low-level compute kernels and inference pipelines for large language models running on custom ML hardware, integrating with frameworks like PyTorch and vLLM.
Builds and optimizes low-level compute kernels and serving infrastructure for large language model inference on custom ML hardware, integrating frameworks like PyTorch and vLLM.
Build and lead the Edge AI ML platform that trains, optimizes, and deploys large generative models on devices and in the cloud using PyTorch, TensorFlow, and Kubernetes.
Build and deploy AI models, RAG pipelines, and LLM-powered systems for early-stage startups backed by a VC firm.
Senior Applied Scientist designing and deploying ML compiler analyzers and tooling to optimize deep learning models across frameworks like PyTorch and TensorFlow, working with custom ML accelerators like Inferentia and Trainium.
Build and optimize low-level compute kernels and serving integrations for large language model inference on custom ML hardware, spanning model execution, memory management, and distributed systems.
Design and advise customers on generative AI and ML solutions using AWS services like Amazon Bedrock and SageMaker, including RAG pipelines and agentic workflows.
Build and ship ML infrastructure for autonomous freight systems: multimodal data pipelines, model training/evaluation workflows, and low-latency inference tooling using PyTorch/TensorFlow/JAX.
Наша команда развивает современные цифровые продукты, применяет smart-модели прогнозирования, которые рассчитаны как для внутренних пользователей Сбера, так для внешнего рынка, с постоянно растущем количеством…
Develops and integrates Java-based business-process automation for real-estate lease management using Platform V Flow and Spring/Hibernate stacks.
Builds and optimizes AI model quantization tooling to deploy neural networks efficiently on Qualcomm’s Snapdragon platforms, using Python/C++ and frameworks like PyTorch.
About us We are building AI systems that can reason, use tools, and complete meaningful work in the real world. Our team works across model post-training, reinforcement-learning infrastructure, large-scale training,…
Build and lead large-scale AI systems for federal programs, designing Python-based LLM solutions, RAG pipelines, and cloud-native architectures while ensuring compliance and security in regulated environments.
Lead the AI software stack, compiler, and platform roadmap for a next-gen AI infrastructure startup, defining toolchains, framework integrations, and runtime engines to optimize AI model execution on custom hardware.
Lead ML initiatives for autonomous-vehicle behavior and planning, designing and deploying models like transformers and reinforcement learning to drive real-world robotaxis and logistics fleets.
Senior Java backend developer building REST/SOAP APIs, JMS integrations, and Camunda workflows on Wildfly/JBoss for enterprise clients.
We couldn't check your fit for this role — add a CV to your profile to see it next time.