Point your AI agent at freehire and let it find you a job. A CLI and an MCP server over the whole job API — no browser.
Research analyst role fine-tuning large language models by solving complex math problems, breaking down solutions into clear steps, and validating claims to improve AI reasoning and accuracy.
Build and optimize Java-based back-end components for training and evaluating large language models, collaborating with cross-functional teams to improve AI performance and alignment.
Build and deploy LLM-based GenAI systems using Python, Langchain, and RAG pipelines, collaborating with teams to shape AI roadmaps and productionize models.
Leads end-to-end GenAI implementation for a Fortune 500 financial-services client, designing architecture, setting security guardrails, and managing delivery from POC to production. Works with AI/ML engineers and client stakeholders to ensure compliant, scalable AI solutions for valuation workflows.
Build and evaluate AI models, create ML tasks, and provide domain-specific feedback for frontier AI research projects on a contract basis.
Join a leading AI lab’s cutting-edge GenAI team to be at the core of the AI revolution, where your expertise fuels the development of the most advanced Large Language Models. 1. Overview Join a leading AI lab’s…
Designs grading criteria and scores AI-generated UI/UX work samples for a research project, focusing on real-world digital products like dashboards and onboarding flows.
1. Role Overview Mercor is partnering with a leading AI research organization to engage experienced accountants for a project focused on evaluating how well AI systems perform real-world accounting work. Rather than…
Design and validate security-review tasks in AI-powered simulated environments for procurement workflows, evaluating vendor SOC 2 reports and questionnaires against buyer standards.
Mercor is working with a leading AI research lab to improve the capabilities of next-generation AI systems. We are seeking experienced Insurance Verification and Eligibility & Benefits Managers to evaluate AI tools…
Red-team frontier AI models to expose hidden failure modes and edge cases, then turn findings into rigorous benchmark tasks for evaluation.
Build and ship full-stack apps and tools in Python, Java, Rust, C#, C++, or TypeScript that exercise cutting-edge GenAI models end-to-end.
Develop and validate C/C++ code for next-gen dialog agents, ensuring quality and scalability while collaborating with cross-functional teams in a fast-paced AI environment.
Design and refine AI training tasks by writing technical documentation, editing complex content, and creating evaluation rubrics for frontier AI agents in corporate settings.
Senior technical editor creates and refines complex AI evaluation tasks, style guides, and documentation to test and improve frontier AI agents' reliability in corporate settings.
About Turing: Based in San Francisco, California, Turing is the world’s leading research accelerator for frontier AI labs and a trusted partner for global enterprises deploying advanced AI systems. Turing supports…
Build and deploy end-to-end ML systems—from data pipelines to production models—using Python and libraries like Pandas, Scikit-learn, and PyTorch.
Designs and refines AI training tasks focused on market research and analytics, creating realistic survey scenarios and data-analysis prompts for frontier AI models.
Evaluate and benchmark AI-generated code for correctness and quality, create coding benchmarks, and provide structured feedback to improve LLM coding capabilities.
About Turing Turing is one of the world’s fastest-growing AI companies, accelerating the advancement and deployment of powerful AI systems. Turing helps customers in two ways: Working with the world’s leading AI labs…
We couldn't check your fit for this role — add a CV to your profile to see it next time.