Tech jobs
Job listings
LLM Engineer (Data and Optimization)
Build, optimize, and deploy large language and multimodal models for industrial use, focusing on training, compression, RAG, and agent workflows.
LLM Research Intern
Research intern experimenting with open-weight LLMs, fine-tuning, RAG, and evaluation to adapt models for enterprise use cases alongside ProCogia’s AI teams.
Software Engineer II — Agentic AI Foundations
Build a secure, vendor-agnostic agent platform for identity verification workflows, integrating LLMs, orchestration, and safety controls in production.
ML/NLP Engineer
Build and deploy NLP models (Transformers, LLMs) for text analysis, RAG, and multi-agent systems on a cloud-native data-processing platform handling 10M events and 10TB daily.
ML Infrastructure Engineer
Build and optimize GPU infrastructure for AI workloads, profiling performance across hardware and frameworks to guide platform decisions and hardware development.
MLOps
Build and run a production-grade ML platform: design AI system architectures, deploy and optimize LLM inference servers, and maintain MLOps pipelines with GPU scheduling and monitoring.
ML Platform Engineer
ML Platform Engineer builds and scales AI model pipelines, deploys ML services on Kubernetes, and maintains FastAPI-based inference APIs for a large insurer’s internal AI ecosystem.
Senior AI Engineer
Build, deploy, and optimize production AI systems including LLMs and generative AI models, focusing on inference pipelines, performance, and cost efficiency.
Staff AI Engineer
Lead architecture and deployment of production AI systems, optimizing inference pipelines and mentoring engineers in a fast-growing company.
Forward Deployed Engineer
Forward Deployed Engineers embed with clients to deploy production-grade AI systems using SeekrFlow, specializing in LLM fine-tuning, agentic workflows, and Kubernetes-based deployments across cloud/on-prem environments.
Staff AI/ML Engineer, Large Language Model
Lead a team to build and deploy large language model applications, fine-tune models, and design RAG systems for mission-critical government use cases.
Senior Principal Backend Development Engineer
Build and own the agentic AI platform for a crypto exchange, designing orchestration engines, MCP servers, and lifecycle toolchains to power conversational agents handling trading, compliance, and customer service.
ML Engineer
Build and deploy ML/LLM services for an e-commerce platform, including RAG, agentic workflows, and quality evaluation, using Python, FastAPI, and cloud infrastructure.
Helix AI Engineer, Training Performance
Optimize distributed AI training for 100B+ parameter models across 100k+ GPUs, writing CUDA/Triton kernels and co-designing hardware-efficient training recipes.
#EG AI Engineer
Build and ship production-grade AI systems: multi-agent orchestration, RAG pipelines, and secure LLM integrations using Python/Go and Kubernetes.
AI/ML Engineer
Build production-grade AI architecture blueprints and open-source Quickstarts with Python, PyTorch, and Kubernetes, focusing on enterprise deployment, security, and regulated environments.
AI Platform Engineer
Build and scale an AI model service platform from scratch, designing APIs, routing, and load balancing for high-volume inference across connected hardware.
AI Platform Engineer
Build and maintain a secure AI platform for law-enforcement investigations, integrating LLMs, RAG, and agentic workflows with backend services and frontend interfaces.
Разработчик ИИ (backend)
Разрабатывает ИИ-агентов и backend-сервисы на Python (FastAPI) для платформы цифрового строительства, интегрируя LLM-модели и аналитику в реальном времени.
Member of Technical Staff — RL Research (New PhD Grad)
Research and build reinforcement-learning and post-training systems for real-time, full-duplex AI avatars that listen, speak, and react with emotional intelligence.