San FranciscoUSAFullTimeashby2026-09-10
Why this is a real AI job: The role is explicitly focused on building and optimizing infrastructure for large-scale LLM inference, a core AI task. The description heavily emphasizes AI/ML technologies and their application.
About Anyscale At Anyscale https://www.anyscale.com/, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray https://docs.ray.io/en/latest/, a popular open-source project that's creating an ecosystem of li…
Details Open source / apply
San FranciscoUSAFullTimeashby2026-09-10
Why this is a real AI job: The role is explicitly focused on building, scaling and optimizing LLM inference workloads for customers. The team directly contributes to the core Baseten codebase related to AI products. Strong emphasis on hands-on technical work with LLMs.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the front…
Details Open source / apply
San FranciscoUSAFullTimeashby2026-09-10
Why this is a real AI job: The role explicitly focuses on building and shipping AI/LLM-powered products, agents, and internal tooling. The responsibilities are heavily centered around AI system design, implementation, and evaluation.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the front…
Details Open source / apply
New York, NYUSAFullTimeashby2026-09-10
Why this is a real AI job: The role is fundamentally focused on enabling customers to build and deploy agent systems using LangChain's core technologies. The job description explicitly requires experience building LLM/agent applications and a deep understanding of agent architectures.
ABOUT US At LangChain, our mission is to make intelligent agents ubiquitous. We build the foundation for agent engineering in the real world, helping developers move from prototypes to production-ready AI agents that teams can rely on. We began as widely adopted open-source tools and have grown to…
Details Open source / apply
New York, NYUSAFullTimeashby2026-09-10
Why this is a real AI job: The role is explicitly focused on building, deploying, and improving AI agents and LLM-powered applications. The description heavily emphasizes AI/ML concepts and technologies. The candidate is expected to have deep understanding of AI system components.
ABOUT US At LangChain, our mission is to make intelligent agents ubiquitous. We build the foundation for agent engineering in the real world, helping developers move from prototypes to production-ready AI agents that teams can rely on. We began as widely adopted open-source tools and have grown to…
Details Open source / apply
Redwood City, CAUSAFullTimeashby2026-09-10
Why this is a real AI job: The role is explicitly focused on building and maintaining infrastructure for ML research and serving, with a strong emphasis on large language models and GPU utilization. The requirements and responsibilities directly relate to core AI/ML engineering tasks.
ABOUT THE ROLE We’re looking for seasoned ML Infrastructure engineers with experience designing, building and maintaining training and serving infrastructure for ML research. Responsibilities: - Provide infrastructure support to our ML research and product - Build tooling to diagnose cluster issues…
Details Open source / apply
San FranciscoUSAgreenhouse2026-09-10
Why this is a real AI job: The role is explicitly focused on building and optimizing AI inference systems for large language models. The responsibilities and requirements heavily emphasize ML engineering, performance optimization, and working with cutting-edge AI technologies.
About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and ensuring they run efficiently and…
Details Open source / apply
San FranciscoUSAgreenhouse2026-09-10
Why this is a real AI job: The role is explicitly focused on building and optimizing the model serving layer for voice applications, working with state-of-the-art voice models and inference engines. The responsibilities are heavily centered around ML engineering tasks.
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a…
Details Open source / apply
San FranciscoUSAgreenhouse2026-09-10
Why this is a real AI job: The role is entirely focused on building and optimizing the model serving layer for voice applications, including LLMs, STT, and TTS. It requires deep expertise in ML engineering, inference optimization, and GPU utilization. The responsibilities and requireme…
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a…
Details Open source / apply
LondonUSA/UK/PolandFullTimeashby2026-09-10
Why this is a real AI job: The role is explicitly focused on building and maintaining datasets and evaluation workflows for AI safety, directly impacting the development and deployment of AI models. The description heavily emphasizes ML, data science, and AI-specific tasks.
ABOUT ELEVENLABS ElevenLabs is an AI research and product company transforming how we interact with technology. We launched in January 2023 with the first human-like AI voice model. Today, we serve millions of users and thousands of businesses - from fast-growing startups to large enterprises like…
Details Open source / apply