San FranciscoUSAFullTimeashby2026-08-01
Why this is a real AI job: The role is explicitly focused on building and optimizing infrastructure for large-scale LLM inference, a core AI task. The description heavily emphasizes AI/ML technologies and their application.
About Anyscale At Anyscale https://www.anyscale.com/, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray https://docs.ray.io/en/latest/, a popular open-source project that's creating an ecosystem of li…
Details Open source / apply
San FranciscoUSAFullTimeashby2026-08-01
Why this is a real AI job: The role is explicitly focused on building, scaling and optimizing LLM inference workloads for customers. The team directly contributes to the core Baseten codebase related to AI products. Strong emphasis on hands-on technical work with LLMs.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the fronti…
Details Open source / apply
Berlin OfficeGermany/GlobalFullTimeashby2026-08-01
Why this is a real AI job: The role is explicitly focused on building AI-powered features into the core product, including LLM integration, prompt engineering, and AI feature lifecycle management. The requirements and bonus points heavily emphasize AI/ML expertise.
The AI orchestration of your wildest imagination. n8n is the open workflow orchestration platform built for the new era of AI. We give technical teams the freedom of code with the speed of no-code, so they can automate faster, smarter, and without limits. Backed by a fiercely inventive community an…
Details Open source / apply
Redwood City, CAUSAFullTimeashby2026-08-01
Why this is a real AI job: The role is explicitly focused on building and maintaining infrastructure for ML research and serving, with a strong emphasis on large language models and GPU utilization. The requirements and responsibilities directly relate to core AI/ML engineering tasks.
ABOUT THE ROLE We’re looking for seasoned ML Infrastructure engineers with experience designing, building and maintaining training and serving infrastructure for ML research. Responsibilities: - Provide infrastructure support to our ML research and product - Build tooling to diagnose cluster issues…
Details Open source / apply
RemoteUSAgreenhouse2026-08-01
Why this is a real AI job: The role is entirely focused on the development and optimization of LLM inference frameworks, distributed systems, and related technologies. The responsibilities and requirements clearly indicate a core AI/ML engineering position.
About the Role At Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize inference frameworks, algorithms, and infrastructure, pushing the boundaries of performance, scalability, and cost-e…
Details Open source / apply
milia.io · Köln, Nordrhein-Westfalen, Deutschland
95/100
Köln, Nordrhein-Westfalen, DeutschlandDeutschlandba_search2026-08-01
Why this is a real AI job: Der Jobtitel und die Beschreibung deuten klar auf eine Kern-KI-Rolle hin, die sich auf Machine Learning konzentriert.
Senior AI Engineer (m/w/d) - Remote Machine Learning Engineer
Details Open source / apply
BelgradeUSAFullTimeashby2026-08-01
Why this is a real AI job: The role explicitly focuses on building and improving search technologies using machine learning models, including LLMs and RAG pipelines. The responsibilities are heavily centered around model training, evaluation, and deployment. The required qualifications…
Perplexity is seeking an experienced Machine Learning Engineer to help build the next generation of advanced search technologies, with a focus on retrieval and ranking. Responsibilities - Relentlessly push search quality forward—through models, data, tools, or any other leverage available - Archite…
Details Open source / apply
New York CityUSAFullTimeashby2026-08-01
Why this is a real AI job: The role is explicitly focused on improving the quality of an LLM-first search engine through data science, metric design, and analysis of user interactions. The responsibilities directly involve working with LLMs and building data pipelines to support model…
Perplexity serves tens of millions of users daily with reliable, high-quality answers grounded in an LLM-first search engine and specialized data sources. The Answer Quality team ensures that our prompts, tools, search, and specialized datasets, combined with both frontier and in-house models, crea…
Details Open source / apply
San FranciscoUSAFullTimeashby2026-08-01
Why this is a real AI job: The role is entirely focused on building and maintaining evaluation pipelines for LLM-based products, designing evaluation sets, and improving answer quality. It requires deep expertise in data science, machine learning, and specifically LLMs.
Perplexity serves tens of millions of users daily with reliable, high-quality answers grounded in an LLM-first search engine and our specialized data sources. We aim to use the latest models as they are released, but the intelligence frontier is a jagged one, and popular benchmarks do not effective…
Details Open source / apply
United StatesCanadaFullTimeashby2026-08-01
Why this is a real AI job: The role explicitly focuses on building and deploying AI models (including small language models), defining AI impact measurement, agentic analytics, and driving the company's AI strategy. The core responsibilities are heavily centered around data science and…
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe th…
Details Open source / apply