AI Job Radar

Inference Jobs – Page 2

Aktuelle KI-Jobs mit Inference, passende Lernpfade und Bewerbungsbezug.

How to use Inference in applications

If a job requires Inference, the skill should be supported by a project, course or portfolio example. The application check reviews whether the skill is actually evidenced in your CV.

19
Results
9
Companies
94.2
Average score
4
Remote

9 results on this page. 19 results in total. More results are available via pagination, company pages, skill pages and job detail pages.

ParisFranceFull-timelever2026-07-10

Why this is a real AI job: The role explicitly focuses on deploying and scaling AI products, working with customers on AI solutions, and contributing to open-source AI codebases. The job description heavily emphasizes AI/ML technologies and their application in production environments.

About Mistral At Mistral AI, we believe in the power of AI to simplify tasks, save time, and enhance learning and creativity. Our technology is designed to integrate seamlessly into daily working life. We democratize AI through high-performance, optimized, open-source and cutting-edge models, produ…

Details Open source / apply

SeoulFrancelever2026-07-10

Why this is a real AI job: The role explicitly focuses on deploying and integrating Mistral AI's products (LLMs) with customers, involving end-to-end execution of AI solutions. The job description heavily emphasizes working directly with AI models, infrastructure for AI, and solving co…

About Mistral At Mistral AI, we believe in the power of AI to simplify tasks, save time, and enhance learning and creativity. Our technology is designed to integrate seamlessly into daily working life. We democratize AI through high-performance, optimized, open-source and cutting-edge models, produ…

Details Open source / apply

Machine Learning Engineer

Together AI · San Francisco

95/100
San FranciscoUSAgreenhouse2026-07-01

Why this is a real AI job: The role explicitly focuses on developing systems for LLM inference and fine-tuning, requiring deep expertise in ML and related technologies. The company is a research-driven AI company.

About the Role Together AI is looking for an ML Engineer who will develop systems and APIs that enable our customers to perform inference and fine tune LLMs. Relevant experience includes implementing runtime systems that perform inference at scale using AI/ML models from simple models up to the lar…

Details Open source / apply

Toronto OfficeUSAgreenhouse2026-06-26

Why this is a real AI job: The role is explicitly focused on LLM inference performance, model evaluation, and optimization on specialized hardware. The responsibilities and required skills are deeply rooted in AI/ML concepts and techniques.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training…

Details Open source / apply

Full Stack LLM Engineer

Cerebras Systems · Toronto Office

95/100
Toronto OfficeUSAgreenhouse2026-06-26

Why this is a real AI job: The role is explicitly focused on bringing up and optimizing large language models (LLMs) on specialized hardware. The responsibilities and required skills are heavily centered around AI/ML concepts and implementation.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training…

Details Open source / apply

SingaporeFrancelever2026-06-23

Why this is a real AI job: The role explicitly focuses on deploying and integrating AI products with customers, working on complex AI solutions, and contributing to open-source AI codebases. The job description heavily emphasizes AI/ML technologies and their application in production e…

About Mistral At Mistral AI, we believe in the power of AI to simplify tasks, save time, and enhance learning and creativity. Our technology is designed to integrate seamlessly into daily working life. We democratize AI through high-performance, optimized, open-source and cutting-edge models, produ…

Details Open source / apply

San FranciscoUSAFullTimeashby2026-07-31

Why this is a real AI job: The role is explicitly focused on AI/LLM inference, solution architecture for AI products, and working with customers deploying AI models. The responsibilities heavily involve technical AI concepts and deployments.

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the fronti…

Details Open source / apply

San FranciscoUSAgreenhouse2026-07-01

Why this is a real AI job: The role is focused on building and optimizing a platform for custom models and inference, specifically for video and audio generation. The responsibilities directly involve ML bottlenecks, model bring-up, optimization, and scaling. The company is a research-…

About the Role Our team focuses on enabling custom models and dedicated inference on Together. We are responsible for building a container platform, optimizing autoscaling, minimizing cold starts, achieving the best end-to-end model performance, and providing a best-in-class developer experience wi…

Details Open source / apply

AI Product Engineer

Fireworks AI · New York, San Mateo

90/100
New York, San MateoUSAgreenhouse2026-06-19

Why this is a real AI job: The role is deeply embedded in building and improving a generative AI platform, focusing on core components like inference, fine-tuning, and model deployment. The job description explicitly mentions working with LLMs and AI infrastructure.

About Us: At Fireworks, we’re building the future of generative AI infrastructure. Our platform delivers the highest-quality models with the fastest and most scalable inference in the industry. We’ve been independently benchmarked as the leader in LLM inference speed and are driving cutting-edge in…

Details Open source / apply