AI Job Radar

Model Optimization Jobs

Aktuelle KI-Jobs mit Model Optimization, passende Lernpfade und Bewerbungsbezug.

How to use Model Optimization in applications

If a job requires Model Optimization, the skill should be supported by a project, course or portfolio example. The application check reviews whether the skill is actually evidenced in your CV.

20
Results
12
Companies
94.8
Average score
2
Remote

10 results on this page. 20 results in total. More results are available via pagination, company pages, skill pages and job detail pages.

US-WA-BellevueGermanyFullTimeashby2026-09-10

Why this is a real AI job: The role is explicitly focused on the research and development of LLM inference systems and optimization, spanning the entire inference stack. The job description details numerous tasks directly related to AI/ML model performance, system design, and AI-driven…

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collabor…

Details Open source / apply

United StatesUSAgreenhouse2026-09-10

Why this is a real AI job: The role is explicitly focused on developing and deploying LLMs, ML models, and AI products. The description details tasks such as fine-tuning LLMs, building AI-powered features, and contributing to the ML platform architecture. The required expertise heavily…

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible f…

Details Open source / apply

US-WA-BellevueUSAFullTimeashby2026-09-10

Why this is a real AI job: The role is explicitly focused on the research and development of LLM inference systems, optimization techniques, and AI-native engineering. The entire job description revolves around core AI/ML concepts and their practical application.

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collabor…

Details Open source / apply

Senior Machine Learning Engineer, AI Performance

Wayve · London, United Kingdom

95/100
London, United KingdomUKgreenhouse2026-09-10

Why this is a real AI job: The role is explicitly focused on machine learning engineering for a company building AI-powered autonomous driving systems. The responsibilities center around training, optimizing, and deploying ML models in production.

About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems. Our vision is to create…

Details Open source / apply

United States and CanadaUSAFullTimeashby2026-09-10

Why this is a real AI job: The role is explicitly focused on applying and improving machine learning techniques (specifically LLMs) at scale. Responsibilities directly involve building ML pipelines, optimizing models, and working with large datasets – all core AI activities.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…

Details Open source / apply

San FranciscoUSAFullTimeashby2026-09-10

Why this is a real AI job: The role explicitly focuses on AI research and development, specifically large language models (LLMs) and their application to Perplexity's core products. The responsibilities are heavily centered around model training, optimization, and implementation, indic…

Perplexity is seeking top-tier AI Research Scientists and Engineers to advance our AI products and capabilities. We're building the future of AI-powered search and agent experiences through our Sonar models, Deep Research Agent, Comet Agent, and Search products. Join us in creating SOTA experiences…

Details Open source / apply

London, UKUSAgreenhouse2026-09-10

Why this is a real AI job: The role is explicitly focused on applied AI research, building agentic systems, optimizing models, and evaluating AI performance in critical applications. The job description heavily emphasizes AI/ML/LLM concepts and their practical implementation.

Scale’s mission is to develop reliable AI systems for the world's most important decisions. Our core work consists of: Creating custom AI applications that will impact millions of citizens Generating high-quality training data for national LLMs Upskilling and advisory services to spread the impact…

Details Open source / apply

San FranciscoUSAFullTimeashby2026-09-10

Why this is a real AI job: The role explicitly focuses on designing, deploying, and optimizing machine learning models, specifically LLMs, for content understanding and abuse prevention. The core responsibilities are deeply rooted in AI/ML engineering.

About the Team The Integrity team at OpenAI is dedicated to ensuring that our cutting-edge technology is not only revolutionary, but also secure from a myriad of adversarial threats. We strive to maintain the integrity of our platforms as they scale. The Integrity team is at the front lines of defe…

Details Open source / apply

RemoteUSAgreenhouse2026-09-08

Why this is a real AI job: The role is entirely focused on the development and optimization of LLM inference frameworks, distributed systems, and related technologies. The responsibilities and requirements clearly indicate a core AI/ML engineering position.

About the Role At Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize inference frameworks, algorithms, and infrastructure, pushing the boundaries of performance, scalability, and cost-e…

Details Open source / apply

San FranciscoUSAgreenhouse2026-08-31

Why this is a real AI job: The role is explicitly focused on building and optimizing AI/ML systems, specifically around inference and reinforcement learning for large language models. The responsibilities and requirements heavily emphasize core AI/ML skills and concepts.

About the Role The Turbo team sits at the intersection of efficient inference (algorithms, architectures, engines) and post‑training / RL systems. We build and operate the systems behind Together’s API, including high‑performance inference and RL/post‑training engines that can run at production sca…

Details Open source / apply