AI Job Radar

Inference Jobs – Seite 2

Aktuelle KI-Jobs mit Inference, passende Lernpfade und Bewerbungsbezug.

Wie du Inference für Bewerbungen nutzt

Wenn ein Job Inference verlangt, sollte der Skill nicht nur als Stichwort im CV stehen. Besser sind ein kurzer Projektbeleg, ein Kursnachweis oder ein Portfolio-Beispiel. Für den Bewerbungscheck wird geprüft, ob der Skill in deinem Lebenslauf wirklich belegbar ist.

16
Treffer
8
Unternehmen
94.1
Durchschn. Score
3
Remote

6 Treffer auf dieser Seite. Insgesamt 16 Treffer. Weitere Treffer sind über die Seitennavigation, Firmen-, Skill- und Jobdetailseiten erreichbar.

Toronto OfficeUSAgreenhouse2026-06-26

Warum echter KI-Job: The role is explicitly focused on LLM inference performance, model evaluation, and optimization on specialized hardware. The responsibilities and required skills are deeply rooted in AI/ML concepts and techniques.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training…

Details Quelle / Bewerbung öffnen

Full Stack LLM Engineer

Cerebras Systems · Toronto Office

95/100
Toronto OfficeUSAgreenhouse2026-06-26

Warum echter KI-Job: The role is explicitly focused on bringing up and optimizing large language models (LLMs) on specialized hardware. The responsibilities and required skills are heavily centered around AI/ML concepts and implementation.

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training…

Details Quelle / Bewerbung öffnen

SingaporeFrancelever2026-06-23

Warum echter KI-Job: The role explicitly focuses on deploying and integrating AI products with customers, working on complex AI solutions, and contributing to open-source AI codebases. The job description heavily emphasizes AI/ML technologies and their application in production e…

About Mistral At Mistral AI, we believe in the power of AI to simplify tasks, save time, and enhance learning and creativity. Our technology is designed to integrate seamlessly into daily working life. We democratize AI through high-performance, optimized, open-source and cutting-edge models, produ…

Details Quelle / Bewerbung öffnen

San FranciscoUSAFullTimeashby2026-07-31

Warum echter KI-Job: The role is explicitly focused on AI/LLM inference, solution architecture for AI products, and working with customers deploying AI models. The responsibilities heavily involve technical AI concepts and deployments.

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the fronti…

Details Quelle / Bewerbung öffnen

San FranciscoUSAgreenhouse2026-07-01

Warum echter KI-Job: The role is focused on building and optimizing a platform for custom models and inference, specifically for video and audio generation. The responsibilities directly involve ML bottlenecks, model bring-up, optimization, and scaling. The company is a research-…

About the Role Our team focuses on enabling custom models and dedicated inference on Together. We are responsible for building a container platform, optimizing autoscaling, minimizing cold starts, achieving the best end-to-end model performance, and providing a best-in-class developer experience wi…

Details Quelle / Bewerbung öffnen

AI Product Engineer

Fireworks AI · New York, San Mateo

90/100
New York, San MateoUSAgreenhouse2026-06-19

Warum echter KI-Job: The role is deeply embedded in building and improving a generative AI platform, focusing on core components like inference, fine-tuning, and model deployment. The job description explicitly mentions working with LLMs and AI infrastructure.

About Us: At Fireworks, we’re building the future of generative AI infrastructure. Our platform delivers the highest-quality models with the fastest and most scalable inference in the industry. We’ve been independently benchmarked as the leader in LLM inference speed and are driving cutting-edge in…

Details Quelle / Bewerbung öffnen