India OfficeUSAFullTimeashby2026-07-31
Warum echter KI-Job: The role is explicitly focused on bringing up and optimizing machine learning models and frameworks on specialized hardware. The responsibilities directly involve core AI/ML tasks like model architecture translation, performance tuning, and debugging.
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…
Details Quelle / Bewerbung öffnen
Toronto OfficeUSAFullTimeashby2026-07-31
Warum echter KI-Job: The role is explicitly focused on bringing up and optimizing large language models (LLMs) on specialized hardware. The responsibilities directly involve working with model architectures, compilers, runtimes, and performance tuning – all core AI/ML tasks.
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…
Details Quelle / Bewerbung öffnen
US and Canada OfficesUSAFullTimeashby2026-07-31
Warum echter KI-Job: The role explicitly focuses on applying and improving machine learning techniques (specifically LLMs) at scale. Responsibilities center around building ML pipelines, optimizing models, and working with large datasets – all core AI activities.
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…
Details Quelle / Bewerbung öffnen
San FranciscoUSAFullTimeashby2026-07-31
Warum echter KI-Job: The role is explicitly focused on building and optimizing infrastructure for large-scale LLM inference, a core AI task. The description heavily emphasizes AI/ML technologies and their application.
About Anyscale At Anyscale https://www.anyscale.com/, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray https://docs.ray.io/en/latest/, a popular open-source project that's creating an ecosystem of li…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-07-31
Warum echter KI-Job: The role is explicitly focused on building and optimizing AI/ML systems, specifically around inference and reinforcement learning for large language models. The responsibilities and requirements heavily emphasize core AI/ML skills and concepts.
About the Role The Turbo team sits at the intersection of efficient inference (algorithms, architectures, engines) and post‑training / RL systems. We build and operate the systems behind Together’s API, including high‑performance inference and RL/post‑training engines that can run at production sca…
Details Quelle / Bewerbung öffnen
RemoteUSAgreenhouse2026-07-31
Warum echter KI-Job: The role is entirely focused on the development and optimization of LLM inference frameworks, distributed systems, and related technologies. The responsibilities and requirements clearly indicate a core AI/ML engineering position.
About the Role At Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize inference frameworks, algorithms, and infrastructure, pushing the boundaries of performance, scalability, and cost-e…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-07-31
Warum echter KI-Job: The role is entirely focused on building and optimizing the model serving layer for voice applications, including LLMs, STT, and TTS. It requires deep expertise in ML engineering, inference optimization, and GPU utilization. The responsibilities and requireme…
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a…
Details Quelle / Bewerbung öffnen
Hannover, Niedersachsen, DeutschlandDeutschlandba_search2026-07-31
Warum echter KI-Job: Der Jobtitel und die kurze Beschreibung deuten stark auf eine Kern-KI-Rolle hin. Der Fokus liegt direkt auf KI-Entwicklung.
AI Engineer (m/w/d) KI-Engineer
Details Quelle / Bewerbung öffnen
Dresden, Sachsen, DeutschlandDeutschlandba_search2026-07-31
Warum echter KI-Job: Der Jobtitel und die kurze Beschreibung deuten stark auf eine Kern-KI-Rolle hin. Der Fokus liegt direkt auf KI-Entwicklung.
AI Engineer (m/f/d) KI-Engineer
Details Quelle / Bewerbung öffnen
AUDI AG · Ingolstadt, Donau, Bayern, Deutschland
95/100
Ingolstadt, Donau, Bayern, DeutschlandDeutschlandba_search2026-07-31
Warum echter KI-Job: Der Jobtitel und das Unternehmen deuten stark auf eine Kern-KI-Rolle hin. Die Beschreibung ist zwar kurz, aber der Fokus liegt eindeutig auf KI.
Audi dual - Künstliche Intelligenz
Details Quelle / Bewerbung öffnen