SF OfficeUSAFullTimeashby2026-08-01
Warum echter KI-Job: The role is explicitly focused on designing, building, and evaluating LLM-powered systems for a core healthcare application. The job description heavily emphasizes LLM APIs, agentic workflows, and model evaluation, indicating a primary focus on AI/ML.
ABOUT ABRIDGE Abridge was founded in 2018 with the mission of powering deeper understanding in healthcare. Our AI-powered platform was purpose-built for medical conversations, improving clinical documentation efficiencies while enabling clinicians to focus on what matters most—their patients. Our e…
Details Quelle / Bewerbung öffnen
Toronto OfficeUSAFullTimeashby2026-08-01
Warum echter KI-Job: The role is explicitly focused on bringing up and optimizing large language models (LLMs) on specialized hardware. The responsibilities directly involve working with model architectures, compilers, runtimes, and performance tuning – all core AI/ML tasks.
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…
Details Quelle / Bewerbung öffnen
San FranciscoUSAFullTimeashby2026-08-01
Warum echter KI-Job: The role is explicitly focused on building and optimizing infrastructure for large-scale LLM inference, a core AI task. The description heavily emphasizes AI/ML technologies and their application.
About Anyscale At Anyscale https://www.anyscale.com/, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray https://docs.ray.io/en/latest/, a popular open-source project that's creating an ecosystem of li…
Details Quelle / Bewerbung öffnen
San FranciscoUSAFullTimeashby2026-08-01
Warum echter KI-Job: The role is explicitly focused on building, scaling and optimizing LLM inference workloads for customers. The team directly contributes to the core Baseten codebase related to AI products. Strong emphasis on hands-on technical work with LLMs.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the fronti…
Details Quelle / Bewerbung öffnen
Berlin OfficeGermany/GlobalFullTimeashby2026-08-01
Warum echter KI-Job: The role is explicitly focused on building AI-powered features into the core product, including LLM integration, prompt engineering, and AI feature lifecycle management. The requirements and bonus points heavily emphasize AI/ML expertise.
The AI orchestration of your wildest imagination. n8n is the open workflow orchestration platform built for the new era of AI. We give technical teams the freedom of code with the speed of no-code, so they can automate faster, smarter, and without limits. Backed by a fiercely inventive community an…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-08-01
Warum echter KI-Job: The role is explicitly focused on building and optimizing AI/ML systems, specifically around inference and reinforcement learning for large language models. The responsibilities and requirements heavily emphasize core AI/ML skills and concepts.
About the Role The Turbo team sits at the intersection of efficient inference (algorithms, architectures, engines) and post‑training / RL systems. We build and operate the systems behind Together’s API, including high‑performance inference and RL/post‑training engines that can run at production sca…
Details Quelle / Bewerbung öffnen
RemoteUSAgreenhouse2026-08-01
Warum echter KI-Job: The role is entirely focused on the development and optimization of LLM inference frameworks, distributed systems, and related technologies. The responsibilities and requirements clearly indicate a core AI/ML engineering position.
About the Role At Together.ai, we are building state-of-the-art infrastructure to enable efficient and scalable inference for large language models (LLMs). Our mission is to optimize inference frameworks, algorithms, and infrastructure, pushing the boundaries of performance, scalability, and cost-e…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-08-01
Warum echter KI-Job: The role is explicitly focused on building and optimizing the model serving layer for voice applications, working with state-of-the-art voice models and inference engines. The responsibilities are heavily centered around ML engineering tasks.
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-08-01
Warum echter KI-Job: The role is entirely focused on building and optimizing the model serving layer for voice applications, including LLMs, STT, and TTS. It requires deep expertise in ML engineering, inference optimization, and GPU utilization. The responsibilities and requireme…
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a…
Details Quelle / Bewerbung öffnen
BMW AG · München, Bayern, Deutschland
95/100
München, Bayern, DeutschlandDeutschlandba_search2026-08-01
Warum echter KI-Job: The job title explicitly mentions 'AI Engineer' and focuses on 'Agentic AI', indicating a core AI role. The company is a major automotive manufacturer, suggesting application to advanced robotics and autonomous systems.
Senior Agentic AI Engineer (f/m/x) Informatiker/in
Details Quelle / Bewerbung öffnen