Toronto, CANUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on bringing up and optimizing large language models (LLMs) on specialized hardware. The responsibilities directly involve the entire ML model lifecycle - from architecture translation to performance tuning. Strong emphasis on de…
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…
Details Quelle / Bewerbung öffnen
United States and CanadaUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on applying and improving machine learning techniques (specifically LLMs) at scale. Responsibilities directly involve building ML pipelines, optimizing models, and working with large datasets – all core AI activities.
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transfor…
Details Quelle / Bewerbung öffnen
San FranciscoUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on building and optimizing infrastructure for large-scale LLM inference, a core AI task. The description heavily emphasizes AI/ML technologies and their application.
About Anyscale At Anyscale https://www.anyscale.com/, we're on a mission to democratize distributed computing and make it accessible to software developers of all skill levels. We’re commercializing Ray https://docs.ray.io/en/latest/, a popular open-source project that's creating an ecosystem of li…
Details Quelle / Bewerbung öffnen
San FranciscoUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on building, scaling and optimizing LLM inference workloads for customers. The team directly contributes to the core Baseten codebase related to AI products. Strong emphasis on hands-on technical work with LLMs.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the front…
Details Quelle / Bewerbung öffnen
San FranciscoUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role explicitly focuses on building and shipping AI/LLM-powered products, agents, and internal tooling. The responsibilities are heavily centered around AI system design, implementation, and evaluation.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the front…
Details Quelle / Bewerbung öffnen
Berlin OfficeGermany/GlobalFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on building AI-powered features into the core product, including LLM integration, prompt engineering, and AI feature lifecycle management. The requirements and bonus points heavily emphasize AI/ML expertise.
The AI orchestration of your wildest imagination. n8n is the open workflow orchestration platform built for the new era of AI. We give technical teams the freedom of code with the speed of no-code, so they can automate faster, smarter, and without limits. Backed by a fiercely inventive community an…
Details Quelle / Bewerbung öffnen
New York, NYUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on building, deploying, and improving AI agents and LLM-powered applications. The description heavily emphasizes AI/ML concepts and technologies. The candidate is expected to have deep understanding of AI system components.
ABOUT US At LangChain, our mission is to make intelligent agents ubiquitous. We build the foundation for agent engineering in the real world, helping developers move from prototypes to production-ready AI agents that teams can rely on. We began as widely adopted open-source tools and have grown to…
Details Quelle / Bewerbung öffnen
Redwood City, CAUSAFullTimeashby2026-09-10
Warum echter KI-Job: The role is explicitly focused on building and maintaining infrastructure for ML research and serving, with a strong emphasis on large language models and GPU utilization. The requirements and responsibilities directly relate to core AI/ML engineering tasks.
ABOUT THE ROLE We’re looking for seasoned ML Infrastructure engineers with experience designing, building and maintaining training and serving infrastructure for ML research. Responsibilities: - Provide infrastructure support to our ML research and product - Build tooling to diagnose cluster issues…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-09-10
Warum echter KI-Job: The role is explicitly focused on building and optimizing AI inference systems for large language models. The responsibilities and requirements heavily emphasize ML engineering, performance optimization, and working with cutting-edge AI technologies.
About the Role Together AI is seeking a Machine Learning Engineer to join our Inference Engine team, focusing on optimizing and enhancing the performance of our AI inference systems. This role involves working with state-of-the-art large language models models and ensuring they run efficiently and…
Details Quelle / Bewerbung öffnen
San FranciscoUSAgreenhouse2026-09-10
Warum echter KI-Job: The role is explicitly focused on building and optimizing the model serving layer for voice applications, working with state-of-the-art voice models and inference engines. The responsibilities are heavily centered around ML engineering tasks.
About the Role Together AI is building the best inference infrastructure for voice applications. Our Voice AI platform powers production-grade, real-time voice agents and applications — serving speech-to-text and text-to-speech models with best-in-class latency and reliability. We're looking for a…
Details Quelle / Bewerbung öffnen