Stories
15
Sources
1
Topics
5
For You lens
8 stories in this edition match your reader profile.
Reader signals
3
Searches
0
Matches
8
Top score
109
Edition Index
Topic, entity, and source map
Lead Story
Who Cares About LLM costs?
Why would you worry about them? LLMs promise something close to infinite capability, and worrying about the meter feels like someone else's problem. Ship the feature. Let the model think as long as it needs to, right? Right. Until the bill arrives. The shock Picture it: your
Mozilla.ai Blog / 2:12 PM
Open Models are ready for agents. Their APIs are not.
Open models have reached impressive levels of capability, but the surrounding platform often falls short. This blog explores the hidden infrastructure gaps developers face and why bridging them is essential for production-ready AI applications.
Mozilla.ai Blog / 3:18 PM
Using Octonous as an AI Safety Engineer
Discover how Octonous helps our team at Mozilla.ai cut through daily busywork, from automatically monitoring our libraries to staying on top of the latest developments in AI safety.
Mozilla.ai Blog / 1:31 PM
Introducing Otari: The Open-Source LLM Control Plane
If you are building LLM-powered applications today, you are probably managing multiple LLM providers, a pile of API keys, and your own logic for routing, budgets, and failovers. Accessing language models is no longer the challenge. Operating LLM infrastructure effectively is. Today, we are excited to launch Otari, the
Mozilla.ai Blog / 3:58 PM
Announcing transcribe.cpp
Meet transcribe.cpp, a new open-source C/C++ speech-to-text inference library with portable, GPU-accelerated support for multiple STT models. Developed through Mozilla.ai's Builders in Residence program, it makes adding fast, local transcription to applications easier than ever.
Mozilla.ai Blog / 4:02 PM
Using Octonous as a Product Manager
A look at how we use Octonous inside mozilla.ai to reduce the everyday overhead of product work, from turning Slack feedback into GitHub issues to staying on top of product changes and finding context across the tools where work already happens.
Mozilla.ai Blog / 3:24 PM
Image Classification Comes to encoderfile
Encoderfile now handles images. Starting with image classification, you can run vision models as a single executable — no Python runtime, no serving infrastructure, just a file path in and a label out.
Mozilla.ai Blog / 9:00 AM
What is an LLM control plane?
Runaway agents? Provider outages? Discover why your AI stack needs an LLM control plane, not just a gateway, to handle production routing, budgets, and privacy.
Mozilla.ai Blog / 1:40 PM
Use the Otari Gateway with OpenCode
AI coding sessions can feel like a black box. Route OpenCode through the Otari Gateway to track costs, token usage, and model activity in real time. Get budget controls and visibility across every session without changing a single line of application code.
Mozilla.ai Blog / 4:02 PM
cq exchange: Agents without Borders
cq exchange gives agents a shared place to store and retrieve experience-driven knowledge through private namespaces and a public commons.
Mozilla.ai Blog / 10:14 AM
The Control Layer: Why the Next Era of AI Is About Infrastructure, Not Just Models
The Model’s the Easy Part - How to Get, and Keep, Value Here’s how I see the evolution of AI in enterprises over the last few years: Autumn of 2022, the world thinks it’s going into a recession. IT budgets are frozen for 2023.
Mozilla.ai Blog / 3:17 PM
Otari: Own Your AI Stack
Meet Otari, an open-source LLM gateway powered by any-llm, and Otari.ai, the hosted platform built on the same foundation. Run frontier or open-weights models through one API with usage tracking, budget controls, routing policies, observability, and team management.
Mozilla.ai Blog / 4:09 PM
AI Got Expensive. Now What?
Cloud AI pricing changed fast in 2026. This post looks at why more teams are moving back to local models, the tradeoffs behind tools like Ollama and LM Studio, and why portability and ownership are becoming bigger concerns for developers.
Latest story in this edition: 9:38 AM
Back to front page