The AI Front Page

Source Edition

Mozilla.ai Blog

21 stories from 1 sources across 6 topics.

Stories

21

Sources

1

Topics

6

For You lens

10 stories in this edition match your reader profile.

Reader signals

3

Searches

0

Matches

10

Top score

81

Tune For You

Lead Story

Introducing Agent Skills in Octonous

Agent Skills in Octonous offer a central, reusable library for your context, guidelines, and procedures. Written in standard Markdown, Skills end repetitive setup. Store team knowledge once, share it across your workspace, and seamlessly attach it to automated tasks or live conversations.

Mozilla.ai Blog10:52 AMHeat 58
ReadSource

Mozilla.ai Blog / 10:01 AM

llamafile v0.10.5

llamafile v0.10.5 is out! Updated llama.cpp core support lets you run two huge models locally: the compressed 6GB Ternary Bonsai 27B and the fast 118B Laguna-S-2.1 coding MoE. This release also fixes docs and adds pre-built transcribefile speech-to-text binaries.

ReadSource

Mozilla.ai Blog / 1:31 PM

Introducing Otari: The Open-Source LLM Control Plane

If you are building LLM-powered applications today, you are probably managing multiple LLM providers, a pile of API keys, and your own logic for routing, budgets, and failovers. Accessing language models is no longer the challenge. Operating LLM infrastructure effectively is. Today, we are excited to launch Otari, the

ReadSource

Mozilla.ai Blog / 3:58 PM

Announcing transcribe.cpp

Meet transcribe.cpp, a new open-source C/C++ speech-to-text inference library with portable, GPU-accelerated support for multiple STT models. Developed through Mozilla.ai's Builders in Residence program, it makes adding fast, local transcription to applications easier than ever.

ReadSource

Mozilla.ai Blog / 1:40 PM

Use the Otari Gateway with OpenCode

AI coding sessions can feel like a black box. Route OpenCode through the Otari Gateway to track costs, token usage, and model activity in real time. Get budget controls and visibility across every session without changing a single line of application code.

ReadSource

Mozilla.ai Blog / 9:38 AM

Who Cares About LLM costs?

Why would you worry about them? LLMs promise something close to infinite capability, and worrying about the meter feels like someone else's problem. Ship the feature. Let the model think as long as it needs to, right? Right. Until the bill arrives. The shock Picture it: your

ReadSource

Mozilla.ai Blog / 4:02 PM

Using Octonous as a Product Manager

A look at how we use Octonous inside mozilla.ai to reduce the everyday overhead of product work, from turning Slack feedback into GitHub issues to staying on top of product changes and finding context across the tools where work already happens.

ReadSource

Mozilla.ai Blog / 3:17 PM

Otari: Own Your AI Stack

Meet Otari, an open-source LLM gateway powered by any-llm, and Otari.ai, the hosted platform built on the same foundation. Run frontier or open-weights models through one API with usage tracking, budget controls, routing policies, observability, and team management.

ReadSource

Mozilla.ai Blog / 4:09 PM

AI Got Expensive. Now What?

Cloud AI pricing changed fast in 2026. This post looks at why more teams are moving back to local models, the tradeoffs behind tools like Ollama and LM Studio, and why portability and ownership are becoming bigger concerns for developers.

ReadSource

Latest story in this edition: 10:52 AM

Back to front page