Every request.
Every millisecond.
See how response times vary across a session, from the typical request to the slowest few.
SMALL MODELS FOR AI-FIRST PRODUCTS
Turn user requests into product actions, without per-token fees. Small models trained for your app run directly on your users’ devices.
Explore the samples02 PERFORMANCE
Response time, task quality, and model size.
Explore a few of the things worth measuring.
Interactive visual samples. Minifield performance results are still to come.
See how response times vary across a session, from the typical request to the slowest few.
Compare task success with the space a model needs. Explore where a smaller footprint meets the quality your product needs.
Task success (%) / relative model size (%)
Track a training run alongside held-out examples. Watch for improvement that carries beyond the training data.
03 FOR PRODUCT TEAMS
Let users complete common tasks in their own words while they’re still learning your product.
Reduce the need for walkthroughs. Give new users a direct way to get work done before they’ve learned where every control lives.
Help users find and use the features you’ve already built. Turn familiar requests into supported actions, right where they’re working.
Give your team fewer routine workflows to explain. Let users filter, organize, and change views in their own words.
Make AI part of every session without a per-user token bill. Small models run on your customers’ devices, keeping hosted inference charges out of local interactions.
No per-token fees refers to local inference. Model training, integration, distribution, and any cloud fallback still have costs.
05 PERMISSIONS & PRIVACY
Bring AI into your product with the access controls you already use. Keep prompts and context on the user’s device during inference.
Use a small model trained for your product’s vocabulary and supported actions. Focus its training on the tasks your customers need.
Process requests on the user’s device, without sending prompts or context to an external AI provider.
The AI can only access what the user can access and perform actions they’re already allowed to take.
Inspired by the idea at the heart of Asimov’s Foundation: knowledge, distributed widely, gives small groups the power to shape what comes next.
Independent by design.The device in your hands should work for you. We believe intelligence should live there, too.
A focused model should know your product’s job deeply. Keep its footprint small enough to make broad access practical.
Users bring the goal. Your product should help them get there, with clear actions and an easy way back.
Keeping inference on your device means your instructions can stay there. That’s a design decision we make at the beginning.
We believe in tools that keep working. Local intelligence makes room for ownership, offline use, and life beyond the meter.