Flat pricing, metered simply
Every plan is a flat rate on the calls you send — never a percentage of your spend, and never a cut of your savings. What Metergraph finds is yours to keep.
For trying Metergraph on a real service and seeing your spend clearly for the first time.
For teams whose product runs on LLM calls and whose margins move when the model mix does.
For organizations with traffic that can't leave their environment.
Building something now? We're onboarding a small number of design partners and working closely with each one — book a call to see if it's a fit.
Questions
One captured LLM request/response pair, recorded by the SDK or exported from your gateway. Spans grouped under a trace each count once.
Because it puts our incentives against yours: percentage pricing rewards vendors for big one-time cuts, not for keeping quality up as models and traffic drift. Flat metering means the recommendations only have to be right.
No. Collection is shadow-only and asynchronous — Metergraph never holds live inference credentials and can't fail your product's calls. Optimizations ship through your normal deploy process, validated first.
That's what Scale is for: start with staging-only ingestion, then deploy Metergraph into your VPC or on-prem so production traces never leave your boundary.