Pricing
Start with a free risk assessment. Evaluations are priced by scope, with continuous monitoring available afterwards.
Risk assessment
Freetwo-hour review
For a planned or live AI deployment.
- Failure modes that apply
- Evidence needed for sign-off
Evaluation
RecommendedFrom $35,000to $90,000, depending on scope
Per evaluation engagement.
- Suite built on your own tasks and data
- Red-teaming
- Pass criteria agreed in advance
Monitoring
$12,000per month
Continuous monitoring of a production system.
- Regression testing on production traffic
MosaicAGI never sells or resells AI systems. Prices in USD.
Questions people actually ask
- What is the free Deployment Risk Assessment?
- A structured two-hour review of a planned or live AI deployment. It identifies the specific failure modes that apply, what evidence would be needed to sign it off, and which of the current evaluation claims are actually meaningless.
- Why not rely on vendor benchmarks?
- Vendor benchmarks are marketing. An evaluation built on your own tasks and data tests how the system performs, how it fails and whether it can be manipulated.
- What does an evaluation test?
- Task accuracy with proper statistics, prompt injection, jailbreaks and data extraction, disparate performance across the populations you serve, calibration and failure behaviour, and the surrounding system including tool use and retrieval.
- What if the results are unflattering?
- They are reported anyway. Pass criteria are agreed in advance, and results are reported whether or not they are flattering.
- What does an evaluation engagement cost?
- 35,000 to 90,000 dollars per engagement, depending on scope.
- Can a system be monitored after it goes live?
- Yes. Continuous monitoring with regression testing on production traffic is 12,000 dollars a month.
- Does MosaicAGI sell AI systems?
- No. It never sells or resells AI systems, which is what preserves its independence.
- Who is it for?
- Enterprises deploying AI in regulated or high-stakes functions, including financial services, healthcare, insurance, legal and government.
- Why should I trust the AI deployment evidence check?
- The score comes from your answers, worked out in your browser. It does not test the system; it shows which evidence exists. The free Deployment Risk Assessment reviews the deployment itself and names the failure modes that apply to it.
Could you show a board why this AI deployment was signed off?
Answer seventeen questions about a planned or live AI deployment, across performance, manipulation, fairness, the surrounding system and sign-off. See where the evidence is thin and what to ask for first. Runs in your browser. No account, no card, no call.
Open the free toolIt runs in your browser. MosaicAGI never sees your inputs.