Braintrust Pricing
Braintrust pricing tells you exactly what scoring your agent costs. It never tells your customer what the agent was worth.
Braintrust's real numbers, verified straight from its own pricing page: Starter is free with unlimited users and ten dollars of model credit, Pro is $249 a month flat with five gigabytes of data and fifty thousand scores included, and Enterprise has no published price at all, custom quoted through sales only. Past the included allowance, data overage runs $3 to $4 per gigabyte and score overage runs $1.50 to $2.50 per thousand depending on tier, and the $249 Pro price currently doubles as a promotional model credit balance, a detail most third party writeups about Braintrust get wrong or skip. This page verifies every number, works the real overage math independently, and then covers the one question no tier, free or Enterprise, was built to answer: what did the agent actually do for the customer paying for it.
Verified against Braintrust's own pricing page
What Braintrust actually costs
Braintrust's pricing runs three tiers, and every number below comes from Braintrust's own current pricing page rather than a list price a third party writeup forgot to update. Starter costs $0 a month, includes ten dollars of model credit, covers one gigabyte of processed data and ten thousand scores, holds fourteen days of retention, and still ships with unlimited users, projects, datasets, playgrounds, and experiments, no seat cap even on the free tier. Pro costs $249 a month flat, no per seat charge as a workspace adds engineers, raises the included allowance to five gigabytes of data and fifty thousand scores, stretches retention to thirty days with paid extended storage past that, and adds custom charts, environments, priority support, basic role based access control, a loop agent for automated evaluation runs, playground annotations, and unlimited human review scores.
Enterprise is where Braintrust's own pricing page stops publishing numbers entirely, listing the tier only as custom pricing, quoted through sales, with no minimum contract value or usage based starting point shown anywhere on the page. What the page does document is the feature set added at that tier: full role based access control instead of Pro's basic version, custom retention rules set per project rather than one account wide window, S3 export for a company's own long term storage, on premises or hosted deployment for teams that cannot send evaluation data to a third party, a guaranteed SLA, and a signed BAA for HIPAA covered workloads. Several third party pricing writeups list a specific Enterprise dollar figure anyway; none of those numbers appear on Braintrust's own site, so treat any exact figure you see elsewhere as unverified until Braintrust's own sales team quotes one against real usage.
Three tiers, two published prices
The three Braintrust tiers, in plain numbers
Straight from Braintrust's own pricing page, not a third party estimate.
Starter: $0 a month
Ten dollars of model credit, one gigabyte of data, ten thousand scores, fourteen day retention. Unlimited users and projects even at this tier.
Pro: $249 a month
Five gigabytes of data and fifty thousand scores included, thirty day retention, custom charts, and basic role based access control. No per seat charge.
Enterprise: custom pricing
No published number. Full RBAC, on premises deployment, an SLA, and a HIPAA BAA, quoted through sales against real usage only.
The number every paid tier bills against
How a 'score' actually gets counted, and what overage costs
Every tier above bills against the same two meters, processed data and scores, and Braintrust's own documentation defines a score precisely: it is any scored output an evaluation run produces, whether an LLM acts as the judge, an autoeval built into the platform runs it, or a team's own custom code scorer grades it. A human reviewer manually grading a trace during review generates a score the same way, which is why Pro lists unlimited human review scores as its own named feature rather than folding that activity into the fifty thousand score allowance. Running one autoeval and one custom scorer across the same dataset produces two scores per row, not one, the detail that most often explains why score usage climbs faster than a team's row count would suggest on its own.
Past the included allowance, Braintrust does not cut a workspace off; it keeps ingesting and bills the overage per unit rather than in blocks. On Starter, data past the first gigabyte costs $4 per additional gigabyte and scores past the first ten thousand cost $2.50 per thousand. On Pro, the same two meters drop to $3 per additional gigabyte and $1.50 per thousand scores, plus fifty cents per gigabyte per month for data retained past the included thirty days. Take a Pro workspace processing eight gigabytes of data and running eighty thousand scores in a month: the base $249 covers the first five gigabytes and first fifty thousand scores, the remaining three gigabytes add $9, and the remaining thirty thousand scores add $45, landing the real platform bill at $303, before counting whatever model credit the team burns past the included $249 running evaluations through Braintrust's own proxy.
The line every third party writeup gets confused about
Why $249 shows up twice on the Pro plan
Braintrust's own pricing page lists $249 as both the subscription fee and the current promotional credit balance. They are not the same $249.
The subscription fee is flat
The $249 a month platform charge does not scale with seats. A five person team and a fifty person team on Pro both pay the same base price before data or score overage.
The credit balance is a promotion
Braintrust's normal included model credit on Pro is $100. A current promotion bumped it to $249 to match the subscription price, which is why the same number appears twice on the pricing page.
Model usage past the credit bills separately
Once a workspace burns through its credit calling a model via Braintrust's own proxy, further usage bills at the provider's token rate, on top of the flat $249. Using a team's own API key skips this meter entirely.
What the score grid doesn't show
Braintrust's pricing tells you what evaluating your agent cost. It tells your customer nothing
Every number on Braintrust's pricing page, the score count climbing toward a plan's allowance, the data overage rate, the thirty day retention window on Pro, describes the same thing: what evaluating and scoring your own agent costs you and your engineering team. That holds true at every tier, including Enterprise, since even project level role based access control and a signed BAA are still built for your own account administrators, not for an outside customer. Unlimited users on Starter and up means unlimited seats for your own company, and even Braintrust's own multi tenancy inside a workspace, unlimited projects and datasets, still assumes every project belongs to your team, not to a paying customer sitting outside it entirely.
If you sell an AI agent inside a product other companies pay for, that gap shows up the moment a customer asks what the agent actually did for their account this month, a question no score grid or retention window answers. AiAgRe's Node SDK reads the same kind of agent events Braintrust already scores, through the same LangChain, LlamaIndex, and CrewAI integrations, and tags each one with an organization identity and a customer identity at the point of ingestion. That turns into white-labeled dashboard components your customer sees inside your own product, showing deflection rate and cost per resolution scoped so tightly that one customer's numbers never reach another customer's view. It does not replace Braintrust's own evaluation and scoring work, and most teams run both: Braintrust for whether the agent's outputs pass, AiAgRe for what that passing activity was worth to the person paying for it.
FAQs
Braintrust pricing: frequently asked questions
Common questions from teams pricing Braintrust against their own usage before they commit to a tier, or before they add a second layer on top.
How much does Braintrust actually cost?
Starter is free and includes ten dollars of model credit, one gigabyte of processed data, ten thousand scores, and fourteen day retention, with unlimited users, projects, datasets, playgrounds, and experiments even at this tier. Pro runs $249 a month flat, no per seat charge, and raises the included allowance to five gigabytes of processed data and fifty thousand scores with thirty day retention, plus custom charts, environments, priority support, basic role based access control, and unlimited human review scores. Enterprise has no published number at all: Braintrust's own pricing page lists it only as custom, quoted through sales, and adds full role based access control, custom retention rules per project, on premises or hosted deployment, a guaranteed SLA, and a signed BAA for HIPAA workloads. Every figure here comes straight from Braintrust's own pricing page as of this page's publish date, worth reconfirming there since usage based rates change.
What counts as a 'score' in Braintrust's pricing?
Braintrust's own documentation defines a score as any scored output an evaluation run produces, whether it comes from an LLM acting as a judge, an autoeval built into the platform, or a custom code scorer a team writes itself. A human reviewer manually grading a trace during review also generates a score the same way, which is why Pro includes unlimited human review scores as a named feature rather than folding them into the fifty thousand score allowance. A team running one autoeval and one custom scorer across the same dataset generates two scores per row, not one, which is the detail that most often explains why score usage climbs faster than the row count in a dataset would suggest.
What happens once a team goes over the included data or scores?
Braintrust keeps ingesting past the included allowance on every tier rather than cutting a workspace off, and bills the overage per unit rather than in blocks. On Starter, processed data past the first gigabyte costs $4 per additional gigabyte and scores past the first ten thousand cost $2.50 per thousand. On Pro, the same two meters drop to $3 per additional gigabyte and $1.50 per thousand scores, and retained data held past the included thirty days costs an extra fifty cents per gigabyte per month for as long as it stays stored. Enterprise sets its own custom limits and rates per contract, so none of the per unit numbers above carry over once a team moves onto that tier.
Is the $249 Pro price really flat, or does model usage cost extra?
Both things are true at once, and it is the single most confusing line on Braintrust's own pricing page. The $249 a month platform fee is flat with no per seat charge, and it also currently doubles as the account's starting model credit balance, since Braintrust is running a promotion that bumped the included credit from its normal $100 up to $249 to match the subscription price. Once a workspace burns through that credit calling a model through Braintrust's own proxy, further model usage bills at the underlying provider's token rate, on top of the flat $249, the same way a phone plan's included minutes run out and start costing extra. A team routing model calls through its own API key instead of Braintrust's proxy skips that meter entirely and only pays the $249 platform fee plus data and score overage.
What's actually included in Enterprise, and does it have a real published price?
No, and every third party writeup that quotes a specific Enterprise dollar figure for Braintrust is guessing at a number Braintrust itself has never published. Braintrust's own pricing page lists Enterprise only as custom pricing, reachable by contacting sales rather than checking a rate card, with no minimum contract value, no per seat estimate, and no usage based starting point shown anywhere. What is documented is the feature set: full role based access control instead of Pro's basic version, custom retention rules set per project rather than one account wide window, S3 export for a company's own long term storage, on premises or hosted deployment for teams that cannot send evaluation data to a third party at all, a guaranteed uptime SLA, and a signed BAA for HIPAA covered workloads. Treat any specific Enterprise dollar figure you see on a third party page as unverified until Braintrust's own sales team quotes a number against real usage.
Is Braintrust cheaper than LangSmith or Langfuse?
At the free tier, Braintrust's Starter plan is more generous on seats, unlimited users against LangSmith's five thousand base trace developer plan and Langfuse's two user cap on Hobby, though thinner on retention at fourteen days against Langfuse's thirty. Past the free tier the three get harder to line up directly, since a Braintrust score, a Langfuse unit, and a LangSmith trace are three different meters billed at three different rates. Braintrust also treats evaluation, the scoring, the regression testing, the dataset management, as its main product with tracing built around it, while LangSmith and Langfuse both start from tracing and add evaluation on top, which matters more than the sticker price for most teams choosing between them. The exact headcount where Braintrust's flat fee beats LangSmith's per-seat pricing is worked out on the LangSmith vs Braintrust page, and the fuller side by side against LangSmith, Langfuse, Arize AI, Galileo AI, and Helicone, verified pricing for all five, is on the Braintrust alternatives page.
Does any Braintrust plan show pricing or ROI data to my own customers?
No, on every tier including Enterprise. Every screen Braintrust renders, the score grid inside an experiment, the dataset browser, the playground where a prompt gets tested against real inputs, lives inside a workspace scoped to your own account and built for your own engineers or evaluators to read. Unlimited users on every tier means unlimited seats for your own team, not a scoped view for an outside customer, and even Enterprise's project level role based access control still assumes every project belongs to your company. If you sell an AI agent inside a product other companies pay for, your own customer wants their own deflection rate and cost per resolution, isolated to their own traffic only, which sits outside what any evaluation tool, Braintrust included, was built to show.
Can I run AiAgRe alongside Braintrust?
Yes, and that is the normal setup rather than a workaround. AiAgRe's Node SDK reads the same kind of underlying agent events Braintrust already scores, through the same LangChain, LlamaIndex, or CrewAI integrations, and tags each one with an organization identity and a customer identity at the point of ingestion. That turns into white-labeled dashboard components your own customer sees inside your product, scoped so tightly that one customer's numbers never reach another customer's view. It does not replace the pre-launch evaluation work Braintrust does well: most teams keep Braintrust for scoring an agent before it ships, and add AiAgRe for proving what it does once real customers are using it.
What would a mid-size Pro workload actually cost?
Take a workspace processing eight gigabytes of data and running eighty thousand scores in a month, a reasonable load for a team evaluating an agent across a few dozen datasets with autoevals running on every row. Pro's base price of $249 covers the first five gigabytes of data and the first fifty thousand scores. The remaining three gigabytes of data cost $3 each, adding $9, and the remaining thirty thousand scores cost $1.50 per thousand, adding $45. The real platform bill for that month lands at $303, not the $249 sticker price on Braintrust's own pricing page, and that number still excludes whatever model credit the team burns through Braintrust's proxy beyond the included $249, which depends entirely on which models the evaluation runs call and how many tokens each one uses.
Do qualifying startups really get Pro free?
Braintrust's own pricing page states that qualifying startups get six to twelve months of Pro free, but does not publish the qualification criteria anywhere on the page itself, no funding stage, no company age, no revenue ceiling. That puts it in the same category as most startup credit programs run by usage based SaaS platforms: real, but gated behind an application a company has to submit and Braintrust's own sales team has to approve, rather than a self-serve discount applied automatically at signup. A team planning around this offer should confirm eligibility directly with Braintrust before budgeting a full year of Pro at $0, since the six to twelve month range itself suggests the exact length varies by company rather than being fixed.
Already paying to score your agent? Now show your customer what it did.
Request access and we'll walk through how AiAgRe's embed tokens map onto the evaluation data Braintrust already produces.
