
Grade your product before your users do
Grade scores your product experience on 100s of user scenarios, calibrated to what real users say and do.
Catch regressions before you ship
Review launches with graded examples
Benchmark competitors on key scenarios
From the team at Voicepanel




Use cases
Where teams use Grade
Regression checks
Catch failure modes on priority scenarios before they ship.
Product launch reviews
Walk through graded examples and decide if you're ready to ship.
Competitive benchmarking
See how your product stacks up on the scenarios that matter most.
Agent feedback loops
A shared bar for what agents should optimize toward.
How it works
From user research to autograder
01
Test scenarios with users
We recruit your target users and record them using your product on the scenarios that matter. What they say and do is captured in detail, and handled for you in a matter of days.
02
Decide what good looks like
We help you build a rubric for how well user needs are met, not how good the AI sounds. It applies across hundreds of real scenarios, rooted in what your users actually say and do.
03
Get your autograder
We deliver an autograder calibrated to your users. Score 100s or 1000s of scenarios automatically, and test new product ideas before they reach the first user.
Why now
User research for AI-native product teams
Ad-hoc user research can't keep up with how modern teams ship. Grade gives you a score you can re-run on every change.
| Ad-hoc user research | Grade | |
|---|---|---|
| Cadence | Tied to big milestones | Runs continuously |
| Cost to repeat | High cost per study | Low cost per score |
| Where insights live | Reports that get filed away | Embedded in your dev process |
| Scenario coverage | Deep on a few studies | Scales to 100s of scenarios |
| Failure modes | Hard to cover every path | User-calibrated, applied broadly |
Join the waitlist
We're onboarding teams in waves. Tell us a bit about what you're building, and we'll reach out to schedule a call.

By Voicepanel