Controlled, reviewable AI assessment
The AI that marks the answers, not just sets the questions.
Scorafy evaluates open-ended responses against your own rubric, cites the evidence behind every score, and keeps a human reviewer in control. Nothing reaches a respondent until your assessor releases it.
Scorafy doesn't generate the questions. Scorafy evaluates the answers.
Free plan included · No credit card required
Nothing released without sign-off
AI scores first. Your assessor reviews, overrides, and releases.
Every override recorded
Who scored what, when, and why it changed - kept as an audit trail.
Your rubric, your criteria
Up to 10 criteria across 6 levels. Not a generic question bank.
Server-enforced timers
Whole-assessment and per-question limits the browser can't bypass.
Sarah Chen
Team Lead - Operations
87/100
The pipeline
How a score gets made.
Every mark on Scorafy travels the same seven stages. You can inspect all of them.
Answer
A respondent writes, speaks, records, or uploads. 26 question types, including video and audio with transcription, and PDF or Word documents.
Rubric
Your criteria and levels - up to 10 criteria across 6 levels, with per-question weighting and your own grading schema.
AI evaluation
Claude by Anthropic reads the actual response against your rubric. Not a score-to-template lookup.
Evidence
Every score cites the specific answers behind it, so you can see why - not just what.
Human review
Your assessor reviews each result, overrides any score, and adds comments. Overrides are recorded.
Release
Respondents see nothing until the assessor releases the result. The release, and who made it, is on the record.
Report
A unique, evidence-backed report for each respondent - and one cohort report across the whole group.
Assessment should be fast and fair. Scorafy does the reading and the first draft. Your assessor keeps the judgement.
See the pipeline run on a real answer.
Start free1 assessment and 10 AI evaluations a month, free. No card.
Built for accountability
Designed for organisations that need controlled, reviewable AI assessment.
Not promises - controls. Each one is live in the product today, and described in full on our security page.
Release gate
Results are held until a human assessor reviews and releases them. Respondents see nothing - not a score, not a report - until sign-off.
Server-enforced timers
Set a limit for the whole assessment, per question, or both. Enforcement happens on the server, not in the browser - an expired timer is rejected even if the page is tampered with. The clock starts only when the respondent presses Start.
Overrides on the record
Assessors can override any AI score, per question, with a comment. The original AI score is kept alongside the human decision.
Audit trail
Who scored what, when it was overridden, and when it was released. The trail is the point.
Retention you control
Set response retention per assessment, from 30 to 365 days. Personal data is removed on your schedule, not ours.
Data residency
EU hosting (Dublin) as standard, with an Australian residency option (Sydney) for organisations that need data onshore.
MFA available. Row-level security on every table. Evaluations run on Claude by Anthropic, and your data is never used to train AI models. Full detail: /security · How scoring works
Proof, honestly
What we can show you, and what we won't claim.
On the record
- Every score arrives with its evidence. Open the example report and check.
- The full scoring pipeline is documented, in plain language, at How scoring works.
- Our security posture is described as controls, not badges, at /security.
- A public API (v1) with webhooks - including a
report.releasedevent that fires only after human sign-off - is documented at /docs/api. - If AI capacity runs out mid-cohort, a submitted response is parked and evaluated when capacity returns. A respondent who has submitted is never failed by our meter.
What you won't find here
- No accuracy percentage. We won't publish one until we can back it with a proper human-agreement benchmark.
- No certification badges we haven't earned.
- No stock-photo testimonials. When customers say something we're allowed to quote, we quote them by role, verbatim.
If a claim on this page isn't linked to something you can inspect, tell us and we'll remove it.
Pricing
Pricing that scales with your cohorts
Start free. Scale to cohort and enterprise volume when you're ready.
Free plan available
1 assessment, 10 AI evaluations per month. No card needed.
Starter
For solo assessors and small teams
- 3 active assessments
- 100 AI evaluations per month
- All 26 question types
- Conditional branching
- CSV export
Growth
For teams assessing multiple cohorts
- Everything in Starter, plus
- 15 active assessments
- 500 AI evaluations per month
- Custom branding
- PDF report export
- Analytics dashboard
- 10 team members
Pro
For firms and training providers
- Everything in Growth, plus
- Unlimited assessments
- 2,000 AI evaluations per month
- Full white-label
- Weighted rubrics with evidence
- API access & webhooks
- Unlimited team members
- Priority support
For organisations
Business & Enterprise
From $1,000 /month, billed annually
Multi-instructor accounts, data residency options, configurable retention, and dedicated onboarding for teams running assessments at scale.
Talk to usAll prices in USD. One AI evaluation is one completed respondent assessment, evaluated. Verified schools and non-profits get 40% off Starter, Growth, and Pro.
Questions
Everything you might be wondering
“No tool let us add our own criteria and rubrics for assessment. This is exactly what we needed.
Instructional Designer, Global Education NFP
One AI evaluation is one completed respondent assessment, evaluated. It is counted when someone completes your assessment and Scorafy generates their report. Partial or abandoned attempts do not count towards your limit.
No. Respondents click a link and start the assessment. No signup, no friction. If an assessment is timed, respondents see a start screen first - the clock only starts when they press Start - and if they return after finishing they see a completion screen rather than dropping back into the questions.
When your assessor releases them. Scorafy holds every result - score, evidence, and report - until a human reviewer signs off. You can override any score before release, and overrides are kept on the record.
Yes. Set a time limit for the whole assessment, individual questions, or both. Limits are enforced on the server, so they hold even if a respondent's browser is manipulated.
We use Claude by Anthropic. Every report is generated fresh from the respondent's specific answers - not matched to pre-written text buckets. The AI's score is a first draft - your assessor reviews and releases every result. See how scoring works at /how-scoring-works.
Yes. You can add context, frameworks, rubrics, and grading schemas. The AI uses these to generate domain-specific reports that sound like your practice.
No. All plans allow unlimited questions. We recommend 8 - 25 for the best balance of respondent experience and AI analysis quality.
All data is encrypted in transit and at rest, with row-level security isolating every organisation’s data, and MFA available on all accounts. You control response retention per assessment, from 30 to 365 days. We never use your data to train AI models. Full detail on our security page at /security.
See it live
See what your reports could look like.
Answer five questions in the interactive demo and read the evidence-backed report it generates - then imagine it reviewed, signed off, and released by your team. Takes about a minute.
Free plan included · No credit card required