Skip to main content

Controlled, reviewable AI assessment

The AI that marks the answers, not just sets the questions.

Scorafy evaluates open-ended responses against your own rubric, cites the evidence behind every score, and keeps a human reviewer in control. Nothing reaches a respondent until your assessor releases it.

Scorafy doesn't generate the questions. Scorafy evaluates the answers.

Free plan included · No credit card required

Nothing released without sign-off

AI scores first. Your assessor reviews, overrides, and releases.

Every override recorded

Who scored what, when, and why it changed - kept as an audit trail.

Your rubric, your criteria

Up to 10 criteria across 6 levels. Not a generic question bank.

Server-enforced timers

Whole-assessment and per-question limits the browser can't bypass.

AI generated

Sarah Chen

Team Lead - Operations

87/100

Strategic thinking92%
Team culture78%
Delegation61%
Recommendation -Sarah's answers on Q3 and Q7 show she absorbs delivery risk personally. Implement weekly structured handoffs.

The pipeline

How a score gets made.

Every mark on Scorafy travels the same seven stages. You can inspect all of them.

1

Answer

A respondent writes, speaks, records, or uploads. 26 question types, including video and audio with transcription, and PDF or Word documents.

2

Rubric

Your criteria and levels - up to 10 criteria across 6 levels, with per-question weighting and your own grading schema.

3

AI evaluation

Claude by Anthropic reads the actual response against your rubric. Not a score-to-template lookup.

4

Evidence

Every score cites the specific answers behind it, so you can see why - not just what.

5

Human review

Your assessor reviews each result, overrides any score, and adds comments. Overrides are recorded.

6

Release

Respondents see nothing until the assessor releases the result. The release, and who made it, is on the record.

7

Report

A unique, evidence-backed report for each respondent - and one cohort report across the whole group.

Assessment should be fast and fair. Scorafy does the reading and the first draft. Your assessor keeps the judgement.

See the pipeline run on a real answer.

Start free

1 assessment and 10 AI evaluations a month, free. No card.

Built for accountability

Designed for organisations that need controlled, reviewable AI assessment.

Not promises - controls. Each one is live in the product today, and described in full on our security page.

Release gate

Results are held until a human assessor reviews and releases them. Respondents see nothing - not a score, not a report - until sign-off.

Server-enforced timers

Set a limit for the whole assessment, per question, or both. Enforcement happens on the server, not in the browser - an expired timer is rejected even if the page is tampered with. The clock starts only when the respondent presses Start.

Overrides on the record

Assessors can override any AI score, per question, with a comment. The original AI score is kept alongside the human decision.

Audit trail

Who scored what, when it was overridden, and when it was released. The trail is the point.

Retention you control

Set response retention per assessment, from 30 to 365 days. Personal data is removed on your schedule, not ours.

Data residency

EU hosting (Dublin) as standard, with an Australian residency option (Sydney) for organisations that need data onshore.

MFA available. Row-level security on every table. Evaluations run on Claude by Anthropic, and your data is never used to train AI models. Full detail: /security · How scoring works

Proof, honestly

What we can show you, and what we won't claim.

On the record

  • Every score arrives with its evidence. Open the example report and check.
  • The full scoring pipeline is documented, in plain language, at How scoring works.
  • Our security posture is described as controls, not badges, at /security.
  • A public API (v1) with webhooks - including a report.released event that fires only after human sign-off - is documented at /docs/api.
  • If AI capacity runs out mid-cohort, a submitted response is parked and evaluated when capacity returns. A respondent who has submitted is never failed by our meter.

What you won't find here

  • No accuracy percentage. We won't publish one until we can back it with a proper human-agreement benchmark.
  • No certification badges we haven't earned.
  • No stock-photo testimonials. When customers say something we're allowed to quote, we quote them by role, verbatim.

If a claim on this page isn't linked to something you can inspect, tell us and we'll remove it.

Pricing

Pricing that scales with your cohorts

Start free. Scale to cohort and enterprise volume when you're ready.

MonthlyAnnualSave 25%

Free plan available

1 assessment, 10 AI evaluations per month. No card needed.

Start free

Starter

$39/month

For solo assessors and small teams

  • 3 active assessments
  • 100 AI evaluations per month
  • All 26 question types
  • Conditional branching
  • CSV export
Start with Starter
Most popular

Growth

$99/month

For teams assessing multiple cohorts

  • Everything in Starter, plus
  • 15 active assessments
  • 500 AI evaluations per month
  • Custom branding
  • PDF report export
  • Analytics dashboard
  • 10 team members
Start growing

Pro

$249/month

For firms and training providers

  • Everything in Growth, plus
  • Unlimited assessments
  • 2,000 AI evaluations per month
  • Full white-label
  • Weighted rubrics with evidence
  • API access & webhooks
  • Unlimited team members
  • Priority support
Go Pro

For organisations

Business & Enterprise

From $1,000 /month, billed annually

Multi-instructor accounts, data residency options, configurable retention, and dedicated onboarding for teams running assessments at scale.

Talk to us
Evaluation volume scoped to your programme
Cohort-level reporting
Public API and webhooks
Negotiated DPA and named support
Data residency options (Enterprise, quoted)
Dedicated onboarding

All prices in USD. One AI evaluation is one completed respondent assessment, evaluated. Verified schools and non-profits get 40% off Starter, Growth, and Pro.

Questions

Everything you might be wondering

No tool let us add our own criteria and rubrics for assessment. This is exactly what we needed.

Instructional Designer, Global Education NFP

One AI evaluation is one completed respondent assessment, evaluated. It is counted when someone completes your assessment and Scorafy generates their report. Partial or abandoned attempts do not count towards your limit.

No. Respondents click a link and start the assessment. No signup, no friction. If an assessment is timed, respondents see a start screen first - the clock only starts when they press Start - and if they return after finishing they see a completion screen rather than dropping back into the questions.

When your assessor releases them. Scorafy holds every result - score, evidence, and report - until a human reviewer signs off. You can override any score before release, and overrides are kept on the record.

Yes. Set a time limit for the whole assessment, individual questions, or both. Limits are enforced on the server, so they hold even if a respondent's browser is manipulated.

We use Claude by Anthropic. Every report is generated fresh from the respondent's specific answers - not matched to pre-written text buckets. The AI's score is a first draft - your assessor reviews and releases every result. See how scoring works at /how-scoring-works.

Yes. You can add context, frameworks, rubrics, and grading schemas. The AI uses these to generate domain-specific reports that sound like your practice.

No. All plans allow unlimited questions. We recommend 8 - 25 for the best balance of respondent experience and AI analysis quality.

All data is encrypted in transit and at rest, with row-level security isolating every organisation’s data, and MFA available on all accounts. You control response retention per assessment, from 30 to 365 days. We never use your data to train AI models. Full detail on our security page at /security.

See it live

See what your reports could look like.

Answer five questions in the interactive demo and read the evidence-backed report it generates - then imagine it reviewed, signed off, and released by your team. Takes about a minute.

Free plan included · No credit card required