TALLYWARE · WAZA

The Art of
Consensus.

Get a panel to judge the same thing and you get a dozen opinions. Waza has them score independently, then shows where they truly agree — and where they don't. Judgment you can stand behind, and explain.

Not about replacing judgment — about making judgment accountable.
A youth hockey coach at the rink boards scoring an evaluation on a tablet while mite-age players run a drill behind the glass
WHICH ONE IS YOU?
THE PROBLEM

Three coaches. Forty kids. Two hours.

“And at the end, we sit in a room and try to agree on who makes the team.”

— every tryout, every season
01

My threes aren't your threes

Same player, same drill — three evaluators, three different numbers. Without a shared scale, you collect noise.

02

Averages hide the truth

A mean of 3 can mean everyone agreed on 3 — or that nobody agreed on anything at all.

03

A feeling and a messy clipboard

When a parent asks why their kid didn't make it, that's all you've got. No trail, no defence.

Waza gives the panel one structured way to score, then surfaces where they agree and where they don't — so the call you make is the one the evidence actually supports.

THE PAYOFF

“Three coaches scored independently. Here's where your kid stood on each skill, where we agreed, and where we saw it differently. We can talk about any of it.

When the answer is transparent, a hard conversation earns respect instead of starting an argument.

THE SHIFT

Same criteria. Same scale. Scored blind.

Everyone scores independently — and no one sees another's numbers until every evaluation is in. You capture what each person actually saw, not who scored first.

M
COACH · MARIA
🔒4
D
COACH · DAN
🔒4
S
COACH · SAM
🔒2
No “Dan gave him a four, so I guess I should too.”
Scores stay sealed until all are submitted — anchoring never gets a chance to start.
THE FULL PICTURE

Score the skill — not the show.

The loudest skater on the ice pulls every eye — and one big impression quietly seeps into every score. A rubric makes each criterion stand on its own, so a dazzling personality can't inflate the skills underneath it, and the quiet, complete player never disappears into someone else's spotlight.

01

Every criterion, on its own

Each skill is scored deliberately, one at a time — so a strong overall impression can't stand in for skills nobody actually watched. A complete read of the whole player, not just the parts that grabbed the room.

02

Dazzle shows up as disagreement

If one evaluator's eye got caught and the others' didn't, that gap surfaces as divergence on a specific criterion — not a tidy average that quietly absorbs it.

03

Unless presence is the point

When charisma or stage command is exactly what you're judging, put it on the rubric. It should count — but only in the column where you decided it counts.

HOW IT WORKS

Four steps, start to consensus.

01

Build the rubric

Define your criteria and 1–5 scales — skating, shooting, game sense. Write your own or clone one from the Marketplace.

02

Create the event

Attach the rubric, add the roster, assign your evaluators. Each organization sees only its own events.

03

Evaluate

Score each criterion by form or by voice — at the boards, in the wings, at the back of the hall, with or without signal.

04

Reach consensus

Waza weights the scores, surfaces agreement, and flags where evaluators diverge — for the whole panel to see.

https://eval.mywaza.com/event/mites-tryout
Skating12345
Shooting12345
Game sense12345
Compete level12345
EVALUATE

Two ways to score. One result.

Half the time there's no Wi-Fi at the venue. Doesn't matter — score how you like, sync when you're back.

⌨️

By form

Tap through the criteria, score 1–5, submit. Installs as a phone app — keep scoring offline and it syncs the moment you reconnect.

  • Criterion by criterion, no ambiguity
  • Drafts saved as you go
  • Offline-capable, installs as a PWA
🎙️

By voice

Press and hold to talk. Transcription runs on your own hardware, Waza extracts structured scores from what you said, and you confirm before anything is recorded.

🎙️
Press & hold to evaluate
“Skating was strong, a four — but the compete level looked like a two today…”
THE RESULTS

Where you agreed — and where you didn't.

Waza weights each evaluator's scores and runs disagreement analysis, so a split panel looks like a split panel — never a tidy mean that hides it. And when one evaluator sees what the others missed, that signal stays in front of you — not flattened into the average.

https://results.mywaza.com/event/mites-tryout
⚑ diverges
Weighted consensus ≈ 4.1 · 1 criterion flagged for review
Dan watched the crossovers. Sam watched the stops. Same player, same skill, different scores — and both were right. That conversation is the part a clipboard can't give you. That's consensus.

Weighted scoring

Not every evaluator counts the same. Consensus reflects how the panel is weighted, not a flat average.

Disagreement analysis

Where the panel splits on a criterion, Waza flags it — so you discuss the real contention instead of burying it.

The right players, not just the best

A crowded tryout overloads any clipboard and any memory — the obvious standouts get kept, the rest blur together. But the lone high score nobody else gave isn't noise to average away, and the player who makes a team isn't always the highest overall — it's the one with the factor you needed. Scored criterion by criterion, that diamond in the rough lives in the data, where memory can't lose it.

Seen enough to want this at your next tryout?
Request access Where it's used ↓
NEW TO THE VOCABULARY?

A few words do all the work here — , , .

Tap a term for the short version — or learn all of the terms used here, each with a plain definition and an example, and the story below reads itself.

Learn the vocabulary →
ACROSS A SEASON

Evaluate the same players — beginning, middle, and end.

“Would you like to evaluate your players beginning, middle, and end of the season? I've got a platform that makes that very, very easy.”

— the question that started it, asked to a mites coach

Same rubric, three checkpoints, one clear line of growth. Development stops being a feeling and becomes something you can show a player, a parent, a board.

PLAYER · #14 · CONSENSUS OVER SEASON
Begin
2.8
Mid
3.4
End
4.1
OUR STORY

The one with the best story wins.

It started at a rink — one coach, one clipboard, and a tryout that came down to memory. Waza exists so the call you make about someone else's kid is one you can defend. [PLACEHOLDER — founder's origin story to come]

Youth hockey coach scoring an evaluation on a tablet at the rink boards
Youth hockey · mites
Score the same skaters all season and show every parent exactly where their kid stands.
Figure-skating judge scoring on a tablet rinkside as a skater performs
Figure skating
Every judge scores blind — no one sees a number until they're all in, so no judge anchors the next.
Clinical assessor scoring a trainee on a tablet at a simulation bedside
Clinical skills
Run it on your own hardware — audio and personal data are processed in-house and never leave the building.
Faculty examiner scoring on a tablet during a thesis defense
Thesis defense
One rubric across the whole committee, with every point of disagreement surfaced instead of buried.
Audition panel scoring independently across a darkened concert hall
Music audition
Weighted consensus that reflects the panel — not a flat average that hides a split vote.
Culinary judge scoring a plated dish on a tablet at the pass
Culinary judging
Score hands-free by voice at the pass and confirm before anything is recorded.
Interview panel scoring a candidate independently on a tablet
Hiring panel
A defensible, auditable record of exactly why each candidate scored the way they did.
Grant reviewer scoring proposals on a tablet at a desk
Grant & research review
Start from a proven rubric in the Marketplace and shape it to your panel — heavier or lighter, your call.
drag / scroll →
MARKETPLACE

One library. Every org inherits it.

Don't build a rubric from scratch — clone one and tweak it to fit. The Marketplace holds shared rubrics you can freely adapt — heavier or lighter — alongside agreements that carry a platform baseline an organization can only strengthen, never weaken.

Enforcement is a ladder. The platform sets a minimum rung for each agreement; an organization can only move it stricter. Clickwrap signing and login gating make sure the right terms are accepted before work begins.

Stricter, always. Weaker, never. An organization can raise the platform's floor but can never drop below it — and if someone tries, the system refuses on its own, every time.
effective = max(platform baseline, org override)
1Offno enforcement
2Notifyremind only
3Gate featuresbaseline ▸ org may raise
4Grace, then gatecountdown
5Block loginstrictest
TRUST & SAFETY

Built to keep data — and people — safe.

The sensitive work stays in-house, and the platform denies by default whenever anything is uncertain.

🔒

On-premises AI

Transcription and analysis run on your own hardware. Audio and personal data are processed in-house and don't leave the building.

HIPAA boundary
🧪

Content & PII safety

Built-in detection screens content and personal information as it enters the platform, before it spreads.

💚

Protection & wellbeing

Guardian controls and wellbeing rules look after the people doing the evaluating, not just the data they produce.

🛡️

Fail-closed access

Every permission check denies on error. The creator can't be the approver. Every action is written to an access audit.

From one panel to many organizations.

Organization isolation

One platform, many orgs — each sees only its own events, rubrics, and people.

Role-based access

Admin, event manager, evaluator, viewer — permissions match the job.

Mobile & offline

Install as a phone app, evaluate offline, sync on reconnect.

BUILT TO BE TRUSTED

Care is in the engineering.

An evaluation platform only earns trust if it's built like one. Every service follows the same disciplined pattern, gets tested against expected outcomes, and ships through a deliberate, reversible release path.

🧱

One build standard

Every service is built to a single template: identity comes from a signed token, never from request parameters; authorization denies on error; and every action is written to a structured audit log.

🧪

Tested like it matters

A scripted acceptance playbook checks role permissions, multi-organization isolation, and full evaluation flows against expected outcomes — and we evaluate Waza with Waza before anyone else does.

exit gate · zero P1 bugs
🚦

Disciplined releases

Changes move through development, validation, and production in reversible steps. Nothing auto-deploys to production — every promotion is a deliberate, approved action, with secrets isolated per environment.

SAST · dependency · secret scan
🔎

Access you can audit

A bidirectional access audit shows exactly who can do what — and separates standard role-based grants from any that bypass them. The creator can't be the approver, and access outside its scope is denied by default.

TALLYWARE · WAZA

Bring consensus
to your evaluations.

Build the rubric. Score independently. See where you agree.
That's Waza.