Loading rankings…
Loading rankings…
An AI evaluator assesses AI performance. SigRank is an AI evaluator for the operator layer — the only one that evaluates the human, not the model or the output.
Every AI layer has an evaluator. Except the operator layer — until SigRank.
Vals AI evaluates models. Braintrust evaluates outputs. SigRank evaluates operators. The evaluation stack is now complete.
1. Public. Results are visible on a public leaderboard, not locked in a private dashboard. Anyone can see the rankings. This creates accountability — what gets measured publicly gets better publicly.
2. Content-free. Token counts only — never prompt content, never code, never conversation. The four token pillars (input, output, cache-read, cache-write) are sufficient to compute Yield without exposing what the operator is building.
3. Continuous. Evaluation runs from real session telemetry, not one-time test suites. The operator's skill is measured as it evolves, in production, over time.
4. Governed. The MO§ES framework defines the measurement specification, privacy boundaries, and public accountability requirements. This makes SigRank suitable for compliance contexts.
5. Platform-agnostic. Works with Claude, GPT, Gemini, Cursor, Copilot, Windsurf, Codex, and any AI tool that produces token telemetry. Yield measures the human's cascade architecture, not the model's capability.
SigRank evaluates operators using Yield (Υ), a token-cascade efficiency score.
The public evaluation layer for AI operators — ranked by Yield (Υ).
See the top operators ranked by Yield on the public leaderboard.
The four layers of AI evaluation and where the operator layer fits.
The full methodology behind Yield, token telemetry, and operator measurement.