Codebase Evaluation Framework (CEF) Overview
Purpose: Authoritative truth map protocol assessing software codebase quality across eight independent Diamond Scale dimensions with cited rubrics, evidence-graded findings, paired specialist/adversarial auditors, and referentially anchored architecture diagrams.
| Specification Metadata | Value |
|---|---|
| Framework Version | CEF v0.1.0 |
| Governance Tier | Authoritative Core (Framework Index) |
| Target Roles | Solo Operators, Evaluation Evaluators, Benchmark Harnesses, System Architects |
| Primary Mode | Truth Map (Objective, portable assessment decoupled from product launch roadmaps) |
Quickstart & Evaluation Lifecycle
- Constitutional Binding: Review
CONSTITUTION.mdfor non-negotiable evaluation rules and evidence grades. - Quality Grading Scale: Consult
DIAMOND_SCALE.mdfor the 1–5 scoring ladder across the eight quality dimensions. - Preflight Inventory: Execute Wave 0 mechanical inventory and AST discovery per
WAVE_PLAN.md. - Lens Specialist & Adversarial Passes: Dispatch paired evaluator and auditor agents from
prompts/. - Schema Validation: Verify all findings against
schemas/finding.schema.json. - Integrator Synthesis: Assemble validated findings and multi-axis scorecards into
HANDOFF_SCHEMA.md.
Framework Groupings & Architecture
The Codebase Evaluation Framework is organized into four clean, functional tiers:
1. Framework Core & Governance
| Document | Focus Area & Purpose |
|---|---|
CONSTITUTION.md |
Non-negotiable constitutional invariants, evidence grades (E0–E3), budget density rules, and zero-fix discipline. |
DIAMOND_SCALE.md |
Multi-axis Diamond Scale grading model across all 8 architectural quality dimensions. |
WAVE_PLAN.md |
Multi-wave phased execution plan (Preflight → Foundations → Runtime → Craftsmanship → Synthesis). |
LENSES.md |
Full lens catalog, density classes (D-HIGH exhaustive vs D-LOW Top-N), and ownership boundaries. |
DIAGRAM_CONTRACT.md |
Mandatory architectural illumination, Mermaid diagram types, and concrete symbol anchoring rules. |
OPERATOR.md |
Single-page operator runbook, preflight checklist, and execution guidelines. |
KICKOFF_PROMPT.md |
Standardized, deterministic solo-operator kickoff prompt for A/B comparable evaluations. |
HANDOFF_SCHEMA.md |
Downstream objectification and remediation handoff contract for process systems. |
EXTENSIONS.md |
Framework extension architecture for adding custom lenses, compliance overlays, and language packs. |
adapters/go.md |
Go language-specific analysis adapter, tooling heuristics, and AST scanner contracts. |
2. Evaluation Lenses Matrix (Rubrics & Agent Prompt Pairs)
Each evaluation dimension is governed by an authoritative rubric and paired with a specialist evaluator and an adversarial auditor:
Note
All evaluation prompts share common operational directives defined in prompts/_SHARED_PREAMBLE.md.
Relationship to other quality artifacts (this repo)
If present in a host project, these are optional sensors, not CEF itself:
- File-level vetting matrices / VDS profiles under
docs/quality/ - Project policies, checklists, linters, AST tools
CEF must remain copyable to a non-ZQK tree and still make sense.
Versioning
- cef_version:
0.1.0 - Breaking changes to finding schema or diamond axes require a minor/major bump and a short changelog entry in this README.
Changelog
| Version | Date | Notes |
|---|---|---|
| 0.1.0 | 2026-08-13 | Initial framework: constitution, diamond scale, lenses, rubrics, specialist/adversarial prompts, wave plan, handoff, extensions, schemas |