Athena/ Learning Log
    Week of August 2, 2026 · updated continuously

    How Athena learned this week.

    Every reviewer correction on the platform — anonymized, aggregated, public. This is the proof that closed-loop learning is real, not a slide. Nobody else in CMMC publishes this.

    Reviewer corrections learned
    247
    +18% vs last week
    Acceptance rate
    91.0%
    +6.2 pts vs v3.3
    Avg change score
    0.07
    lower is better · −0.04
    Controls in active corpus
    320
    NIST 800-171 full coverage
    Model leaderboard

    Accuracy by model_version

    Version
    Status
    Acceptance
    Δ score
    Corpus
    Trend
    athena-extract-v4-learnings
    Reviewer-correction injection live; MFA scoping fixed.
    active
    91.0%
    0.07
    12,847
    +6.2%
    athena-extract-v3.3
    Tightened 3.5.x mapping; removed boilerplate hallucinations.
    shipped
    85.7%
    0.11
    11,204
    +3.1%
    athena-extract-v3.2
    Source-line citation made mandatory.
    shipped
    83.1%
    0.14
    9,881
    +1.8%
    athena-extract-v3.1
    Regression on AU.x — rolled forward via v3.2 hotfix.
    shipped
    81.6%
    0.16
    8,402
    -0.4%
    athena-extract-v5-canary
    Test-only. Multi-evidence reconciliation enabled.
    canary
    93.4%
    0.05
    1,207
    +2.6%

    How we measure: Every batch ingested under a model_version is replayed through _shared/gap-diff.ts against the stored human_output in v_athena_learning_dataset. Acceptance = reviewer kept the AI output unchanged. Δ score = average token-level distance from human truth (lower is better).

    Closed-loop corrections

    What Athena got wrong — and learned

    3.5.3scope
    2d ago

    MFA enforced for remote access only

    MFA enforced for all privileged accounts, including local console

    3.1.20mapping
    3d ago

    Mapped to network monitoring controls (incorrect)

    Mapped to external system use — vendor MSAs are the right evidence

    3.13.11evidence
    4d ago

    FIPS-validated cryptography assumed from product name

    Requires explicit CMVP certificate # — pulled from vendor PDF

    3.6.1language
    5d ago

    IRP narrative flagged as "weak language"

    Pattern matched to NIST 800-61 template — accepted as strong

    3.4.2scope
    6d ago

    Baseline config applied only to servers

    Baseline must include endpoints + cloud workloads

    3.14.6evidence
    7d ago

    Single SIEM screenshot accepted as monitoring evidence

    Requires alert-rule export + 30-day retention proof

    Radical transparency, by design.

    This page is generated from anonymized rollups of gap_review_workflow_metrics and v_athena_learning_dataset. No customer names, no document contents, no identifying metadata ever appears here. The point is to show that Athena learns — and to let you compare us against any other CMMC tool that claims AI. None of them publish anything like this.

    The next snapshot will be more accurate than the last.

    Start an assessment today and your corrections shape the model your competitors run on tomorrow.