Implementation plan

Phases, resources, risks, and mitigations for the completed ControlSift capstone.

Phases

Phase 0 / Done
Problem framing, MMC option 4 / UN SDG 10 mapping, synthetic benchmark, assurance artifacts, and capstone hub.
Phase 1 / Done
Classical benchmark hardening and v1.1 majority / TF-IDF baselines.
Phase 2 / Done
Hugging Face access, free GPU path, Gemma v1.0 zero-shot and few-shot evaluation.
Phase 3 / Done
Gemma v1.0 QLoRA train/eval; few-shot remains the strongest Gemma rung.
Phase 4 / Done
Error analysis, Failure Lab, external research sources, report, paper, slides, and final version-boundary disclosure.
Phase 5 / Done
Final narrated presentation completed, published with the v1.0-capstone release, and submission package assembled.

Resources used

Risk management

RiskImpactMitigation / outcome
Lexical shortcutArtificially easy benchmarkHardened the generator before sealing v1.1 classical results.
Overclaiming AI performanceFalse success narrativePreserved the QLoRA negative result and low parse-success rate.
Cross-version comparisonMisleading leaderboardClassical v1.1 and Gemma v1.0 are disclosed separately; cross-version scores are descriptive only.
Synthetic-to-real overgeneralizationUnsafe production interpretationExplicit non-production scope, Limitations, Intended Use, and human-review requirements.
Secret exposureCredential compromiseTokens remain in platform secret stores and are excluded from git.
Format missCapstone rejectionRubric-facing report, exactly eight slides, and the final under-five-minute narrated presentation are complete.