Evaluation & Scoring Rubric
The formal evaluation framework designed to incentivize forensically accurate observations, methodological discipline, and defensive voice in open-source incident research.
Point Distribution (Max 100 Points)
Correct Analysis Flag
Forensically identifying and submitting the correct simulated flag code matching the incident dataset.
Evidence & Source Quality
Notebook records explicitly specify which official reports were used, record the correct access date, and match observed timeline telemetry.
Confidence Level & Defensive Voice
Assigning lower confidence ratings to ambiguous indicators. Formatting the report to strictly segregate verified facts from inferential leaps.
Scoring Penalty Parameters
Unsupported Over-Claims & Early Attribution
Asserting actor attribution as absolute truth based solely on network proxy overlaps, or drawing definite conclusions from ambiguous diagnostic log entries.
Privacy & Ethical Bounds Violations
Including real personal identifiers, victim details, or leaked database entries in your report notes. The lab runs strictly on synthetic, simulated values.
Investigative Standard
This playground is built to test investigative discipline rather than simple trivia. The final grade depends on how cleanly you map variables under pressure, protect local privacy guidelines, and present conclusions without over-attributing operations. Cyber defense starts with slow, verifiable documentation.

