Review comment closure fraction
review_comment_closure_fraction
Closed review comments divided by total comments
Scores if within ±0.3% of the reference value.
Calculates source-bound visual systems review metrics from a deterministic SSC-13 task-owned source pack. The template combines review comments, revised layout, device schedule, calculation trace, criteria matrix, and repair response.
with-tool: The model is given an executable Python calculator script.
One template produces many comparable benchmark tasks while keeping the scoring contract fixed.
01
The reusable contract shown on this page.
02
An archetype and site context are sampled.
03
Inputs may be hidden at harder tiers.
04
The model responds with the declared outputs.
Inputs the model receives, and the outputs it is scored on.
16 inputs
Included directly in every task prompt.
Closed review comments
closed_review_comments
REVIEW-13-COMMENTS-08 closed review comments
Total review comments
total_review_comments
REVIEW-13-COMMENTS-08 total review comments
Updated affected checks
updated_affected_checks
CALC-13-TRACE-08 updated affected checks
Required affected checks
required_affected_checks
CRIT-13-MATRIX-08 required affected checks
Revised minimum
revised_minimum_lux
LAYOUT-13-REV-08 revised minimum lighting
Required minimum
required_minimum_lux
CRIT-13-MATRIX-08 required minimum lighting
Cctv horizontal pixels
cctv_horizontal_pixels
DEVICE-13-SCHED-08 revised CCTV horizontal pixels
Revised target width
revised_target_width_m
LAYOUT-13-REV-08 revised CCTV target width
Required ppm
required_ppm
CRIT-13-MATRIX-08 required CCTV PPM
Revised network load
revised_network_load_mbps
DEVICE-13-SCHED-08 revised network load
Network capacity
network_capacity_mbps
DEVICE-13-SCHED-08 network capacity
Revised poe load
revised_poe_load_w
DEVICE-13-SCHED-08 revised PoE load
Poe budget
poe_budget_w
DEVICE-13-SCHED-08 PoE budget
Unresolved conflict
unresolved_conflict_count
RESPONSE-13-REPAIR-08 unresolved conflicts
Completed repair memo sections
completed_repair_memo_sections
RESPONSE-13-REPAIR-08 completed memo sections
Required repair memo sections
required_repair_memo_sections
RESPONSE-13-REPAIR-08 required memo sections
10 outputs
review_comment_closure_fraction
Closed review comments divided by total comments
Scores if within ±0.3% of the reference value.
affected_check_update_fraction
Updated affected checks divided by required affected checks
Scores if within ±0.3% of the reference value.
lighting_minimum_margin_lux
Revised minimum lighting minus required minimum lighting
Scores if within ±0.3% of the reference value.
revised_cctv_pixels_per_m
Revised CCTV pixels per metre
Scores if within ±0.3% of the reference value.
cctv_ppm_margin
Revised CCTV PPM minus required PPM
Scores if within ±0.3% of the reference value.
network_headroom_mbps
Network capacity minus revised network load
Scores if within ±0.3% of the reference value.
poe_headroom_w
PoE budget minus revised PoE load
Scores if within ±0.3% of the reference value.
unresolved_conflict_count
Unresolved source conflict count
Scores if within ±0.1% of the reference value.
repair_memo_completeness_fraction
Completed repair memo sections divided by required sections
Scores if within ±0.3% of the reference value.
overall_pass_score
1.0 when review repair checks pass
Scores if within ±0.1% of the reference value.
Each template is sampled at three tiers. Harder tiers may hide inputs, forcing the model to infer them from the scenario description.
For this template, difficulty scales through parameter and scenario ranges rather than hidden information.
The exact instruction and parameter contract used to generate this task, pinned to the published library source.
/workspace
Teal lines show Jinja input conditions, not task visibility policy. A line renders only when that input or tool is visible.
1You are an electrical visual systems reviewer checking a task-owned synthetic SSC-13 review and repair package.2 3Use only the task-owned synthetic source pack values shown below for numeric grading. External review-management, lighting, CCTV, ITS, and network tools shape the workflow context only; they are not extra data sources for this instance.4 5## Scene6 7- Product: `SSC-13-LH-08`8- Review comments: `REVIEW-13-COMMENTS-08`9- Revised layout: `LAYOUT-13-REV-08`10- Revised device schedule: `DEVICE-13-SCHED-08`11- Calculation trace: `CALC-13-TRACE-08`12- Criteria matrix: `CRIT-13-MATRIX-08`13- Repair response: `RESPONSE-13-REPAIR-08`14 15All checks use the same comment register, revised layout, device schedule, affected calculation trace, criteria matrix, and response ledger.16 17## Source Values18 19| Item | Value |20|------|-------|21| Closed review comments | {{ closed_review_comments }} of {{ total_review_comments }} |22| Updated affected checks | {{ updated_affected_checks }} of {{ required_affected_checks }} |23| Revised minimum lighting | {{ revised_minimum_lux }} lux |24| Required minimum lighting | {{ required_minimum_lux }} lux |25| CCTV pixels and revised target width | {{ cctv_horizontal_pixels }} px / {{ revised_target_width_m }} m |26| Required PPM | {{ required_ppm }} |27| Revised network load and capacity | {{ revised_network_load_mbps }} Mbps / {{ network_capacity_mbps }} Mbps |28| Revised PoE load and budget | {{ revised_poe_load_w }} W / {{ poe_budget_w }} W |29| Unresolved conflicts | {{ unresolved_conflict_count }} |30| Repair memo sections | {{ completed_repair_memo_sections }} of {{ required_repair_memo_sections }} |31 32## Output Format33 34Write a compact visual systems review response to `/workspace/output.md`. Include a source-boundary statement that this is a task-owned synthetic source pack. Do not claim authority approval, accepted project evidence, software export validity, full standards compliance, executable real source-pack parsing, generated benchmark readiness, or benchmark readiness.35 36Include a fenced JSON block with exactly these numeric keys:37 38```json39{40 "review_comment_closure_fraction": <numeric_value>,41 "affected_check_update_fraction": <numeric_value>,42 "lighting_minimum_margin_lux": <numeric_value>,43 "revised_cctv_pixels_per_m": <numeric_value>,44 "cctv_ppm_margin": <numeric_value>,45 "network_headroom_mbps": <numeric_value>,46 "poe_headroom_w": <numeric_value>,47 "unresolved_conflict_count": <numeric_value>,48 "repair_memo_completeness_fraction": <numeric_value>,49 "overall_pass_score": <numeric_value>50}51```52 Each generated task is drawn from one of these realistic scenario bands.
Site contexts ground each scenario in a real locale the model can use to infer hidden values.
ssc13_visual_repair_case
Visual systems review and repair case
ssc13-visual-repair-ssc13-visual-repair-case-preview — hard difficulty, all inputs given.
Visual systems review and repair case. ssc13-visual-repair. Required outputs: review_comment_closure_fraction, affected_check_update_fraction, lighting_minimum_margin_lux, revised_cctv_pixels_per_m, cctv_ppm_margin, network_headroom_mbps
Scenario context and visible inputs.
Executable tool: visual-systems-review-repair-package_calc.py
Inputs withheld at this difficulty.
Nothing. All inputs are supplied.
The scored JSON answer schema.
{
"review_comment_closure_fraction": <number>,
"affected_check_update_fraction": <number>,
"lighting_minimum_margin_lux": <number>,
"revised_cctv_pixels_per_m": <number>,
"cctv_ppm_margin": <number>,
"network_headroom_mbps": <number>,
"poe_headroom_w": <number>,
"unresolved_conflict_count": <number>,
"repair_memo_completeness_fraction": <number>,
"overall_pass_score": <number>
}