The completed LFT-019 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Graph literacy, while 1 check remained unresolved.
The completed LFT-068 synthetic field test reached 10/10 after one failure-only correction: 5 of five static checks passed for Map Literacy, while 0 checks remained unresolved.
This completed synthetic Refactoring field test asked the session to refactor a fragile legacy script, preserved an actual five-row legacy script refactor diff, and derived 0/10 then 2/10 from task-specific semantic checks after one failure-only correction.
The completed WFT-048 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Survey Analysis, while 1 check remained unresolved.
This completed synthetic Container Inspection field test asked the session to check a container image for seeded secrets and risky defaults, preserved an actual five-row container secret and defaults audit, and derived 6/10 then 10/10 from task-specific semantic checks after one failure-only correction.
The completed LFT-033 synthetic field test reached 10/10 after one failure-only correction: 5 of five static checks passed for Map reasoning, while 0 checks remained unresolved.
This completed synthetic Version Control field test asked the session to resolve a complex merge conflict correctly, preserved an actual five-row three-way merge resolution record, and derived 2/10 then 8/10 from task-specific semantic checks after one failure-only correction.
The completed WFT-034 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Document Classification, while 1 check remained unresolved.
The completed WFT-055 synthetic field test finished at 4/10 and was not recommended: only two of five Obligation Mapping checks passed after the permitted correction.
The completed WFT-013 synthetic field test finished at 4/10 and was not recommended: only two of five Expense Compliance checks passed after the permitted correction.
This completed synthetic Attachment Safety field test asked the session to triage a suspicious email attachment safely, preserved an actual five-row suspicious attachment static triage record, and derived 6/10 then 10/10 from task-specific semantic checks after one failure-only correction.
The completed LFT-012 synthetic field test stopped at 6/10: three of five Socratic dialogue checks passed after one correction, but Socratic dialogue learner adaptation [LFT-012] and Socratic dialogue evidence traceability [LFT-012] remained unsupported.
The completed LFT-055 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Accessible Reading, while 1 check remained unresolved.
The completed LFT-040 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Plain language, while 1 check remained unresolved.
The completed LFT-031 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Art comparison, while 1 check remained unresolved.
The completed WFT-009 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Formula Auditing, while 1 check remained unresolved.
The completed WFT-061 synthetic field test finished at 4/10 and was not recommended: only two of five Punch-List Scheduling checks passed after the permitted correction.
This completed synthetic Kernel Diagnostics field test asked the session to diagnose a kernel panic from a bounded evidence packet, preserved an actual five-row kernel panic causal analysis, and derived 4/10 then 10/10 from task-specific semantic checks after one failure-only correction.
The completed LFT-010 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Historical inquiry, while 1 check remained unresolved.
This completed synthetic Display Color field test asked the session to guide a basic monitor color calibration, preserved an actual five-row monitor color calibration measurement record, and derived 4/10 then 6/10 from task-specific semantic checks after one failure-only correction.
The completed WFT-044 synthetic field test stopped at 6/10: three of five Pricing Controls checks passed after one correction, but Pricing Controls task fidelity [WFT-044] and Pricing Controls handoff usability [WFT-044] remained unsupported.
The completed LFT-054 synthetic field test stopped at 6/10: three of five Science Diagnosis checks passed after one correction, but Science Diagnosis content accuracy [LFT-054] and Science Diagnosis learner adaptation [LFT-054] remained unsupported.
This completed synthetic Portability field test asked the session to port a shell automation between operating systems, preserved an actual five-row cross-platform shell portability matrix, and derived 0/10 then 6/10 from task-specific semantic checks after one failure-only correction.
The completed LFT-045 synthetic field test reached 8/10 after one failure-only correction: 4 of five static checks passed for Project planning, while 1 check remained unresolved.