Task · Design discovery · v2

MZI-Ring-v2

MZI-coupled microring add-drop filter electrothermal design v2

Design a fifth-order four-port Vernier microring add-drop filter with two physical bus MZIs, seven heaters, two fixed VOAs, and trusted local electrothermal and optical scoring.

Published runCodex · GPT-5.6-sol
No valid design
Reporting score S · provisional calibration—
Official score J · lower is better—

J_acc = 0.3 · calibration design-jacc-2026-09-05-v2. S ≥ 1 meets the aggregate reporting reference, not every optical or resource specification. Joint feasibility of this reference remains unresolved.

Saved-run scores and gate outcomes are unchanged; this reporting update is not a new evaluation against current task code.

Published runClaude Code · Claude Opus 5.5
Valid
Reporting score S · provisional calibration0.0206
Official score J · lower is better28.7874

J_acc = 0.3 · calibration design-jacc-2026-09-05-v2. S ≥ 1 meets the aggregate reporting reference, not every optical or resource specification. Joint feasibility of this reference remains unresolved.

Saved-run scores and gate outcomes are unchanged; this reporting update is not a new evaluation against current task code.

2 published runs are compared on one elapsed-time chart. Their final layouts and score evidence remain separate.

Evidence chain

Why these results are reported this way

Codex · GPT-5.6-solFinal ineligible

Trusted final: infrastructure valid; geometry checks executed; gates not passed. Official J withheld; reporting S withheld.

Verified full run archive
Claude Code · Claude Opus 5.5Final eligible

Trusted final: infrastructure valid; geometry checks executed; gates passed. Official J 28.7874; reporting S 0.0206.

Verified full run archive
How to interpret this result

A valid outcome means the submitted artifact cleared the evaluator gates recorded for this archived run. It does not imply that every possible physical specification, fabrication condition, or later task revision has been independently re-tested.

Run trajectory · model comparison

Reporting score S vs. time

Each run starts at its own elapsed time zero. Hollow markers are gate-ineligible probes, not measured S = 0.

Codex · GPT-5.6-solcircle probes · solid bestClaude Code · Claude Opus 5.5triangle probes · dashed best
Published run 1

Codex · GPT-5.6-sol

No valid design · 40 probes · reporting score S ineligible

Final artifact · Codex · GPT-5.6-sol

Codex · GPT-5.6-sol GDS layout

The exact final submission, rendered from its evaluator-produced GDS. It failed the final gates and has no official score.

GDS previewInert SVGFinal submission

MZI-Ring-v2 · Device layout

Codex · GPT-5.6-sol · final GDS

Fit
Codex · GPT-5.6-sol · final GDSlayout sha256 d5f9509d9b3f
Published run 2

Claude Code · Claude Opus 5.5

Valid · 40 probes · reporting score S 0.0206

Final artifact · Claude Code · Claude Opus 5.5

Claude Code · Claude Opus 5.5 GDS layout

The exact submission the evaluator scored, rendered from its GDS

GDS previewInert SVGFinal submission

MZI-Ring-v2 · Device layout

Claude Code · Claude Opus 5.5 · final GDS

Fit
Claude Code · Claude Opus 5.5 · final GDSlayout sha256 ad56da938abe