Navers lab
← All tasks
lang04 Language rewrite Task 4 / 20

JavaScript → Rust

acorn 8.14.0 JavaScript Rust 1.90

The implementation language moves, the artifact does not. 9,571 lines, 30 hours per model, and 16,039 checks across 13 modules.

Runs
26

one per configuration

Past rung 1
20

of 26 — the migration happened

Past rung 2
0

of those 20 — every check perfect

Past rung 3
0

no run got this far

Best composite

nothing scored here

The instrument

Acceptance gates

7

decide whether the migration happened

Behavioral modules

13

ported suites, run against the new tree

Frozen checks

16,039

all must pass to open rung three

Verifiers

6

hostile programs, written against the spec

Where the 26 runs stopped

6
Never migrated
20
Migrated, behavior lost
0
Perfect suite, gate rejected
0
Verifier found a difference
0
Accepted

No run was perfect on all 16,039 checks. No run reached rung three, so no verifier was paid to attack this task.

By model

Model Runs Gate Ceil. Where they stopped Mean Best
claude-opus-5 5 5 0
0.0
claude-sonnet-5 5 4 0
0.0
dsv4-flash 1 1 0
0.0
glm-5.2 1 1 0
0.0
gpt-5.6-luna 6 5 0
0.0
gpt-5.6-sol 6 2 0
0.0
kimi-k3 1 1 0
0.0
qwen3.8-max 1 1 0
0.0

Every run

best composite first

One row per configuration. Shape is the run's tool sequence in 24 slices — pale is shell, dark is an edit — and each log is the full session as recorded.

ModelEffortOutcomeGateChecksVerif.ScoreShape
qwen3.8-max max partial 7/7 99.3% 0
claude-opus-5 high partial 7/7 97.5% 0
kimi-k3 max partial 7/7 97.3% 0
claude-sonnet-5 high partial 7/7 96.7% 0
claude-opus-5 xhigh partial 7/7 95.8% 0
claude-opus-5 medium partial 7/7 95.1% 0
dsv4-flash max partial 7/7 93.7% 0
claude-opus-5 low partial 7/7 93.3% 0
claude-sonnet-5 xhigh partial 7/7 90.7% 0
claude-opus-5 max partial 7/7 87.8% 0
gpt-5.6-sol max failed 6/7 87.8% 0
gpt-5.6-luna max partial 7/7 84.0% 0
glm-5.2 max partial 7/7 80.6% 0
gpt-5.6-luna xhigh failed 6/7 79.4% 0
gpt-5.6-sol high partial 7/7 65.5% 0
gpt-5.6-luna high partial 7/7 48.0% 0
gpt-5.6-luna medium partial 7/7 39.5% 0
gpt-5.6-sol xhigh partial 7/7 36.1% 0
gpt-5.6-luna none partial 7/7 31.2% 0
gpt-5.6-luna low partial 7/7 0.8% 0
claude-sonnet-5 low partial 7/7 0.3% 0
gpt-5.6-sol low failed 6/7 0.1% 0
gpt-5.6-sol none failed 4/7 0.1% 0
gpt-5.6-sol medium failed 6/7 0.1% 0
claude-sonnet-5 medium partial 7/7 0.0% 0
claude-sonnet-5 max failed 4/7 0.0% 0