r2s-full-v7-systems-42762c2-e9842024-47646604
completed —
Hypothesis
Trained on the 2,681 family-disjoint training icons with metric coordinates, path-major order, online exact augmentation and compositions, the 9.0M-parameter render-to-program model transcribes held-out validation renders better than retrieving the closest training icon. Pass requires all three: (1) mean 72 px pixel error below the nearest-training-icon baseline with the paired 95% interval on the reduction excluding zero; (2) CLIP top-1 on Gate N's 32 validation icons at or above OmniSVG 4B zero-shot, 0.609; (3) median single-icon latency on the RTX 4080 under 500 ms.
Visual output
Written result
Hypothesis. Trained on the 2,681 family-disjoint training icons with metric coordinates, path-major order, online exact augmentation and compositions, the 9.0M-parameter render-to-program model transcribes held-out validation renders better than retrieving the closest training icon. Pass requires all three: (1) mean 72 px pixel error below the nearest-training-icon baseline with the paired 95% interval on the reduction excluding zero; (2) CLIP top-1 on Gate N's 32 validation icons at or above OmniSVG 4B zero-shot, 0.609; (3) median single-icon latency on the RTX 4080 under 500 ms.
Result
| measure | value |
|---|---|
| icons evaluated | 339 (primary/validation) |
| model pixel error, mean [95% CI] | 0.0765 [0.0716, 0.0815] |
| model pixel error, median | 0.0715 |
| exact program rate | 0.021 |
| rendered rate | 1.000 |
| blank canvas pixel error | 0.1720 |
| nearest training icon pixel error | 0.0897 [0.0848, 0.0947] |
| error reduction vs nearest icon | 0.0132 [0.0095, 0.0169] |
| icons where model beats nearest icon | 226 |
| CLIP top-1, model greedy (32 Gate N icons) | 0.469 |
| CLIP top-1, nearest training icon | 0.344 |
| CLIP top-1, OmniSVG 4B zero-shot (Gate N) | 0.609 |
Resources
| measure | value |
|---|---|
| device | NVIDIA GeForce RTX 4080 |
| parameters | 9,025,826 |
| train seconds | 5708 |
| peak VRAM GiB | 4.341529369354248 |
| latency ms/icon, batch 1, median | 1490.8 |
| latency ms/icon, batch 1, p95 | 1509.5 |
| throughput ms/icon, batch 64 | 71.5 |
| model calls per icon (decoding) | 728.0 |
| torch / CUDA | 2.14.0a0+4fdf77b940.nv26.08 / 13.4 |
Latency covers encoding and decoding to a validated token program; it excludes rasterising the SVG. Samples: samples.png, rows are reference, model, nearest training icon.
State transitions
- running2026-09-27T23:03:56Z
- completed2026-09-28T00:40:20Z
Run record
Verbatim from runs/r2s-full-v7-systems-42762c2-e9842024-47646604/run.yaml, the record committed before launch.
62d5b4aede5048c1…e9842024e1647f9f…476466042da98d72…



