openmoji-g1-gpu-e7dc920-0bafd5c-9b9b1699
failed —
Hypothesis
The exact locally verified Gate G pipeline completes one CUDA train step on the owned RTX 4080, restores a canonical checkpoint in durable storage, and preserves locked paths exactly.
Why run it
Validates the real data, selected normalizers, conditioned model, optimizer, checkpoint, and locked-edit path in the target GPU environment before larger work. Also tests the single corrected factor: the staged wrapper now resolves relative input paths against the immutable source root rather than the image working directory.
Measured behaviour
Table view
| optimizer step | train accuracy |
|---|---|
| 1 | 0.0051 |
Measured result
Verbatim from the run's summary.json.
d11efa006657dafc…0bafd5c3b3de29da…71ee91a4c99b9f53…Written result
Status: failed during the first CUDA step; no checkpoint written.
This run is bounded to the exact one-step local pilot and stages only its six selected raw SVGs plus the pinned palette with the committed source snapshot.
It retries openmoji-g1-gpu-cd3250e-0bafd5c-9b9b1699, which verified its stage but failed before model construction because the staged pilot resolved relative input paths against the image working directory. The only changed factor is the wrapper's working directory; the config, fixture, seed, step budget, image, and cap are unchanged.
The working-directory correction worked: the staged config, pinned palette, six raw SVGs, selected normalizers, conditioned model, and optimizer all loaded from the immutable source root. The run then failed inside the first CUDA step because the pipeline declares torch.use_deterministic_algorithms(True) and CUDA >= 10.2 cuBLAS needs CUBLAS_WORKSPACE_CONFIG to honor that declaration. The container left an empty artifact directory, no checkpoint, and no running container.
This is a launcher environment gap, not a weakening of the determinism contract. The correction sets CUBLAS_WORKSPACE_CONFIG=:4096:8 explicitly in the owned Docker smoke launcher, which makes the declared determinism achievable rather than relaxing it.
State transitions
- planned2026-09-20T11:01:52Z
- staged2026-09-20T11:03:17Z
- failed2026-09-20T11:04:35Zcuda_deterministic_algorithms_require_cublas_workspace_config
Run record
Verbatim from runs/openmoji-g1-gpu-e7dc920-0bafd5c-9b9b1699/run.yaml, the record committed before launch.
7c8c71136016300d…f74a8a69ede13b32…0bafd5c3b3de29da…9b9b1699677a6f97…e0cdb2a3cc8f00df…