Source: evals/results.json · provider gemini · model gemma-4-26b-a4b-it · ran on: laptop

24tasks
20accepted
4rejected
0errors
functioncategorystatusattemptsoracle casesspeeduplower bound
linear_scorefloatrejected32061.41x1.37x
sigmoid_approxfloataccepted22061.65x1.60x
collatz_stepsintaccepted120922.68x21.79x
digital_rootintaccepted12093.32x3.24x
popcountintaccepted12094.99x4.72x
reverse_digitsintaccepted12092.19x2.16x
sum_digits_sqintaccepted12093.01x2.83x
triangularintaccepted120947.14x46.97x
bucket_hashlistaccepted22042.23x2.17x
count_peakslistaccepted12042.13x2.02x
longest_runlistaccepted22041.64x1.54x
max_subarraylistaccepted12041.72x1.56x
median_doubledlistrejected32041.38x1.35x
prefix_max_sumlistrejected32041.46x1.41x
sum_squareslistaccepted12041.74x1.66x
caesar_shiftstringaccepted122511.51x11.48x
count_vowel_runsstringaccepted12255.96x5.28x
count_wordsstringaccepted12257.36x6.74x
is_palindrome_normstringaccepted12252.40x2.39x
longest_alpha_runstringaccepted12256.77x6.30x
normalize_spacesstringrejected32251.22x1.20x
reverse_wordsstringaccepted12251.61x1.58x
snake_to_camelstringaccepted12252.66x2.49x
top_wordstringaccepted22254.96x4.74x

Speedups — measured, not modeled

The gold mark is the conservative lower bound — the acceptance rule reads it, not the median.

triangular
47.14x
collatz_steps
22.68x
caesar_shift
11.51x
count_words
7.36x
longest_alpha_run
6.77x
count_vowel_runs
5.96x
popcount
4.99x
top_word
4.96x
digital_root
3.32x
sum_digits_sq
3.01x
snake_to_camel
2.66x
is_palindrome_norm
2.40x
bucket_hash
2.23x
reverse_digits
2.19x
count_peaks
2.13x
sum_squares
1.74x
max_subarray
1.72x
sigmoid_approx
1.65x
longest_run
1.64x
reverse_words
1.61x

False-accept gate

Wrong rewrites are rejected

13/13 deliberately wrong rewrites rejected by the meta-eval (plus a 16/16 tuple battery). The oracle catches wrong values, wrong exceptions, and panics — so a plausible-looking candidate cannot slip through.

Source: scripts/meta_eval.py · full corpus: evals/results.json

CLI reference

Commands, flags and exit codes. The help text is a captured snapshot of smelt --help; regenerate it with python3 site/scripts/gen_cli_ref.py --refresh.

Commands


smelt run <file.py:fn> [--provider p] [--model m] [--candidate lib.rs] [--cases N] [--json]
smelt capture <file.py:fn> [--cases N] [--json]      # freeze an oracle, no generation
smelt resume <run-dir> [--json]                       # continue an interrupted run
smelt report <run-dir>                                # print the run report (markdown)
smelt scan <path> [--json]                            # candidates with tier + skip reasons
smelt apply <run-dir>                                 # VIEW ONLY diff + fallback shim
smelt doctor [--probe] [--json]                       # environment + provider checks
smelt evals [--provider p] [--model m] [--max-calls N] [--pause-s S] [--out-tag T]
smelt furnace                                         # launch the Go TUI if installed

Exit codes

codemeaning
0accepted — verified, speedup clears the conservative 1.5x lower bound
2config error (bad target, dotted qualname, unsupported annotation on a required path)
3refused — impure target or unsupported-but-understood shape (typed, with a hint)
4rejected — no acceptable rewrite within budget (a correct, honest outcome)
5harness/internal error (report the run dir)

Live smelt --help

                                                                                
 Usage: smelt [OPTIONS] COMMAND [ARGS]...                                       
                                                                                
 Smelt a slow pure Python function into a verified Rust extension. Only proven  
 metal ships. Bare `smelt` launches the Furnace TUI.                            
                                                                                
╭─ Options ────────────────────────────────────────────────────────────────────╮
│ --help          Show this message and exit.                                  │
╰──────────────────────────────────────────────────────────────────────────────╯
╭─ Commands ───────────────────────────────────────────────────────────────────╮
│ run      Run the full smelt loop on one function: path/to/file.py:function   │
│ capture  Capture + freeze the oracle only (no generation, no build).         │
│ resume   Resume an interrupted run from its checkpoint.                      │
│ report   Print a run's report (markdown).                                    │
│ scan     List candidate functions under <path> (source for the TUI picker).  │
│ apply    Print the proposed change for an accepted run — VIEW ONLY, never    │
│          writes.                                                             │
│ furnace  Launch the Furnace TUI (Go binary) if installed; print build steps  │
│          otherwise.                                                          │
│ doctor   Check toolchain, sandbox, local server, and provider keys. Green =  │
│          no FAIL rows.                                                       │
│ evals    Run the frozen corpus sequentially; writes evals/results.json +     │
│          results.md.                                                         │
╰──────────────────────────────────────────────────────────────────────────────╯