Commit 8d752a7
Grug-native template cleanup and legacy path retirement (#3054)
## Scope
This PR is the full grug-native/template transition slice on this branch
(not just a small cleanup).
## What changed
- Promote `experiments/grug/base/` as the canonical grug edit surface:
- add/expand `experiments/grug/base/model.py`,
`experiments/grug/base/train.py`, `experiments/grug/base/launch.py`
- add `experiments/grug/README.md` and package init files
- Retire legacy Grugformer-era library paths and old comparison script:
- remove `lib/levanter/src/levanter/grug/main.py`
- remove `lib/levanter/src/levanter/grug/data.py`
- remove `lib/levanter/src/levanter/models/grug_wrapper.py`
- remove
`experiments/speedrun/grugformer_vs_hackable_125m/grugformer_vs_hackable_125m.py`
- Remove obsolete grugformer-focused test suite and replace with
template-focused coverage:
- remove `lib/levanter/tests/grug/test_grugformer*.py`
- add/update `tests/test_grug_base_template.py`
- Add callback state-adapter path needed by grug template training:
- add `lib/levanter/src/levanter/callbacks/state_adapter.py`
- update callbacks/tensorstore callback wiring and tests
- Include supporting runtime/parity adjustments used by the template
flow:
- `lib/levanter/src/levanter/eval.py`
- `lib/levanter/src/levanter/utils/jax_utils.py` (+ tests)
- minor compatibility updates in
`lib/levanter/src/levanter/compat/hf_checkpoints.py`
- minor fused CE API touch in
`lib/levanter/src/levanter/kernels/pallas/fused_cross_entropy_loss/api.py`
- Update project/docs guidance to template-first grug workflow:
- `.agents/projects/grugformer.md`
- `docs/recipes/change_grug.md`
- `docs/reports/grug-archive.md`
## Validation run on this branch
- `uv run pytest tests/test_eval.py` (from `lib/levanter`)
- `uv run pytest tests/test_grug_base_template.py`
- `uv run python infra/pre-commit.py --all-files`
## Notes
- This PR intentionally contains the accumulated grug-native transition
work on `codex/grug-native-template-cleanup`.
- Local scratch/monitoring files were not included.
---------
Co-authored-by: Moo Jin Kim <moojink@stanford.edu>1 parent cf99ce5 commit 8d752a7
File tree
29 files changed
+1514
-1711
lines changed- .agents/projects
- docs
- recipes
- reports
- experiments
- grug
- base
- speedrun/grugformer_vs_hackable_125m
- lib/levanter
- src/levanter
- callbacks
- compat
- grug
- models
- utils
- tests
- grug
- tests
29 files changed
+1514
-1711
lines changedLarge diffs are not rendered by default.
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
1 | | - | |
| 1 | + | |
2 | 2 | | |
3 | | - | |
| 3 | + | |
4 | 4 | | |
5 | 5 | | |
6 | 6 | | |
7 | | - | |
8 | | - | |
| 7 | + | |
| 8 | + | |
9 | 9 | | |
10 | | - | |
| 10 | + | |
11 | 11 | | |
12 | | - | |
13 | | - | |
14 | | - | |
15 | | - | |
16 | | - | |
17 | | - | |
| 12 | + | |
| 13 | + | |
| 14 | + | |
| 15 | + | |
| 16 | + | |
| 17 | + | |
| 18 | + | |
| 19 | + | |
18 | 20 | | |
19 | | - | |
| 21 | + | |
20 | 22 | | |
21 | | - | |
| 23 | + | |
22 | 24 | | |
23 | | - | |
| 25 | + | |
24 | 26 | | |
25 | | - | |
| 27 | + | |
| 28 | + | |
| 29 | + | |
| 30 | + | |
| 31 | + | |
26 | 32 | | |
27 | | - | |
28 | | - | |
29 | | - | |
30 | | - | |
31 | | - | |
| 33 | + | |
32 | 34 | | |
33 | | - | |
| 35 | + | |
| 36 | + | |
| 37 | + | |
34 | 38 | | |
35 | | - | |
| 39 | + | |
36 | 40 | | |
37 | | - | |
| 41 | + | |
38 | 42 | | |
39 | | - | |
| 43 | + | |
| 44 | + | |
| 45 | + | |
| 46 | + | |
40 | 47 | | |
41 | | - | |
| 48 | + | |
42 | 49 | | |
43 | | - | |
44 | | - | |
45 | | - | |
46 | | - | |
47 | | - | |
48 | | - | |
| 50 | + | |
49 | 51 | | |
50 | | - | |
| 52 | + | |
| 53 | + | |
| 54 | + | |
51 | 55 | | |
52 | | - | |
| 56 | + | |
53 | 57 | | |
54 | | - | |
| 58 | + | |
| 59 | + | |
| 60 | + | |
| 61 | + | |
55 | 62 | | |
56 | | - | |
57 | | - | |
58 | | - | |
59 | | - | |
60 | | - | |
| 63 | + | |
61 | 64 | | |
62 | | - | |
| 65 | + | |
63 | 66 | | |
64 | | - | |
| 67 | + | |
| 68 | + | |
65 | 69 | | |
66 | | - | |
| 70 | + | |
67 | 71 | | |
68 | | - | |
69 | | - | |
70 | | - | |
| 72 | + | |
71 | 73 | | |
72 | | - | |
73 | | - | |
74 | | - | |
75 | | - | |
76 | | - | |
77 | | - | |
78 | | - | |
79 | | - | |
80 | | - | |
81 | | - | |
82 | | - | |
83 | | - | |
84 | | - | |
85 | | - | |
86 | | - | |
87 | | - | |
88 | | - | |
89 | | - | |
90 | | - | |
91 | | - | |
92 | | - | |
93 | | - | |
94 | | - | |
95 | | - | |
96 | | - | |
97 | | - | |
98 | | - | |
99 | | - | |
100 | | - | |
101 | | - | |
102 | | - | |
103 | | - | |
104 | | - | |
105 | | - | |
106 | | - | |
| 74 | + | |
107 | 75 | | |
| 76 | + | |
108 | 77 | | |
109 | 78 | | |
110 | | - | |
| 79 | + | |
111 | 80 | | |
112 | | - | |
| 81 | + | |
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
1 | 1 | | |
2 | 2 | | |
3 | | - | |
| 3 | + | |
4 | 4 | | |
5 | 5 | | |
6 | 6 | | |
7 | | - | |
8 | | - | |
9 | | - | |
10 | | - | |
11 | | - | |
12 | | - | |
13 | | - | |
14 | | - | |
15 | | - | |
16 | | - | |
17 | | - | |
18 | | - | |
19 | | - | |
20 | | - | |
| 7 | + | |
| 8 | + | |
| 9 | + | |
21 | 10 | | |
22 | 11 | | |
23 | 12 | | |
24 | | - | |
25 | | - | |
26 | 13 | | |
27 | 14 | | |
28 | | - | |
| 15 | + | |
29 | 16 | | |
30 | 17 | | |
31 | 18 | | |
32 | 19 | | |
33 | | - | |
34 | | - | |
35 | | - | |
| 20 | + | |
| 21 | + | |
36 | 22 | | |
37 | 23 | | |
38 | 24 | | |
39 | 25 | | |
40 | | - | |
41 | | - | |
42 | | - | |
43 | | - | |
44 | | - | |
45 | | - | |
46 | | - | |
47 | | - | |
48 | | - | |
49 | | - | |
| 26 | + | |
| 27 | + | |
50 | 28 | | |
51 | 29 | | |
52 | 30 | | |
53 | | - | |
| 31 | + | |
54 | 32 | | |
55 | 33 | | |
56 | 34 | | |
57 | 35 | | |
58 | 36 | | |
59 | | - | |
60 | | - | |
61 | | - | |
| 37 | + | |
| 38 | + | |
| 39 | + | |
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
| 1 | + | |
| 2 | + | |
| 3 | + | |
| 4 | + | |
| 5 | + | |
| 6 | + | |
| 7 | + | |
| 8 | + | |
| 9 | + | |
| 10 | + | |
| 11 | + | |
| 12 | + | |
| 13 | + | |
| 14 | + | |
| 15 | + | |
| 16 | + | |
| 17 | + | |
| 18 | + | |
| 19 | + | |
| 20 | + | |
| 21 | + | |
| 22 | + | |
| 23 | + | |
| 24 | + | |
| 25 | + | |
| 26 | + | |
| 27 | + | |
| 28 | + | |
| 29 | + | |
| 30 | + | |
| 31 | + | |
| 32 | + | |
| 33 | + | |
| 34 | + | |
| 35 | + | |
| 36 | + | |
| 37 | + | |
| 38 | + | |
| 39 | + | |
| 40 | + | |
| 41 | + | |
| 42 | + | |
| 43 | + | |
| 44 | + | |
| 45 | + | |
| 46 | + | |
| 47 | + | |
| 48 | + | |
| 49 | + | |
| 50 | + | |
| 51 | + | |
| 52 | + | |
| 53 | + | |
| 54 | + | |
| 55 | + | |
| 56 | + | |
| 57 | + | |
| 58 | + | |
| 59 | + | |
| 60 | + | |
| 61 | + | |
| 62 | + | |
| 63 | + | |
| 64 | + | |
| 65 | + | |
| 66 | + | |
| 67 | + | |
| 68 | + | |
| 69 | + | |
| 70 | + | |
| 71 | + | |
| 72 | + | |
| 73 | + | |
| 74 | + | |
| 75 | + | |
| 76 | + | |
| 77 | + | |
| 78 | + | |
| 79 | + | |
| 80 | + | |
| 81 | + | |
| 82 | + | |
| 83 | + | |
| 84 | + | |
| 85 | + | |
| 86 | + | |
| 87 | + | |
| 88 | + | |
| 89 | + | |
| 90 | + | |
| 91 | + | |
| 92 | + | |
| 93 | + | |
| 94 | + | |
| 95 | + | |
| 96 | + | |
| 97 | + | |
| 98 | + | |
| 99 | + | |
| 100 | + | |
| 101 | + | |
| 102 | + | |
| 103 | + | |
| 104 | + | |
| 105 | + | |
| 106 | + | |
| 107 | + | |
| 108 | + | |
| 109 | + | |
| 110 | + | |
| 111 | + | |
| 112 | + | |
| 113 | + | |
| 114 | + | |
| 115 | + | |
| 116 | + | |
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
| 1 | + | |
| 2 | + | |
| Original file line number | Diff line number | Diff line change | |
|---|---|---|---|
| |||
| 1 | + | |
| 2 | + | |
0 commit comments