 George HotzandGitHub
|
57ae1bc7a7
|
rename MULTI to UNSHARD (#17267)
* rename MULTI to UNSHARD
* comment updates (glm)
* rename method to unshard
|
2026-07-28 16:51:41 -07:00 |
|
 chenyuandGitHub
|
33b635d23a
|
Tensor.train -> TRAINING [PR] (#16705)
* Tensor.train -> TRAINING [PR]
* doc
|
2026-06-22 15:13:22 -04:00 |
|
 sirhcmandGitHub
|
e0fe6e542e
|
ci: fewer pydeps (#16654)
|
2026-06-17 22:52:14 -04:00 |
|
 sirhcmandGitHub
|
acf239e4d2
|
specify renderer in DEV, <dev>_<ren>=1 is deprecated (#15551)
|
2026-03-31 18:35:14 -04:00 |
|
 hoovedandGitHub
|
1e8945a28c
|
Training loop for Stable Diffusion mlperf (#12315)
* add diff
* fix edit error
* match master
* point reference to specific commit
* simplify wandb logging
* remove lr test, dehardcode device
* increase stack size limit
|
2025-10-03 02:45:38 -04:00 |
|
 hoovedandGitHub
|
5d9035f5a6
|
Eval for Stable Diffusion mlperf (#12316)
* add diff
* rerun ci
* refactor beam workaround, add test
* fix conflict
* linting
|
2025-10-02 02:35:38 -04:00 |
|
 
|
0f804c9a83
|
Stable Diffusion model init for mlperf (#12314)
* include clip pr diff
* updated unet and sd init
* dehardcode default device
* revert beam hang workaround
---------
Co-authored-by: chenyu <[email protected]>
|
2025-10-02 02:28:41 -04:00 |
|
 hoovedandGitHub
|
c2689c505e
|
Clip model updates for Stable Diffusion mlperf training (#12313)
* stable diffusion mlperf clip changes
* add clip tests
* set gelu as attribute
* add more tests
* factor out GPUS
* rerun CI
* add imports to if blocks
* remove unneeded axis
* add clip tests to CI
* move clip tests
* add deps, disable max buf size
|
2025-09-29 21:50:14 -04:00 |
|