Files
tinygrad/extra/models
chenyuandGitHub f88506e630 move gpt2/llama sampling inside the model call (#3013)
* move gpt2/llama sampling inside the model call

* argmax uses one more kernel
2024-01-04 17:01:50 -05:00
..
2023-11-28 17:36:55 -08:00
2023-11-28 17:36:55 -08:00
2024-01-01 14:58:48 -08:00
2023-11-28 17:36:55 -08:00
2023-11-28 17:36:55 -08:00
2023-11-28 17:36:55 -08:00
2023-11-28 17:36:55 -08:00
2023-11-28 17:36:55 -08:00