This website requires JavaScript.
Explore
Help
Register
Sign In
ros
/
nanochat
Watch
1
Star
0
Fork
0
You've already forked nanochat
mirror of
https://github.com/karpathy/nanochat.git
synced
2026-01-30 04:22:02 +00:00
Code
Issues
Packages
Projects
Releases
Wiki
Activity
Files
ba4f40bf588a83ed3ee4d3c02cb7581edfb105ba
nanochat
/
scripts
History
Andrej Karpathy
cf587acb1a
move eval bundle download to be lazy and inside the python code so that we can substantially simplify the run bash scripts
2025-11-01 16:04:38 +00:00
..
base_eval.py
move eval bundle download to be lazy and inside the python code so that we can substantially simplify the run bash scripts
2025-11-01 16:04:38 +00:00
base_loss.py
many small tweaks. base, eval, core work now i think
2025-10-16 15:46:18 -07:00
base_train.py
fix tok/sec calculation bug when grad accum steps > 1
2025-10-30 08:36:32 -07:00
chat_cli.py
upgrading all other files to be able to use cpu/mps as well as cuda. various minor other changes ,e.g. changing max_iterations to num_iterations in sft script for consistency in naming
2025-10-20 10:15:17 -07:00
chat_eval.py
typo fixes in scripts
2025-10-28 20:17:31 +01:00
chat_rl.py
typo fixes in scripts
2025-10-28 20:17:31 +01:00
chat_sft.py
add the SpellingBee task so that nanochat can count r in strawberry etc. along the way we had to add a bunch of new functionality, e.g. extend the calculator to support the count function of python. possibly the current TaskMixture uses way too many synthetic examples of SpellingBee because the eval gives us exactly 100% performance on spelling. We can tune this later to reclaim some wall clock time here I think
2025-10-24 14:02:48 +00:00
chat_web.py
upgrading all other files to be able to use cpu/mps as well as cuda. various minor other changes ,e.g. changing max_iterations to num_iterations in sft script for consistency in naming
2025-10-20 10:15:17 -07:00
mid_train.py
Fix tok/sec metrics for base_train and mid_train when gradient accumulation is not 1
2025-10-26 01:43:49 -05:00
tok_eval.py
initial commit
2025-10-13 06:49:24 -07:00
tok_train.py
initial commit
2025-10-13 06:49:24 -07:00