mirror of
https://github.com/karpathy/llm.c.git
synced 2026-07-26 20:15:08 -04:00
adjust readme to new scripts/ dir
This commit is contained in:
parent
a0d5dbbda6
commit
fc40ffc4aa
1 changed files with 1 additions and 1 deletions
|
|
@ -4,7 +4,7 @@ LLM training in simple, pure C/CUDA. There is no need for 245MB of PyTorch or 10
|
|||
|
||||
Our current goal is to reproduce GPT-2. For an overview of current ongoing work, see the latest [State of the Union](https://github.com/karpathy/llm.c/discussions/344) post.
|
||||
|
||||
Update May 28, 2024: A useful recent post may be ["Reproducing GPT-2 (124M) in llm.c in 90 minutes for $20"](https://github.com/karpathy/llm.c/discussions/481) where I detail the steps to go from scratch to reproducing the GPT-2 miniseries for 124M/350M models. The files themselves that show the launch commands are [run124M.sh](run124M.sh) and [run350M.sh](run350M.sh).
|
||||
Update May 28, 2024: A useful recent post may be ["Reproducing GPT-2 (124M) in llm.c in 90 minutes for $20"](https://github.com/karpathy/llm.c/discussions/481) where I detail the steps to go from scratch to reproducing the GPT-2 miniseries for 124M/350M models. E.g. have a look at the `scripts` directory, example: [scripts/run_gpt2_124M.sh](scripts/run_gpt2_124M.sh).
|
||||
|
||||
I'd like this repo to only maintain C and CUDA code. Ports of this repo to other languages are very welcome, but should be done in separate repos, and then I am happy to link to them below in the "notable forks" section, just like I did in [llama2.c notable forks](https://github.com/karpathy/llama2.c/tree/master?tab=readme-ov-file#notable-forks).
|
||||
|
||||
|
|
|
|||
Loading…
Add table
Add a link
Reference in a new issue