Simon Willison’s Weblog

Subscribe

Monday, 2nd January 2023

nanoGPT. “The simplest, fastest repository for training/finetuning medium-sized GPTs”—by Andrej Karpathy, in about 600 lines of Python.

# 11:27 pm / python, ai, gpt-3, andrej-karpathy, generative-ai, llms

Petals (via) The challenge with large language models in the same scale ballpark as GPT-3 is that they’re large—really large. Far too big to run on a single machine at home. Petals is a fascinating attempt to address that problem: it works a little bit like BitTorrent, in that each user of Petal runs a subset of the overall language model on their machine and participates in a larger network to run inference across potentially hundreds of distributed GPUs. I tried it just now in Google Colab and it worked exactly as advertised, after downloading an 8GB subset of the 352GB BLOOM-176B model.

# 11:29 pm / ai, gpt-3, generative-ai, llms, bloom, gpus

2023 » January

MTWTFSS
      1
2345678
9101112131415
16171819202122
23242526272829
3031