Karpathy likes procedural code for storyboarding and control, with video-to-video models for texturing
“@rainisto @DavidmComfort agree!! i quite like the idea of procedural code for storyboarding and control, and then video to video models for texturing and looksmaxxing.”
Karpathy clarifies he used Eleven Labs for audio, with LLMs able to use APIs
“@cunkpyber Eleven Labs for the audio. LLMs can easily use the APIs (here I did that part manually because I felt picky about the voice).”
Karpathy shares browser-playable source for his Lord of the Rings test and jokes about GTA Hobbiton
“More on the pelican on the bicycle test from @simonw: https://t.co/OXmtODyTKj I uploaded the source here so it's playable in the browser, forkable etc. https://t.co/w3Nctc888d Look out for GTA Hobbiton dropping before GTA VI :)”
Karpathy wonders if ngrams or decision trees give better log probs than neural models at 25KB constraint
“@MattBeton so fun! :) at some point i wonder if ngram (tables) or even something like decision trees start to give superior log probs, and at much smaller program lengths overall (sum of program + weights). i.e. what is the best val loss model overall, for 25KB of user space. fun q!”
Karpathy tests Opus 5 with first paragraph of Lord of the Rings and 1M token budget at roughly $10 cost
“We're starting to leave the territory where you'd test an LLM by e.g. "create an svg of pelican on a bicycle". As one idea to generalize it, I was interested what Opus 5 would do if I gave it the first paragraph of the Lord of the Rings, a 1M token budget (~$10) and asked for https://t.co/ybIpXhYSsj”
Karpathy shares that he gave Three.js feedback to Opus and learned about instanced meshes
“@threejs I was curious and gave your feedback to Opus. TIL! https://t.co/t5HvyZsBlh”