Embedded AI - Intelligence at the Deep Edge

Large Language Monkeys: Why Noise Yields No Knowledge

David Such Season 6 Episode 7

Use Left/Right to seek, Home/End to jump to start or end. Hold shift to jump forward or backward.

0:00 | 22:31

Send us Fan Mail

There is an old claim that a truly random source contains all knowledge: give monkeys enough time at typewriters and Shakespeare falls out. In 2024 two Sydney mathematicians did the arithmetic and found the universe ends first. But the idea has a modern tail. We now have language models that can spot meaningful text instantly, so why not let randomness generate and an LLM extract? This episode works through why that fails, and why the failure is precise: in a random stream, the address of any text costs as many bits as the text itself. Along the way: Borges' Library of Babel, a website that actually built it, DeepMind systems that made the generate-and-filter idea work by cheating in exactly the right way, and what your brain does with noise that an LLM cannot.

Support the show

If you are interested in learning more then please subscribe to the podcast or head over to https://medium.com/@reefwing, where there is lots more content on AI, IoT, robotics, drones, and development. To support us in bringing you this material, you can buy me a coffee or just provide feedback. We love feedback!