Settings

Theme

The Unreasonable Effectiveness of Recurrent Neural Networks (2015)

karpathy.github.io

2 points by Kotlopou · 1 comment

Reader

1 thread
KotlopouOP

I thought this is a cool premonition of the eventual success of language models. One can nicely see a microcosm of everything with generative AI (text generated in a domain you don't understand looks fine at first glance), already a mention of attention models, and there's even a reference to the model "hallucinating" some URL.

It's also all charmingly tiny. I was surprised at the quality of the LaTeX generated from a network trained on a single book -- it looks like a proof, has sequential lemmas, closes all brackets...

Also interesting to read the original discussion with a decade of hindsight: https://news.ycombinator.com/item?id=9584325

Keyboard Shortcuts

j
Next item
k
Previous item
o / Enter
Open selected item
?
Show this help
Esc
Close modal / clear selection