AI can't cross this line and we don't know why.
Have we discovered an ideal gas law for AI? Head to https://brilliant.org/WelchLabs/ to try Brilliant for free for 30 days and get 20% off an annual premium subscription. Welch Labs Imaginary Numbers Book! https://www.welchlabs.com/resources/i... Welch Labs Posters: https://www.welchlabs.com/resources Support Welch Labs on Patreon! / welchlabs Special thanks to Patrons: Juan Benet, Ross Hanson, Yan Babitski, AJ Englehardt, Alvin Khaled, Eduardo Barraza, Hitoshi Yamauchi, Jaewon Jung, Mrgoodlight, Shinichi Hayashi, Sid Sarasvati, Dominic Beaumont, Shannon Prater, Ubiquity Ventures, Matias Forti, Brian Henry, Tim Palade, Petar Vecutin Learn more about WelchLabs! https://www.welchlabs.com TikTok: / welchlabs Instagram: / welchlabs REFERENCES A Neural Scaling Law from the Dimension of the Data Manifold: https://arxiv.org/pdf/2004.10802 First 2020 OpenAI Scaling Paper: https://arxiv.org/pdf/2001.08361 GPT-3 Paper: https://arxiv.org/pdf/2005.14165 Second 202 OpenAI Scaling Paper: https://arxiv.org/pdf/2010.14701 Google Deepmind “Chinchilla Scaling” Paper: https://arxiv.org/abs/2203.15556 Nice summary of Chinchilla Scaling: https://www.lesswrong.com/posts/6Fpvc... GPT-4 Technical Report: https://arxiv.org/pdf/2303.08774 Nice Neural Scaling Laws Summary: https://www.lesswrong.com/posts/Yt5wA... Explaining Neural Scaling Laws: https://arxiv.org/pdf/2102.06701 High Cost of Training GPT-4: https://www.wired.com/story/openai-ce... Nvidia V100 FLOPs: https://lambdalabs.com/blog/demystify... Nvidia V100 Original Price: [https://www.microway.com/hpc-tech-tip... GPU model,Key Points](https://www.microway.com/hpc-tech-tip...) Great paper on scaling up training infrastructure: https://arxiv.org/pdf/2104.04473 Eight Things to Know about LLMs: https://arxiv.org/abs/2304.00612 Emergent Properties of LLMs: https://arxiv.org/abs/2206.07682 Theoretical Motivation for Cross Entropy (Section 6.2): https://www.deeplearningbook.org/ Some papers that appear to pass the compute efficient frontier https://arxiv.org/pdf/2206.14486 https://arxiv.org/abs/2210.11399 CFAQJOTYQHT7JYIT Leaked GPT-4 training info https://patmcguinness.substack.com/p/... https://www.semianalysis.com/p/gpt-4-... https://epochai.org/blog/tracking-lar...

What Did Ilya See?

THIS is why large language models can understand the world
![Yann LeCun's $1B Bet Against LLMs [Part 1]](https://i.ytimg.com/vi/kYkIdXwW2AE/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLDbV4izF3i-wxevCVIn7FJjoy1vlA)
Yann LeCun's $1B Bet Against LLMs [Part 1]
![Inside the World's Smartest Robot Brain [VLA]](https://i.ytimg.com/vi/2mrGMMmrVNE/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLBsKsDGBMWqIscKPoqdk4iZtfZeeQ)
Inside the World's Smartest Robot Brain [VLA]

Why AI Tokens are so Expensive - Computerphile

LLMs Don't Need More Parameters. They Need Loops.

Visualizing transformers and attention | Talk for TNG Big Tech Day '24

The Pyramids Were Easy To Build, Actually
![The Real Reason Huge AI Models Actually Work [Prof. Andrew Wilson]](https://i.ytimg.com/vi/M-jTeBCEGHc/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLBhGNRMnPq6KLzUK1NQBFqKWiYhMA)
The Real Reason Huge AI Models Actually Work [Prof. Andrew Wilson]

"Software Fundamentals Matter More Than Ever" — Matt Pocock
![The Dark Matter of AI [Mechanistic Interpretability]](https://i.ytimg.com/vi/UGO_Ehywuxc/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLBkSvGfku9uu1v4EkqTxrcfZ6YBMA)
The Dark Matter of AI [Mechanistic Interpretability]

But how do AI images and videos actually work? | Guest video by Welch Labs
![What the Books Get Wrong about AI [Double Descent]](https://i.ytimg.com/vi/z64a7USuGX0/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLA-4fwCE2AD2Ap9tqWfLdo7_PPLKA)
What the Books Get Wrong about AI [Double Descent]

The Uncomfortable Truth About AI “Reasoning” | World Science Festival

Why AI Can Never Escape Turing's 1936 Proof
![The Misconception that Almost Stopped AI [How Models Learn Part 1]](https://i.ytimg.com/vi/NrO20Jb-hy0/hqdefault.jpg?sqp=-oaymwEjCNACELwBSFryq4qpAxUIARUAAAAAGAElAADIQj0AgKJDeAE=&rs=AOn4CLCiksXndIEYQZVVoTfArQwhou-eWw)
The Misconception that Almost Stopped AI [How Models Learn Part 1]

The Brain’s Learning Algorithm Isn’t Backpropagation

AI Bubble vs Dot Com Crash. History is REPEATING

AlphaFold - The Most Useful Thing AI Has Ever Done

