How Did DeepSeek Make V4 So Cheap?
Need to fine-tune a model without the hassle? Try out Crusoe's serverless fine-tuning today! https://www.crusoe.ai/contact-sales/s... Learn AI intuitively, intro into LLMs with minimal math! https://intuitiveai.academy/ limited time code "SUMMER" for 25% off yearly plan We just wrote a new piece on Optimizers After a month of delay, here is my part 1 breakdown of the DeepSeek-V4 paper. In this video, I'll be covering all the key developments they've made that you should know if you want to keep up with the frontier of AI. I will have a part 2 that is a deep dive into their infrastructure side of developments that is a lot more advanced so stay tuned! *thumbnail: this is the price per million tokens when the input is a cache hit. The normal price for input (cache miss) is $0.435. My Newsletter https://mail.bycloud.ai/ My Patreon / bycloud DeepSeek-V4 [Paper] https://www.alphaxiv.org/abs/deepseek-v4 Try out my new fav place to learn how to code https://scrimba.com/?via=bycloudAI This video is supported by the kind Patrons & YouTube Members: 🙏Spam Maj, Alex, Chris LeDoux, DX Research Group, Poof N' Inu, Deagan, Robert Zawiasa, Ryszard Warzocha, Tobe2d, Louis Muk, Akkusativ, Kevin Tai, Mark Buckler, NO U, Tony Jimenez, Ângelo Fonseca, jiye, Anushka, Asad Dhamani, Binnie Yiu, Calvin Yan, Clayton Ford, Diego Silva, Etrotta, Gonzalo Fidalgo, Handenon, Hector, Jake Disco very, Michael Brenner, Nilly K, OlegWock, Daddy Wen, Shuhong Chen, Sid_Cipher, Stefan Lorenz, Sup, tantan assawade, Thipok Tham, Thomas Di Martino, Thomas Lin, Richárd Nagyfi, Paperboy, mika, Leo, Berhane-Meskel, Kadhai Pesalam, mayssam, Bill Mangrum, nyaa, Toru Mon, Lame Plane, Matej Macak, Len Mo, saylikhapekar, ZyanSheep, THEVIERAOS, Ricardo Raphael Corona-Moreno [Discord] / discord [Twitter] / bycloudai [Patreon] / bycloud [Business Inquiries] [email protected] [Other Inquiries] [email protected] [Profile & Banner Art] / pygm7 [Video Editor] @Booga04 Manim Animations created with Manimate https://www.manimate.ai/ [Ko-fi] https://ko-fi.com/bycloudai

GLM-5.2: DeepSeek Was Wrong About RL?

Why AI Tokens are so Expensive - Computerphile

My Honest Thoughts about Deepseek

How Did A Chinese Phone Company Topped Open Source LLM?

WebAssembly Is Quietly Killing Docker (Millisecond Startup)

Android 17 sucks. So I put Linux on a phone.

"Software Fundamentals Matter More Than Ever" — Matt Pocock

Putin trapped in crisis as terrified oligarchs fear collapse | Philip Ingram

I Built an LLM From Scratch

Deepseek drops another HUGE breakthrough

Making AMD's Ryzen AI Halo Do Work

Starship Flight 13 - Where's The Kaboom?

The insane engineering of Deepseek V4

NVIDIA didn't want me to do this

Microsoft Just Dropped LLM's Frontier Data Engineering Secrets

How Meta Went From Open Source Hero to AI's Biggest Villain

you need to use Hermes RIGHT NOW!! (goodbye OpenClaw!!)

How DeepSeek Runs a 284B LLM on a Laptop (Run AI Locally)

The Insane Infrastructure Design of DeepSeek V4

