How to approach post-training for AI applications
My talk during NeurIPs at Infer -- the Vancouver AI Engineering group: https://infervan.com/ This was a fun one. I was trying to think of "what to say" to AI engineers. What are the things I'm learning that actually translates to useful advice? What will people building AI applications get wrong if they see new RLHF or finetuning papers and think they're going to try it? When will we have a research ecosystem around finetuning from instruct? This is a fun one, I hope you like it. As usual, reach out if you have questions! (This also has a bunch of my content on OpenAI's reinforcement finetuning API) Slides: https://docs.google.com/presentation/... For more, subscribe here and to my primary distribution channel, Interconnects.ai. Get Interconnects (https://www.interconnects.ai/)... ... on YouTube: / @interconnects ... on Twitter: https://x.com/interconnectsai ... on Linkedin: / interconnects-ai ... on Spotify: https://open.spotify.com/show/2UE6s7w... … on Apple Podcasts: https://podcasts.apple.com/us/podcast...

The Agent Development Lifecycle: Build, Test, Deploy, Monitor | Interrupt 26

RLHF and Post-training Overview | RLHF & Post-Training Book Course, Lecture 1

Claude Is Putting Startups Out of Business

RFT, DPO, SFT: Fine-tuning with OpenAI — Ilan Bigio, OpenAI

Everything You Wanted to Know About LLM Post-Training, with Nathan Lambert of Allen Institute for AI

Transformers, the tech behind LLMs | Deep Learning Chapter 5

Andrej Karpathy: Software Is Changing (Again)

GRPO's new variants and implementation secrets

RL for Agents Workshop - Deep Dive on Training Agents with RL and Open Source

Introducing live-1: Streaming Speaker Diarization model and builds a real-time system live

Why Netflix is betting on systems thinkers—not specialists—in the AI era | Elizabeth Stone (CPTO)

CHOSEN ONE!! YOUR IDENTITY REVEAL JUST SHOOK THE INTERNET... AND THEIR MINDS

Andrej Karpathy: From Vibe Coding to Agentic Engineering w/ Stephanie Zhan

Building Better AI Agents: Artificial Intelligence Observability How To

Stanford Webinar - Large Language Models Get the Hype, but Compound Systems Are the Future of AI

REVEALED: IS GERMANY HOLDING TALKS WITH RUSSIA?

How language model post-training is done today

Complete Terraform Course - From BEGINNER to PRO! (Learn Infrastructure as Code)

Stanford CS25: V4 I Aligning Open Language Models

How AI agents & Claude skills work (Clearly Explained)

10 tips to level up your ai-assisted coding - Aleksander Stensby - NDC Copenhagen 2026

