arrow_back
Courses
Courses 95
Courses
“Dream-RSI: Recursive Self-Improvement through Evolving Worlds”
Courses
looped transformer unlocks a new scaling dimension - recurrent depth; i think it received less love
Courses
i’m hiring for 8 new roles…
Courses
Day 255/365 of GPU Programming
Courses
Superintelligence should learn from experience through RL.
Courses
Reasoning from scratch round 3: This time, I cover generating a verifier for...
Courses
🧪 Six Stages, 12.7T Tokens: What Marin's Fully Open 8B Run Actually Teaches
Courses
https://t.co/JTCDbbjKAS is finally out.
Courses
A 30-question breakdown of how embeddings, vector search, and retrieval actually work - similarity m
Courses
@0xSero: Thanks to
Courses
The paper from ICML26 identifies a fundamental limit of RL post-training:with outcome-only rewards,
Courses
Hugging Face Transformers End to End Course
Courses
Harness, Memory, Context Fragments, & the Bitter Lesson
Courses
Thanks to the great @_fracapuano, stable-worldmodel now supports the @LeRobotHF dataset format, enab
Courses
One of my passions is that education should be dispersed freely and as widely as possible, especiall
Courses
Coding Transformer From Scratch - Pytorch Tutorial
Courses
Attention → Mamba cross-architecture distillation is real
Courses
dude this was such a banger...viv is absolutely blast of a person!
Courses
Another quick lecture -- I've been asked many times for prereq's to my book and what you should know
Courses
I launched 3 more videos in my post-training course!
Courses
Liquid AI's head of post-training explained how they built a small model that runs on-device under 1
Courses
Someone correct my analogy if it’s wrong here: Using reverse KL divergence for on policy distillatio
Courses
Recently been training reasoning models, and these loops are super annoying... especially bad for sm
Courses
excellent post by @BerenMillidge on why rl actually works for llm's and how it's different from pre-
auto_stories
Select a bookmark to read
Click any item from the list