Digest | Garden
arrow_back

Courses

Courses 95

Courses

“Dream-RSI: Recursive Self-Improvement through Evolving Worlds”

Sep 17

Courses

looped transformer unlocks a new scaling dimension - recurrent depth; i think it received less love

Sep 16

Courses

i’m hiring for 8 new roles…

Sep 16

Courses

Day 255/365 of GPU Programming

Sep 15

Courses

Superintelligence should learn from experience through RL.

Sep 14

Courses

Reasoning from scratch round 3: This time, I cover generating a verifier for...

Sep 13

Courses

🧪 Six Stages, 12.7T Tokens: What Marin's Fully Open 8B Run Actually Teaches

Sep 14

Courses

https://t.co/JTCDbbjKAS is finally out.

Sep 13

Courses

A 30-question breakdown of how embeddings, vector search, and retrieval actually work - similarity m

Aug 26

Courses

@0xSero: Thanks to

Aug 29

Courses

The paper from ICML26 identifies a fundamental limit of RL post-training:with outcome-only rewards,

Aug 28

Courses

Hugging Face Transformers End to End Course

Aug 23

Courses

Harness, Memory, Context Fragments, & the Bitter Lesson

Apr 12

Courses

Thanks to the great @_fracapuano, stable-worldmodel now supports the @LeRobotHF dataset format, enab

Apr 12

Courses

One of my passions is that education should be dispersed freely and as widely as possible, especiall

Apr 14

Courses

Coding Transformer From Scratch - Pytorch Tutorial

Apr 15

Courses

Attention → Mamba cross-architecture distillation is real

Apr 19

Courses

dude this was such a banger...viv is absolutely blast of a person!

Apr 17

Courses

Another quick lecture -- I've been asked many times for prereq's to my book and what you should know

Jun 24

Courses

I launched 3 more videos in my post-training course!

Jun 15

Courses

Liquid AI's head of post-training explained how they built a small model that runs on-device under 1

Jul 30

Courses

Someone correct my analogy if it’s wrong here: Using reverse KL divergence for on policy distillatio

Jul 12

Courses

Recently been training reasoning models, and these loops are super annoying... especially bad for sm

Jul 7

Courses

excellent post by @BerenMillidge on why rl actually works for llm's and how it's different from pre-

Aug 15
95 items
← Prev Next →
auto_stories

Select a bookmark to read

Click any item from the list