Digest | Garden
arrow_back

Articles

Articles 143

Articles

Currently doing a write up on scaling laws for RL. Here are the papers I'm covering so far:

Apr 17

Articles

very high quality and detailed content if you want to learn how to post train llm!! https://t.co/9mD

Apr 14

Articles

HOW TO GO VIRAL ON X: complete basics easy mode

Apr 13

Articles

I didn’t used to have friends btw, had to learn how to meet people like this from first principles

Apr 15

Articles

For more than 200 yrs, thousands of books & articles have been written about the history of the

May 9

Articles

Grant Sanderson, (@3Blue1Brown) created one of the most beloved math channels on the internet.

Apr 30

Articles

Wrote up some flashcards and practice problems to help myself retain what @reinerpope taught.

Apr 29

Articles

Meta researchers cut LLM RL training compute by 40% using experience replay

Apr 19

Articles

@teortaxesTex: excellent writeup

May 1

Articles

These are also the two books I recommend to people wanting a foundation in post-training. Both comin

May 8

Articles

lovely article going deeper into the RL-SFT-OPD spectrum with some very nice intuitions + experiment

May 10

Articles

New blog post! Wrote about how SFT, RL, OPD relate to generalization and catastrophic forgetting :)

May 10

Articles

How to actually land a remote job in 30 days.

Jul 6

Articles

Incredible how Z. ai literally has their RL infrastructure open source.

Jun 19

Articles

There is a really banger article on On-Policy Distillation. Came out on HF a few months back. https:

Jun 27

Articles

Songlin Yang's video explanations

Aug 3

Articles

Anthropic engineer just released a 2-hour workshop on "Graph Engineering" for agentic systems:

Jul 25

Articles

NVIDIA has lost it..

Aug 16

Articles

It’s late, so I’ll quietly give you the opening lesson from something I normally teach inside my $20

Aug 10

Articles

ravi got to know about competition from my blog and then proceeded to mog me on the leaderboard. (i

Aug 10

Articles

How does a model understand and explore a world it has never seen?

Aug 6

Articles

Reading "Hyperagents". I am so glad I created Paper Breakdown.

Mar 27

Articles

just finished a 3-hour /office-hours session using the prompt from the article:

Mar 29

Articles

there is something very beautiful about gpt-oss-120b that makes it most hackable model for numerous

Apr 3
auto_stories

Select a bookmark to read

Click any item from the list