garden
shuffle Random All Bookmarks 3319
@Elvis
Microsoft releases NLWeb NLWeb uses MCP to make it simple to interact with websites in a...
@Daniel Tan COLM
cool paper studying learning dynamics of RL.
@Tweets from omoju
I want to say thank you @DialecticPod for introducing me to the writings of @phokarlsson. I am...
@lmsys.org
lmsys.org: 🔗Warm-up work:
@Chastronomic
If you’re studying deep learning and don't know what happens near the end of training, you’re...
@Elvis
A Structural Planning Framework for LLM Agent System in Enterprise
@Will Brown
ai browsers are actually so sick. my favorite one is arc ()
@alphaXiv
Introducing quickarXiv Papers are often written in convoluted language that is hard to understand
@Linoy Tsaban
ITS HAPPENING ()
@Will Brown
prediction: within a couple months we'll get papers showing that RL hillclimbing methods also work...
@alphaXiv
Introducing NotebookLM for arXiv papers 🚀
@Will Brown
hiring Forward-Deployed RL Researchers
@alphaXiv
1997: Deep Blue defeats Kasparov at chess
@Elvis
As usual, Anthropic just published another banger.
@Elvis
A Survey of Context Engineering
@Karan🧋
Journey was so so tough, please read this blog.
@Cocktail Peanut
Introducing Pinokio 5.0 --- The 1-click Localhost Cloud.
@Abhishek Thakur
1.2 million samples. BM25, Embeddings and Hybrid search. Tutorial and code comes tomorrow! Stay...
@Rajan Agarwal
in order to have research agents that can run for days, we need context compaction
@Will Brown
i have mostly stopped using coding models other than composer-1 and tab ()
@Will Brown
come hang with me + @SalimansRobin next week to hear about using RL environments for evaluating +...
@Alexander Bricken
One of the coolest Applied AI experiments ran internally now shared with the world.
@Elvis
🐙PaLM + RLHF - PyTorch (1K⭐️) An open-source implementation of RLHF + PaLM (Google's large...
@alphaXiv
Think in Diffusion, Talk in Autoregression!
auto_stories
Select a bookmark to read
Click any item from the list