- Thomas Wolf / Hugging Face Cofounder2161.who's doing serious ai-thropology research on @moltbook rn?›LLM score 72 · 6 months ago
- Lucas Beyer / Meta Researcher2162.PSA: never, ever write "we use the same learning rate across all methods for fair comparison"›LLM score 85 · 6 months ago
- Andrew White / Edison Scientific Cofounder2163.Another nice example showing how our agents can reproduce analysis and figures from papers.›LLM score 92 · 6 months ago
- Ben Burtenshaw / Hugging Face Researcher2164.PSA: skills are not docs. skills are for the hardest problems an agent can solve.›LLM score 65 · 6 months ago
- alphaXiv
- Jason Weston / Meta Research Scientist2166.📈Self-Improving Pretraining 📈 ✍️: https://t.co/GsvYMuMT4b›LLM score 92 · 6 months ago
- Cyris Kissane / Researcher at Flapping Airplanes2167.Hot take: the creation of Adam put AI research back multiple years.›LLM score 65 · 6 months ago
- Zhaocheng Zhu / Nvidia Research Scientist2168.ICML bidding observations: LLMs are emerging as a field separate from deep learning.›LLM score 80 · 6 months ago
- Hang Gao / ex MTS at xAI2169.
- Sherwin Wu / OpenAI API, Head of Engineering
- Omar Khattab / MIT CSAIL Asst professor2171.New updates for the RLM paper: We post-trained RLM-Qwen3-8B at tiny scale, the first natively recursive LM.›LLM score 85 · 6 months ago
- Alex Zhang / MIT CSAIL PhD2172.We just updated the RLM paper with some new stuff.›LLM score 85 · 6 months ago
- alphaXiv2173.2026 is the year of continual learning And we are getting some amazing papers towards that›LLM score 85 · 6 months ago
- Andrew Lampinen / Research Scientist at DeepMind2174.
- Sayak Paul / Hugging Face Researcher
- Cameron Wolfe / Researcher at Netflix2176.Trinity large is very sparse (400B-A13B, 256 experts w/ 4 active per token).›LLM score 85 · 6 months ago
- Roland Gavrilescu / ex MTS at xAI2177.Models haven’t been post-trained on progressive disclosure yet.›LLM score 92 · 6 months ago
- John Carmack2178.#PaperADay 13 2020: DREAM TO CONTROL: LEARNING BEHAVIORS BY LATENT IMAGINATION›LLM score 85 · 6 months ago
- Asher Spector / Cofounder of Flapping Airplanes
- Cyris Kissane / Researcher at Flapping Airplanes2180.I use muon when making decisions to minimize my regret Adam and SGD just weren't fast enough.LLM score 70 · 6 months ago
- Kushal Thaman / Researcher at Flapping Airplanes2181.I spent a bunch of time a year ago thinking about the data wall.›LLM score 85 · 6 months ago
- Mehtaab Sawhney / OpenAI for Science2182.I've recently gone on leave from Columbia to join OpenAI, working on OpenAI for Science.›LLM score 75 · 6 months ago
- Ben Burtenshaw / Hugging Face Researcher2183.We got Claude to teach open models how to write CUDA kernels.›LLM score 85 · 6 months ago
- alphaXiv2184.BIG new idea in interpretability called Patterning›LLM score 75 · 6 months ago
- John Carmack2185.#PaperADay 12 2019: Learning Latent Dynamics for Planning from Pixels (PlaNet)›LLM score 85 · 6 months ago
- Boris Cherny / Creator of Claude Code
- Lucas Beyer / Meta Researcher
- alphaXiv2188."LLM-in-Sandbox Elicits General Agentic Intelligence"›LLM score 85 · 6 months ago
- Ethan Shen / Ai2 Researcher2189.
- Lucas Beyer / Meta Researcher