- Thomas Wolf / Hugging Face Cofounder1261.json is so token inefficient it hurts these days man, these braces and quotes are costing me real $$LLM score 82 · 4 months ago
- Lewis Tunstall / Hugging Face Researcher1262.We've rebuilt TRL's on-policy distillation trainer from the ground up to:›LLM score 92 · 4 months ago
- Merve Noyan / Hugging Face ML Engineer1263.MiniMax 2.7 is out 🔥 it sits in the open frontier in score per token efficiency 🥵 https://t.co/izzcB1I3vfLLM score 15 · 4 months ago
- Alex Zhang / MIT CSAIL PhD1264.
- Eric Jang / ex VP of AI at 1X Robotics1265.just tried this with Claude code + computer use mcp on my Mac.›LLM score 92 · 4 months ago
- Skyler Miao / MiniMax Head of Engineering1266.M2.7 weights are live. hope you all enjoy it 😎 been grinding sleepless on M3 and the harness›LLM score 82 · 4 months ago
- Jerry Tworek / ex OpenAI VP of RL1267.
- Niklas Muennighoff / AI Researcher at Stanford1268.There's a wave of omni embedding models (gemini, nemotron, bidirlm).›LLM score 82 · 4 months ago
- Alex Zhang / MIT CSAIL PhD1269.this is a sick idea applying a paper I think is very cool (attention matching) to RLMs ›LLM score 92 · 4 months ago
- Ben Burtenshaw / Hugging Face Researcher
- Noam Brown / OpenAI Research Scientist1271.What we really need is a benchmark where AI models make AI models that play poker.›LLM score 92 · 4 months ago
- Lewis Tunstall / Hugging Face Researcher
- Sayak Paul / Hugging Face Researcher1273.
- Alex Zhang / MIT CSAIL PhD1274.The "Mismanaged Geniuses" Hypothesis(blog)LLM score 100 · 4 months ago
- Cameron Wolfe / Researcher at Netflix1275.Can't wait to get my physical copy! Definitely order the RLHF book.›LLM score 92 · 4 months ago
- Zixuan Li / Lead Z.ai1276.Often discuss my three-level vision for opening GLM to the community:›LLM score 92 · 4 months ago
- John Carmack1277.Making a scatter plot of 400_000 data points, some of the plots had odd gaps in coverage.›LLM score 92 · 4 months ago
- Boris Cherny / Creator of Claude Code1278.Just got a nice DM from a big enterprise customer using Claude Code in one of the world's biggest codebases›LLM score 95 · 4 months ago
- Alex Zhang / MIT CSAIL PhD1279.
- Andrej Karpathy / AI researcher1280.Judging by my tl there is a growing gap in understanding of AI capability.›LLM score 92 · 4 months ago
- Eric Jang / ex VP of AI at 1X Robotics1281.
- Niklas Muennighoff / AI Researcher at Stanford
- Shuchao Bi / Meta Researcher
- Noam Brown / OpenAI Research Scientist1284.
- Cameron Wolfe / Researcher at Netflix1285.My reaction to muse spark is similar to how I felt about llama 4.›LLM score 85 · 4 months ago
- Sebastien Bubeck / OpenAI MTS1286.The world of mathematics is rapidly changing.›LLM score 92 · 4 months ago
- Mehtaab Sawhney / OpenAI for Science
- Shuchao Bi / Meta Researcher1288.the model is incredible at generating playable mini-games.›LLM score 92 · 4 months ago
- Alex Zhang / MIT CSAIL PhD1289.but I thought Claude Code and RLMs were the same thing ›LLM score 85 · 4 months ago
- Yang Chen / Nvidia Research Scientist