- Sholto Douglas / Researcher at Anthropic571.we don’t even run evals anymore we just ask Claude what the score will be https://t.co/1VXOEUsVZQLLM score 35 · about 2 months ago
- Andrej Karpathy / AI researcher572.This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards.›LLM score 65 · about 2 months ago
- Boris Cherny / Creator of Claude Code573.Fable 5 is now available in Claude Code and Cowork›LLM score 72 · about 2 months ago
- Jeff Dean / Chief Scientist at DeepMind574.Speech translation has been one of the longest-running ML efforts at Google, and we’ve come a long way.›LLM score 65 · about 2 months ago
- Andy Jones / Anthropic Research Engineer575.did you know? you can just ask fable what its benchmark score will be https://t.co/3bWG6Oy2ffLLM score 25 · about 2 months ago
- Sholto Douglas / Researcher at Anthropic
- Fei-Fei Li
- Lewis Tunstall / Hugging Face Researcher578.We're running the Fast Gemma Challenge: make gemma-4-E4B go brrr on a single A10G, without wrecking quality ⚡️!›LLM score 65 · about 2 months ago
- Jason Weston / Meta Research Scientist
- Leandro von Werra / Hugging Face Head of Research
- Noam Brown / OpenAI Research Scientist581.We've known about LLM test-time compute scaling since @OpenAI o1.›LLM score 72 · about 2 months ago
- Ben Burtenshaw / Hugging Face Researcher582.Yesterday, the open source community is backed OpenEnv for agentic RL.›LLM score 35 · about 2 months ago
- Julian Schrittwieser / Anthropic Researcher583.Plotting benchmark results with inference cost on the x-axis is absolutely the right thing to do, great writeup by @polynoamial !›LLM score 65 · about 2 months ago
- Boris Cherny / Creator of Claude Code584.Just landed nested subagent support in Claude Code›LLM score 65 · about 2 months ago
- Leandro von Werra / Hugging Face Head of Research585.Deep dive into FNS: building a tokenizer that chunks text efficiently but has character level resolution!›LLM score 72 · about 2 months ago
- Andrew Ma / ex MTS at xAI586.feishu is notion if everyone decided in the first place to use notion to manage work collaboration not slackLLM score 15 · about 2 months ago
- Keller Jordan / OpenAI Researcher587.Thank you for this result! Here's one initial correction:›LLM score 65 · about 2 months ago
- Julian Schrittwieser / Anthropic Researcher588.LLMs are increasingly good at restating obtuse academic writing in easily accessible language!›LLM score 65 · about 2 months ago
- Keller Jordan / OpenAI Researcher
- Eric Jang / ex VP of AI at 1X Robotics590.my stock picking strategy of late is to monitor my friends (SWEs not in AI) group chat.›LLM score 15 · about 2 months ago
- Noam Brown / OpenAI Research Scientist591.Implications of Large-Scale Test-Time Compute(blog)LLM score 85 · about 2 months ago
- Eric Jang / ex VP of AI at 1X Robotics592.In time, people will come to understand "why humanoid" https://t.co/d1X5TlVcF7LLM score 45 · about 2 months ago
- Asher Spector / Cofounder of Flapping Airplanes593.great post, couldn't agree more :) https://t.co/xYsbhRzUeZLLM score 15 · about 2 months ago
- John Carmack594.I admire Fabrice Bellard. He is almost certainly a better overall programmer than I am.›LLM score 25 · about 2 months ago
- Jakub Pachocki / OpenAI Chief Scientist595.The north stars we're working towards at OpenAI all center around the mission: ensure AGI benefits all of humanity.›LLM score 35 · about 2 months ago
- Keller Jordan / OpenAI Researcher596.I've added two optimizers to the public benchmark:›LLM score 72 · about 2 months ago
- Sherwin Wu / OpenAI API, Head of Engineering597.The Product Design Codex plugin works best with Figma! Instantly rev on a design and then export it to your Figma canvas.›LLM score 25 · about 2 months ago
- Boris Cherny / Creator of Claude Code598.When we first demoed Claude Code internally, it got two reactions on Slack.›LLM score 72 · about 2 months ago
- Thomas Wolf / Hugging Face Cofounder
- Thomas Wolf / Hugging Face Cofounder