- Behnam Neyshabur / Anthropic Researcher1411.We have been heads down but wanted to share a bit about what we are doing 🧵 https://t.co/ICZlLwZ3FqLLM score 75 · 4 months ago
- Andrej Karpathy / AI researcher1412.One common issue with personalization in all LLMs is how distracting memory seems to be for the models.›LLM score 92 · 4 months ago
- Merve Noyan / Hugging Face ML Engineer1413.does anyone know of a good multimodal tool calling dataset to fine-tune models on?›LLM score 92 · 4 months ago
- Ben Burtenshaw / Hugging Face Researcher1414.Give coding agent a sharable workspace with persistent storage.›LLM score 92 · 4 months ago
- Merve Noyan / Hugging Face ML Engineer1415.fav papers from 3DV #2 SAIL-Recon does reconstruction for inputs above 100 images (video)›LLM score 92 · 4 months ago
- Sander Dieleman / DeepMind Research Scientist1416.
- Leandro von Werra / Hugging Face Head of Research1417.Which LLM would be better: - today's best architecture trained on 2023's best data›LLM score 82 · 4 months ago
- Thomas Wolf / Hugging Face Cofounder1418.
- Sayak Paul / Hugging Face Researcher1419.Introducing the first discrete diffusion pipeline for text in Diffusers -- LLaDA2 by @TheInclusionAI 🔥›LLM score 75 · 4 months ago
- Demis Hassabis / CEO of DeepMind
- Naman Jain / Researcher at Cursor
- Dan Fu / VP of Kernels at Together1422.
- Niklas Muennighoff / AI Researcher at Stanford1423.One gem from Composer paper is that RL improved both pass@k & pass@1.›LLM score 92 · 4 months ago
- Cameron Wolfe / Researcher at Netflix
- Damek Davis / Assoc. Professor Wharton Stats1425.Codex took less than a week to formalize these 41000 lines and find the error.›LLM score 95 · 4 months ago
- Jim Fan / NVIDIA Director of Robotics1426.
- Jonathan Ross / TPU Creator1427.A pilot operates $100M in equipment and nobody blinks.›LLM score 82 · 4 months ago
- Leandro von Werra / Hugging Face Head of Research1428.Auto-research for ML training models is all the rage now, but underrated is: auto-research for data!›LLM score 85 · 4 months ago
- Damek Davis / Assoc. Professor Wharton Stats1429.'Proved' something new and had codex formalize it lean.›LLM score 92 · 4 months ago
- Andrej Karpathy / AI researcher1430.Software horror: litellm PyPI supply chain attack.›LLM score 97 · 4 months ago
- Jerry Tworek / ex OpenAI VP of RL1431.Todays AIs have a taste of maximally bland and median appeal RLHF.›LLM score 72 · 4 months ago
- Merve Noyan / Hugging Face ML Engineer1432.my fav papers from 3DV because why not 🤝 MapAnything by Meta›LLM score 92 · 4 months ago
- Sayak Paul / Hugging Face Researcher1433.
- Boris Cherny / Creator of Claude Code1434.Little known fact, the Anthropic Labs team (the team I joined Anthropic to be on) shipped:›LLM score 85 · 4 months ago
- Cameron Wolfe / Researcher at Netflix1435.Interesting observation on instruction following behavior induced by preference tuning versus RLVR.›LLM score 92 · 4 months ago
- Yang Chen / Nvidia Research Scientist
- Jim Fan / NVIDIA Director of Robotics
- Merve Noyan / Hugging Face ML Engineer1438."train rt-detrv2 on mobile-ui-design" is all it takes to train an object detector 🔥 ›LLM score 92 · 4 months ago
- Horace He / Thinking Machines Founding Engineer
- Jason Weston / Meta Research Scientist1440.🌐Unified Post-Training via On-Policy-Trained LM-as-RM🔧›LLM score 92 · 4 months ago