- Andrej Karpathy / AI researcher2461.@Jess_Riedel @goakhmad Yeah I think granted, it’s why I call them prompts.›LLM score 30 · 9 months ago
- Andrej Karpathy / AI researcher
- Andrej Karpathy / AI researcher
- Xiang Fu / Researcher at Periodic Labs2464.@khoomeik Checkout old PGM works if this is your thing›LLM score 20 · 9 months ago
- Xiang Fu / Researcher at Periodic Labs2465.Scaling laws are often presented as a story of data, model size, and compute.›LLM score 80 · 9 months ago
- Emmanuel Ameisen / Anthropic Interpretability Researcher2466.
- Andrej Karpathy / AI researcher
- Niklas Muennighoff / AI Researcher at Stanford
- Tri Dao / Chief Scientist at Together2469.Tons of effort from IBM and vLLM folks to make these hybrid models go fast.›LLM score 20 · 9 months ago
- Tri Dao / Chief Scientist at Together
- Richard Song / DeepMind Research Scientist2471.Reliable performance is crucial for doing great math.›LLM score 85 · 9 months ago
- Noam Brown / OpenAI Research Scientist
- Noam Brown / OpenAI Research Scientist2473.In 2019 @hughbzhang sent me a detailed personalized cold email asking to intern with me.›LLM score 20 · 9 months ago
- Noam Brown / OpenAI Research Scientist
- Naman Jain / Researcher at Cursor
- Andrej Karpathy / AI researcher
- Emmanuel Ameisen / Anthropic Interpretability Researcher2477.
- Hang Gao / ex MTS at xAI2478.The value of eval lies in a distributional look at the model performance.›LLM score 92 · 9 months ago
- Victoria Lin / Thinking Machines Researcher
- Xiang Fu / Researcher at Periodic Labs2480.Second @simonbatzner. FF architecture was no longer a bottleneck for AI materials discovery since nequip.›LLM score 80 · 9 months ago
- Noam Brown / OpenAI Research Scientist2481.@lineardiff @ericzelikman I don’t view that as incompatible.›LLM score 20 · 9 months ago
- Emmanuel Ameisen / Anthropic Interpretability Researcher2482.Striking result, which changed how I think about LLMs:›LLM score 85 · 9 months ago
- Niklas Muennighoff / AI Researcher at Stanford
- Niklas Muennighoff / AI Researcher at Stanford2484.To scale data-constrained LLMs, repeating & denoising objectives can help.›LLM score 85 · 9 months ago
- Tri Dao / Chief Scientist at Together2485.State space architecture for state of the art voice model! https://t.co/QjdOdrIFSMLLM score 80 · 9 months ago
- Emmanuel Ameisen / Anthropic Interpretability Researcher2486.Cool paper using attribution graphs to automatically detect which reasoning steps contain mistakes!›LLM score 85 · 9 months ago
- Lilian Weng / Thinking Machines Cofounder
- Mira Murati / Thinking Machines CEO
- Xiang Fu / Researcher at Periodic Labs2489.To adapt to today's world, crank down discount factor and crank up learning rateLLM score 75 · 9 months ago
- John Schulman2490.