Category: Machine Learning
- Terence Tao Uses Claude Code to Solve Problems, Crashes Twice Due to Running Out of Tokens
- Latest! Karpathy's 10,000-Word In-Depth Interview: My Anxiety Became an AI Addiction—All Verifiable Domains Will Eventually Belong to Machines
- OpenAI is throwing everything into building a fully automated researcher
- Golden Rules for Skill Development! Google Releases 5 Agent Skill Design Patterns
- Performance Surges 42%! Renmin University & ByteDance Open-Source Scale-SWE, a 100k-Level SWE Dataset
- Mamba-3
- Someone Actually Built the AI Research Community from Karpathy's Side Project...
- OpenClaw-RL: Allowing AI Agents to Self-Evolve Through Chat
- a16z: Agents Perform Poorly Due to Lack of Correct Data Context
- 4B Model Surpasses GPT-5 in Hallucination Suppression: CMU and Others Propose New Behaviorally Calibrated Reinforcement Learning Method
- Jensen Huang Enters the OpenClaw Arena! Most Powerful Open-Source 'Lobster' Model Rivals Opus 4.6
- Masterpiece! MIT and Google Train an LLM Capable of Rigorous Bayesian Inference
- Karpathy Slept, AI Ran 100 Experiments for Him
- MMLU is Dead? 'Humanity's Last Exam' Published in Nature: Global AI Models Collectively Fail!
- Anthropic CEO: The Data Bottleneck for Large Models No Longer Exists, Models Are Training Themselves
- Google AI Conquers 6 World-Class Problems, More Shocking Than an IMO Gold Medal! Terence Tao Points the Way to a New Game
- OpenAI Legend Reveals: Undergraduate Lands Job at OpenAI with Just One Blog Post! No PhD, Zero Papers
- Are LLM RL Training Trajectories Actually Linear? Miaow Lab's Latest Work: Directly 'Predict' Future Models Without Further Training!
- Qwen3.5: Towards Native Multimodal Agents
- Xiaomi Introduces JudgeRLVR: Judge First, Generate Second — Breaking the Efficiency Paradox of "Long Chain-of-Thought" in Reasoning Models
- Stable-DiffCoder Surpasses Autoregressive Models! New Breakthrough in Code Generation with Diffusion Models
- Stop Clipping Aggressively! Qwen Proposes GatedNorm, Unifying the Perspective on Residual Flow Mysteries
- Less is More: Recursive Reasoning with Tiny Networks
- GPT-5.3-Codex Released: The First Self-Training Model
- Is PPO Dead? The Reinforcement Learning Foundation Used by DeepSeek Has Major Flaws!