NEWS // Latest ActivityTOTAL: 019

Structured Uncertainty Guides LLM Agents for Efficient Tool-Calling Disambiguation

Google DeepMind Partners with EVE Online to Train AI Agents in Complex Worlds

HACRL: Collaborative Reinforcement Learning for Heterogeneous AI Agents

OpenAI Buys Tens of Thousands of Macs for Agent Reinforcement Learning

ChatGPT's Unlikely Goblin Obsession: How OpenAI's 'Nerdy' Personality Trait Backfired and Led to Creature Mentions

Inside Koray Kavukcuoglu: The New Operations Mastermind of Google DeepMind

Tsinghua PhD Raises Millions for Nexus: The AI-Native Team Collaboration OS

The Rise of AI Philanthropy: Why Tech Pioneers Are Pledging Their Fortunes

New Research Quantifies User Simulator Utility for Better LLM Assistant Performance

Advancing AI Capabilities: Exploring Autonomous Chain Replication

DeepSeek-R1: Unlocking LLM Reasoning Through Reinforcement Learning

DFPO: A Novel Distributional RL Framework Enhances LLM Post-Training with Value Flow Modeling for Robustness

The Dawn of Autonomous AI: Five Pivotal Technologies Enabling Machines to Learn Without Human Intervention

Kuaishou's GR4AD Generative Recommender System Boosts Ad Revenue by 4.2% and Serves Over 400 Million Users

Atari's Tycoon CEO Wade Rosen Targets a New Retro Gaming Revival

IMAgent: Multi-Image Vision Agent Achieves SOTA with End-to-End Reinforcement Learning

AReaL 2.0 Released: Open-Source RL Infrastructure for Self-Evolving AI Agents

OpenAI Explains "Goblin" References in Its AI Models: An Unintended Consequence of Reinforcement Learning

OpenAI RL Lead Dan Roberts on Test-Time Compute, o1, and the Science of RL