Edited by humans. Written by AI. How our editing works

Reinforcement Learning

18 stories tagged Reinforcement Learning.

IBM Granite 4.2: Open Reasoning Models With an Agent Brain

IBM Granite 4.2: Open Reasoning Models With an Agent Brain

Yuki Okonkwo6 days ago
A large blue C logo wrapped by a green snake with code editor window, network nodes, and programming icons on a dark blue…

Building a Reinforcement Learning Library in C from Scratch

Yuki Okonkwo2 weeks ago
Older man with long gray beard wearing colorful floral shirt holds rifle with data visualization overlays, with "TRAINING…

Rich Sutton Says AI Models Have Stopped Learning

Yuki Okonkwo2 weeks ago
Three graphs showing validation loss, Pass@1, and Pass@16 metrics across model sizes, comparing compute-optimal locus with…

Joint Scaling Laws for Pre-training and RL Explained

Marcus Chen-Ramirez4 weeks ago
Green humanoid figure leaps over gray buildings while red figures lie scattered below, with "Two Minute Papers" logo in…

AI Parkour Research Solves the Imitation Problem

Bob Reynolds4 weeks ago
Man holding microphone speaking to camera with quote "Would it try to take power?" overlaid, discussing AI research findings

Can AI Do the Right Thing for the Wrong Reason?

Yuki Okonkwo1 month ago
Man in plaid shirt presenting Data Curator interface for Bespoke Labs at AI Engineer World's Fair, with "Same Question, 16…

Why AI Training Data Quality Beats Raw Compute

Bob Reynolds1 month ago
Technical architecture diagram showing neural network components including Stable LatentMoE, Gated MLA, KDA blocks, and…

Kimi K3's Post-Training Techniques Examined

Bob Reynolds1 month ago
Google AI Teaches Quantum Computers to Learn From Errors

Google AI Teaches Quantum Computers to Learn From Errors

Marcus Chen-Ramirez1 month ago
Woman presenting AI engineering concepts with pipeline architecture diagrams and performance metrics displayed behind her…

An RL Agent for ETL Pipeline Self-Healing

Dev Kapoor2 months ago
Two men face each other across a Go board with mathematical equations on a blackboard behind them, illustrating the…

AlphaGo From Scratch: What Go Teaches Modern AI

Yuki Okonkwo4 months ago
Red robotic claw bursting through white sphere battles blue claw with Chinese flag, "50x POWERFUL" text above, dramatic…

MiniMax M2.7: The AI That Trained Itself Is Now Available

Rachel "Rach" Kovacs6 months ago
Street scene with holographic AI figure and ball, NVIDIA logo above red "Decision: Stop!" neon sign, parked cars visible

NVIDIA's AI Revolutionizes Self-Driving Cars

Amelia Nwofor6 months ago
Man in red plaid shirt holding microphone against blurred blue background with text "It beat me.

A Retired Engineer Built Superhuman AI in His Garage

Samira Barnes6 months ago
Man with glasses presenting a research paper on GLM-5 training pipeline with diagrams showing pre-training, mid-training,…

GLM-5's Self-Distillation Trick Solves AI's Memory Problem

Rachel "Rach" Kovacs6 months ago
Dark digital landscape with interconnected nodes and "THE BITTER LESSON" text overlaid in yellow and white typography

AI's Bitter Lesson: Reinvention or Repetition?

Mike Sullivan7 months ago
Two men in conversation with "Poolside" branding and "Next-Gen Coding Models" text overlaid on dark background for AI…

AGI's Next Step: Poolside's Malibu Agent in Action

Mike Sullivan8 months ago
Man in black shirt against dark background with yellow and white text describing AI paradigm shifts

2025's AI Shifts: LLMs Evolve with New Paradigms

Marcus Chen-Ramirez9 months ago