Back
Models6/day by emailPrimary

CL-Bench, a new benchmark testing whether AI agents actually improve with experience across sequential real-world tasks, found that even the best system -- Claude Sonnet 4.6 using full-context in-context learning -- achieved only a 25.4% normalized gain over its own stateless baseline Best AI agent's continual-learning gain on CL-Bench: 25.4%

Verified 2026-09-23

THE INTELLIGENCE CLUB

Facts like this, six days a week

The Daily Receipts on weekdays, The Weekly Brief on Saturday — six sourced facts a day by email, two more than the free site shows. Every one traced to the filing it came from. Free.

Free forever. Six emails a week — The Daily Receipts on weekdays, The Weekly Brief on Saturday — drop either in one click.

Read this morning's edition →