Expand ↗
Page list (1404)

FTRL

Follow-The-Regularized-Leader — a family of online-learning algorithms that, each round, play the action minimising total past loss plus a regularisation term. With suitable regularisation FTRL achieves sublinear regret (No-Regret Learning), and it generalises and stabilises Follow-The-Leader. The benchmark against which “Do LLM Agents Have Regret?” measures whether language-model agents behave as no-regret learners in repeated games.

In this vault

Backlinks