Category & trade
Primary
Tree path
Technology & Computing › Artificial Intelligence
Group (tier-1)
Tech stack
AI readiness
AI-Readiness score
45 / 100 · partial · see similar
AI training policy
allowed
AI-bot protection
AI files
Evidence
llms.txt
Compliance (GEO / GDPR)
TLD
Overview
Title
How’s it going? Reinforcement learning in language models recruits a functional welfare axis
Description
How does reinforcement learning shape a language model’s internal representations? We present evidence that RL recruits a pre-existing representation of functional welfare: an estimate of how well or badly the system is doing, relative to its goals. We train several language models in a novel, semantically neutral maze environment. We then extract concept vectors for rewarded and punished trajectories, and evaluate those vectors in settings unrelated to the maze environment. The punishment vec
Final URL
Language
en (html)
Scanned at
2026-07-24 03:30:01
Embed AI-status badge
Show this site's AI-access status with a free, live badge — it reflects AI-Accessible and updates automatically as the site changes.
Embed code (HTML)