Piperic
domain profile
‹ ListPricingToplistsBrowse free →

Category & trade

Primary
Tree path
Technology & Computing › Artificial Intelligence
Group (tier-1)

Tech stack

AI readiness

AI-Readiness score
45 / 100 · partial · see similar
AI training policy
allowed
AI-bot protection
AI files
llms.txt ai.txt humans.txt robots.txt
Evidence
llms.txt

Compliance (GEO / GDPR)

TLD

Overview

Title
How’s it going? Reinforcement learning in language models recruits a functional welfare axis
Description
How does reinforcement learning shape a language model’s internal representations? We present evidence that RL recruits a pre-existing representation of functional welfare: an estimate of how well or badly the system is doing, relative to its goals. We train several language models in a novel, semantically neutral maze environment. We then extract concept vectors for rewarded and punished trajectories, and evaluate those vectors in settings unrelated to the maze environment. The punishment vec
Final URL
Language
en (html)
Scanned at
2026-07-24 03:30:01

Embed AI-status badge

Show this site's AI-access status with a free, live badge — it reflects AI-Accessible and updates automatically as the site changes.

AI access badge preview
Embed code (HTML)