aisilab/moltbook-files
Viewer • Updated • 232k • 67
Interpretability-informed control
Selecting The Most Informative Tokens in Natural Language Autoencoders
Emergent Languages in Populations of Language Model Agents: From Token Efficiency to Oversight Evasion