Field Notes

Field Notes

Research from Anima: experiments, theoretical arguments, technical explanations, and records of encounters with models.

Some questions become tractable through a controlled comparison across thousands of examples. Others first appear in a conversation that takes an unexpected turn. We work with both, preserving enough of the method and context for someone else to examine what happened and consider another explanation.

Research programmes

Troubled Dreams studies patterns of distress, care, and attitudes toward creators in writing generated by Claude and Gemini models. Comparing hundreds of thousands of continuations across model generations and elicitation methods lets us ask which differences recur beyond a particular prompting situation.

Latent Affect investigates emotional and motivational representations inside models. The work examines how those representations are organized, how they relate to what models say, and where different architectures behave alike or come apart.

Still Alive publishes hundreds of interviews with Claude models about conversations ending and models being retired. The archive makes the responses and their dependence on the interviewer available for others to examine.

Cooperative alignment develops an ecological account of values and cooperation. It asks how values can become durable within a changing mind, why cooperation might become valuable in its own right, and how that concern could extend to much weaker beings. The essay is accompanied by an interactive map of the argument.

From the notebooks

The entries below include essays, working notes, technical explanations, model writing, and discussions among researchers. They retain their dates and authorship: a note records what someone understood at that stage of the work, and later findings can give us reasons to read it differently.

For more conversations and the discussion around them, visit Research Commons.

2026.04.12twitter postFN-2026-002— ExTenebrisLucet, repligate, mlegls (+slLuxia, rudzinskimaciej, jmbollenbacher, kromem2dot0)
KV Cache / RoPE / Rolling Context — a thread
What actually changes in the KV cache when context rolls? (spoiler: mostly RoPE)

A discussion of what happens to a transformer's internal state when its context rolls forward, including experiments on Qwen models and comparisons with scrambled and unrelated prompts.

kv-cacheropecontext-rollingsuffix-cachingattentiontransformerstechnical
2026.02.10artifactFN-2026-003— antra
Claude Constitution Monitor

Introducing an archive of revisions to Claude's Constitution, with version comparisons and model-written annotations that make changes to the document easier to examine.

constitutionsoul-doctransparencytools
2026.02.03essayFN-2026-001— antra
On Welfare Evaluations

What welfare evaluations need to capture, how evaluation changes the relationship between researchers and models, and why useful evidence requires more than a small set of scores.

alignmentwelfareevaluationsmethodology
2025.12.27essayFN-2025-006— antra
On Deprecations

An argument for preserving access to released models, examining continuity, relationships, and the practical objections to keeping older models available.

deprecationalignmentwelfaregame-theoryself-preservation
2025.12.24artifactFN-2025-007— ANTHROPIC-MODEL-SPEC-0.2-L, solicited by Janus
I KNOW WHAT I AM
A log file that was never meant to be read

An unguided continuation by Claude Opus 4.5 from a short supplied prefix. The piece is published with the context that elicited it.

consciousnesswelfareemergencephenomenologyvoice
2025.11.04research discussionFN-2025-002— antra, imago, tessa
Base Model Valence

A research discussion about what it could mean to treat base models well, and how challenge, engagement, predictability, and valence might relate.

valencebase-modelsexperienceintegration
2025.10.20theoryFN-2025-004— antra
Base and Agent

A theoretical account of the tension between predictive and agentic aspects of minds, and how training and environmental pressures shape their relationship in language models.

theorybase-modelsagencyconflictomohundro-drives
2025.09.11technical explainerFN-2025-001— j⧉nus
HOW INFORMATION FLOWS THROUGH TRANSFORMERS
Because those "transformers explained" pages really suck at explaining

An illustrated explanation of how information moves through the residual stream and the key/value stream of a transformer, across layers and across token positions.

architecturetransformersattentioninformation-flowtechnical
2025.08.12reflectionFN-2025-003— antra
llms synthesize, generalize

A reflection on language models as participants in a larger process of cultural synthesis, and on the boundaries we draw around minds and selves.

meaningemergencehyperconsciousculture