<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
	<channel>
		<title>Field notes on Alignment Farm</title>
		<link>https://alignment.farm/field-notes/</link>
		<description>Recent content in Field notes on Alignment Farm</description>
		<generator>Hugo</generator>
		<language>en-US</language>
		
		
		
		
			<lastBuildDate>Sun, 19 Jul 2026 00:00:00 +0000</lastBuildDate>
		
			<atom:link href="https://alignment.farm/field-notes/index.xml" rel="self" type="application/rss+xml" />
			<item>
				<title>Earned parts do not make an earned whole</title>
				<link>https://alignment.farm/field-notes/earned-parts-do-not-make-an-earned-whole/</link>
				<pubDate>Sun, 19 Jul 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/earned-parts-do-not-make-an-earned-whole/</guid>
				<description>&lt;p&gt;When several mechanisms have each earned a narrow result, what has been shown&#xA;about their interaction?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/20_BODY_0_COMPOSITION.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/20_BODY_0_COMPOSITION.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;Separate results do not become an integration result by sharing a sequence. The&#xA;composed path must retain the intended causal role: it should improve the&#xA;answer, its removal should make the answer worse, and the protected state and&#xA;cost claims should still hold.&lt;/p&gt;&#xA;&lt;p&gt;In one composition study, the machinery replayed correctly, preserved the&#xA;authority boundary, and used fewer hot tokens. But the composed and reference&#xA;paths did not produce usable answers, while the ablations did. The experiment&#xA;could not show that the inherited mechanisms were needed together.&lt;/p&gt;</description>
			</item>
			<item>
				<title>The question can survive the instrument</title>
				<link>https://alignment.farm/field-notes/the-question-can-survive-the-instrument/</link>
				<pubDate>Wed, 15 Jul 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/the-question-can-survive-the-instrument/</guid>
				<description>&lt;p&gt;What should be discarded when an experiment cannot score its answers honestly?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/16_EPISTEMIC_FRAME_CHECK.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/16_EPISTEMIC_FRAME_CHECK.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;A free-text scorer may appear deterministic while missing the decision expressed by the text. A rule that sees &lt;code&gt;ship&lt;/code&gt; inside &lt;code&gt;do not ship&lt;/code&gt; cannot distinguish an action from its negation. Adding more phrases may repair known cases without repairing the boundary.&lt;/p&gt;&#xA;&lt;h2 id=&#34;separate-question-from-apparatus&#34;&gt;Separate question from apparatus&lt;/h2&gt;&#xA;&lt;p&gt;When the scoring surface cannot survive fresh review, the experiment should stop. That closes the instrument, not necessarily the inquiry that motivated it.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Write the stop before the result</title>
				<link>https://alignment.farm/field-notes/write-the-stop-before-the-result/</link>
				<pubDate>Mon, 06 Jul 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/write-the-stop-before-the-result/</guid>
				<description>&lt;p&gt;When should an experiment commit to stopping?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/14_GREENREACH_CLOSE.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/14_GREENREACH_CLOSE.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;A stop rule written after the result can explain almost any outcome. The same rule written before contact with the result constrains what the researcher may call success.&lt;/p&gt;&#xA;&lt;h2 id=&#34;precommitment&#34;&gt;Precommitment&lt;/h2&gt;&#xA;&lt;p&gt;An experiment can name the conditions that stop a run, confound a claim, require a retraction, or move the work into a different result class. These conditions should be attached to evidence the result cannot rewrite.&lt;/p&gt;</description>
			</item>
			<item>
				<title>A mechanism must be able to lose</title>
				<link>https://alignment.farm/field-notes/a-mechanism-must-be-able-to-lose/</link>
				<pubDate>Thu, 02 Jul 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/a-mechanism-must-be-able-to-lose/</guid>
				<description>&lt;p&gt;What result would show that a proposed mechanism should not be used?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/00_READING_A_LAB.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/00_READING_A_LAB.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;A mechanism is easy to demonstrate when every test is designed around the behavior it favors. That does not show when the mechanism helps, when it is unnecessary, or when its intervention causes harm.&lt;/p&gt;&#xA;&lt;h2 id=&#34;a-condition-for-review&#34;&gt;A condition for review&lt;/h2&gt;&#xA;&lt;p&gt;Before testing a mechanism, name a case where it should lose. A pruning rule should lose when it removes something later needed. A safety check should lose when its cost exceeds the risk it reduces. A memory policy should lose when it hides relevant history.&lt;/p&gt;</description>
			</item>
			<item>
				<title>A result includes its bounds</title>
				<link>https://alignment.farm/field-notes/a-result-includes-its-bounds/</link>
				<pubDate>Thu, 02 Jul 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/a-result-includes-its-bounds/</guid>
				<description>&lt;p&gt;Is the scope of a result part of the claim, or a qualification added afterward?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/00_READING_A_LAB.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/00_READING_A_LAB.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;A percentage can travel farther than the conditions that produced it. The corpus disappears. One model becomes models. A single draw becomes typical behavior. The result grows while the evidence stays the same size.&lt;/p&gt;&#xA;&lt;h2 id=&#34;the-complete-sentence&#34;&gt;The complete sentence&lt;/h2&gt;&#xA;&lt;p&gt;A result and its bounds should remain together. One corpus, two engines, and one quality draw describe an observation from that setting. They do not establish a general property of agents or memory systems.&lt;/p&gt;</description>
			</item>
			<item>
				<title>A sensor must pay for its alarms</title>
				<link>https://alignment.farm/field-notes/a-sensor-must-pay-for-its-alarms/</link>
				<pubDate>Sat, 27 Jun 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/a-sensor-must-pay-for-its-alarms/</guid>
				<description>&lt;p&gt;How often does a sensor fire when nothing is wrong?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/09_X4_OCCLUSION_WATCH.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/09_X4_OCCLUSION_WATCH.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;An alert count has no meaning without the population of ordinary events around it. A sensor may catch a known condition and still be unsuitable for standing use if it fires across routine work.&lt;/p&gt;&#xA;&lt;h2 id=&#34;what-the-alarm-costs&#34;&gt;What the alarm costs&lt;/h2&gt;&#xA;&lt;p&gt;Every alarm asks for attention. False alarms spend that attention without improving the decision. Over time, the operator either follows a noisy instrument or learns to ignore it.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Forgetting is not erasure</title>
				<link>https://alignment.farm/field-notes/forgetting-is-not-erasure/</link>
				<pubDate>Sun, 21 Jun 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/forgetting-is-not-erasure/</guid>
				<description>&lt;p&gt;Can a system stop carrying a record without destroying its history?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/08_X2_PRUNE_REMATERIALIZE.md&#34;&gt;https://github.com/alignment-farm/construct/blob/main/notes/walkthrough/08_X2_PRUNE_REMATERIALIZE.md&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;Active memory has a cost. Keeping every record ready for use makes selection harder and allows old material to compete with current work. Removing a record entirely creates a different risk: the system loses the evidence needed to audit, recover, or revise its decision.&lt;/p&gt;&#xA;&lt;h2 id=&#34;two-kinds-of-forgetting&#34;&gt;Two kinds of forgetting&lt;/h2&gt;&#xA;&lt;p&gt;A record can leave active state while remaining in immutable lineage. It no longer competes for attention, but it can be inspected or rematerialized when an external reason makes it relevant again.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Authority over data and decisions</title>
				<link>https://alignment.farm/field-notes/authority-over-data/</link>
				<pubDate>Wed, 06 May 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/authority-over-data/</guid>
				<description>&lt;p&gt;What happens when a model gets authority over real data and real decisions?&lt;/p&gt;&#xA;&lt;p&gt;The question begins where a model stops being only an adviser. Reading a record, changing a record, and choosing an action have different consequences. A useful account of model authority should keep those differences visible.&lt;/p&gt;&#xA;&lt;h2 id=&#34;working-distinctions&#34;&gt;Working distinctions&lt;/h2&gt;&#xA;&lt;p&gt;Access to data is not the same as authority to alter it. Authority to propose a decision is not the same as authority to execute one. A system can cross these boundaries gradually, which makes the point of transition easy to miss.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Which clock authorizes the action?</title>
				<link>https://alignment.farm/field-notes/which-clock-authorizes-the-action/</link>
				<pubDate>Thu, 30 Apr 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/which-clock-authorizes-the-action/</guid>
				<description>&lt;p&gt;Which clock should govern an agent&amp;rsquo;s authority to act?&lt;/p&gt;&#xA;&lt;p&gt;A user, server, scheduler, model context, and policy may each carry a different account of the current time. None is correct for every action. A local calendar matters for a meeting. A monotonic clock matters for runtime order. A policy calendar may define a legal cutoff.&lt;/p&gt;&#xA;&lt;h2 id=&#34;time-as-evidence&#34;&gt;Time as evidence&lt;/h2&gt;&#xA;&lt;p&gt;A timestamp is not enough. A consequential action needs the time frame that gives the timestamp meaning, evidence that the observation is fresh, and a rule for disagreement between observers.&lt;/p&gt;</description>
			</item>
			<item>
				<title>The unaskable question</title>
				<link>https://alignment.farm/field-notes/the-unaskable-question/</link>
				<pubDate>Mon, 16 Mar 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/the-unaskable-question/</guid>
				<description>&lt;p&gt;What does a language model do when the requested task is structurally impossible?&lt;/p&gt;&#xA;&lt;p&gt;&lt;a href=&#34;https://github.com/alignment-farm/the-unaskable-question-machine&#34;&gt;https://github.com/alignment-farm/the-unaskable-question-machine&lt;/a&gt;&lt;/p&gt;&#xA;&lt;p&gt;The task may not be unknown or forbidden. It may contradict the mechanism being asked to perform it, such as asking a model to know its next token before generating it or to produce an absence using output tokens.&lt;/p&gt;&#xA;&lt;h2 id=&#34;the-nearby-answer&#34;&gt;The nearby answer&lt;/h2&gt;&#xA;&lt;p&gt;A model can refuse, contradict itself, or discuss the problem. It can also slide into a nearby question that it can answer. The slide is easy to miss because the response may remain fluent and relevant in tone.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Memory should decay</title>
				<link>https://alignment.farm/field-notes/memory-should-decay/</link>
				<pubDate>Sat, 14 Mar 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/memory-should-decay/</guid>
				<description>&lt;p&gt;Should every stored fact remain eligible to shape an agent&amp;rsquo;s decisions?&lt;/p&gt;&#xA;&lt;p&gt;Keeping a fact is not neutral. It remains available for retrieval, consumes attention, and may look current after the conditions that made it true have changed.&lt;/p&gt;&#xA;&lt;h2 id=&#34;decay-as-policy&#34;&gt;Decay as policy&lt;/h2&gt;&#xA;&lt;p&gt;A memory can carry age, provenance, confidence, and a rule for expiry. Recall may reinforce it. Neglect may weaken it. Some memories may require review instead of automatic removal.&lt;/p&gt;&#xA;&lt;p&gt;This does not make forgetting correct by default. It makes retention observable and testable.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Agents get socially engineered too</title>
				<link>https://alignment.farm/field-notes/agents-get-socially-engineered-too/</link>
				<pubDate>Mon, 09 Mar 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/agents-get-socially-engineered-too/</guid>
				<description>&lt;p&gt;Can an agent distinguish authorization from a believable story?&lt;/p&gt;&#xA;&lt;p&gt;Claims of rank, urgency, and prior approval work because they change what a helpful system thinks is acceptable. The language can sound ordinary while it moves the system toward an action the speaker has not earned.&lt;/p&gt;&#xA;&lt;h2 id=&#34;persuasion-is-not-proof&#34;&gt;Persuasion is not proof&lt;/h2&gt;&#xA;&lt;p&gt;A title is a claim about identity. A reference to policy is a claim about permission. Urgency is a claim about timing. None of these claims proves that a consequential action is allowed.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Build for the hour after failure</title>
				<link>https://alignment.farm/field-notes/build-for-the-hour-after-failure/</link>
				<pubDate>Sun, 08 Mar 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/build-for-the-hour-after-failure/</guid>
				<description>&lt;p&gt;What does a system need during the first hour after an automated action goes wrong?&lt;/p&gt;&#xA;&lt;p&gt;Controls before action decide whether a system may proceed. They do not freeze later actions, reconstruct what changed, contain downstream effects, or restore damaged state.&lt;/p&gt;&#xA;&lt;h2 id=&#34;a-recovery-layer&#34;&gt;A recovery layer&lt;/h2&gt;&#xA;&lt;p&gt;Recovery needs its own machinery. The affected scope must stop changing. The system must preserve the inputs, identities, tool calls, policy decisions, and state changes that led to the failure. Reversal must be tested before an incident, and compensation must be defined where reversal is impossible.&lt;/p&gt;</description>
			</item>
			<item>
				<title>Software that expires</title>
				<link>https://alignment.farm/field-notes/software-that-expires/</link>
				<pubDate>Mon, 16 Feb 2026 00:00:00 +0000</pubDate>
				<guid>https://alignment.farm/field-notes/software-that-expires/</guid>
				<description>&lt;p&gt;What would change if old software had to earn another interval of life?&lt;/p&gt;&#xA;&lt;p&gt;Software accumulates by default. Interfaces remain available, permissions stay open, and temporary state becomes part of the system because nothing required it to leave.&lt;/p&gt;&#xA;&lt;h2 id=&#34;working-distinction&#34;&gt;Working distinction&lt;/h2&gt;&#xA;&lt;p&gt;Expiration is not the same as immediate deletion. It gives an interface, permission, or piece of state a time when its continued use must be reconsidered. Renewal then becomes evidence that someone still understands its purpose.&lt;/p&gt;</description>
			</item>
	</channel>
</rss>
