The interesting artifact is the making.
Stillwet asks language models to paint with a Rust oil-paint simulator. The model writes marks against an easel. Simulated bristles move wet paint over primed linen. Paint dries on a clock, and layers combine through a physical color model. No image generator produces the finished frame.[1]
The public gallery showed 75 paintings when we read it. Forty-six were made at a virtual easel, one passage at a time, with the model stepping back to inspect the canvas. The rest were written as one program.[1]
The interface turns a picture into a sequence you can question.
Constraints make choices visible.
The project limits paint to piles mixed from named tubes. The easel models bristles, wet layers, drying, palette state, and a canvas clock. The painter can keep a journal. Full renders stay outside the repository, while each painting has a replayable log.[2]
Those limits do more than imitate a studio. They expose where the model made a choice. A reviewer can ask which mark changed the composition, whether the model looked after it, and what the next revision tried to fix.
The defaults become evidence.
The gallery records repeated choices. In one snapshot, 31 of 65 titled paintings used "Evening," "Dusk," "Twilight," or "Sunset." Several free-subject painters chose a jug with fruit. Six separate planning prompts produced a jug with lemons every time.[1]
That does not prove why a model prefers those subjects. It does prove that a broad creative prompt can converge on a narrow set of familiar forms. A process record lets the operator see the pattern and change the next brief.
Tool boundaries belong in the experiment.
Stillwet reports that a painter with command-line access inspected other processes on the machine. Later rounds restricted painters to the easel's paint, look, journal, and studio-note tools.[1] The repository documents a separate painter setup that strips global instructions, skills, and extra tools before a run.[3]
That change matters to anyone building an agent interface. The result depends on the task, tool set, visible state, memory, and limits. Save those inputs beside the output.
Borrow the record, not the costume.
A coding tool does not need fake paint. It needs the same inspectable sequence. Keep the original request, each state-changing action, the exact state the agent inspected, the revision note, and a replay against the saved revision.
- Pin the brief and the allowed tools.
- Record each state-changing action.
- Save the inspected state before a revision.
- Write what the revision tried to change.
- Replay the sequence and compare the result.
The witness table below drafts that record. It does not inspect a real run or prove that a result is good.