
Tap image to enlarge이미지를 누르면 크게 볼 수 있습니다
Application & Purpose작품의 목적
A multi-agent system ran this study end to end, with one person in the loop. The runs, the checks and the work that was sent back all sit in one archive beside the paper. That archive is what lets anyone re-run the study, and building it that way was the work. The poster shows the archive, not the finding.
다중 에이전트 시스템이 이 연구를 처음부터 끝까지 수행했고, 사람 한 명이 루프 안에 있었습니다. 실행, 검사, 그리고 되돌려 보낸 작업이 모두 논문 옆 하나의 아카이브에 놓여 있습니다. 그 아카이브가 누구든 연구를 다시 실행할 수 있게 하며, 그렇게 만드는 것 자체가 작업이었습니다. 이 포스터는 발견이 아니라 아카이브를 보여줍니다.
AI Tools & WorkflowAI 도구와 작업 흐름
Claude models in terminals, inside the study's own repository. Separate runners under a clean environment, a stored expectation for every printed value, and a figure script that recomputes its own number and halts the build if the two disagree.
연구 저장소 안의 터미널에서 실행된 Claude 모델들. 깨끗한 환경의 독립 실행기, 출력되는 모든 값에 대한 저장된 기댓값, 그리고 자기 숫자를 다시 계산해 둘이 다르면 빌드를 중단하는 그림 스크립트.
Process & Iterations제작 과정과 반복
Nothing here converged by agreement. Each result was handed to another agent to break, and several were broken: a correlation that independent recomputations would not reproduce, a verification chain reading the wrong table, a figure standing on a value that had gone stale. The retracted draft was not deleted. It carries a header naming the number that no longer holds, and the significance test was withdrawn rather than defended. The poster went through the same loop. Drafts were read by agents given no context, and a machine gate counted the markers of AI-sounding prose. The person in the loop stopped the work more than once, and every stop is in the record.
여기서 합의로 수렴한 것은 없습니다. 각 결과는 다른 에이전트에게 넘겨져 깨뜨려 보게 했고, 실제로 여럿이 깨졌습니다. 독립 재계산으로 재현되지 않는 상관관계, 잘못된 표를 읽는 검증 체인, 오래된 값 위에 서 있던 그림. 철회된 초안은 삭제하지 않았습니다. 더 이상 성립하지 않는 숫자를 명시한 헤더를 달고 있으며, 유의성 검정은 방어하는 대신 철회했습니다. 포스터도 같은 루프를 거쳤습니다. 초안은 맥락을 모르는 에이전트들이 읽었고, 기계 게이트가 AI스러운 문장의 표지를 셌습니다. 루프 안의 사람은 작업을 여러 번 멈췄고, 모든 정지는 기록에 있습니다.
A Researcher's Perspective연구자의 관점
What it was like: the models were never the bottleneck. What surprised me: when agents came back with work that looked alike, the resemblance was in their shared brief, not in them. One word in it, cyanotype, made every sheet look a century old. What felt limiting: one vocabulary reaching every agent at once. What I would tell colleagues: cut the brief back to the question, and ask someone to say plainly when the image is not worth looking at.
어땠는가: 모델은 결코 병목이 아니었습니다. 놀라웠던 것: 에이전트들이 비슷한 작업을 가져왔을 때, 그 닮음은 에이전트가 아니라 공유된 브리프에 있었습니다. 브리프의 단어 하나, 시아노타입이 모든 시트를 백 년 묵은 것처럼 보이게 했습니다. 한계로 느낀 것: 하나의 어휘가 모든 에이전트에 동시에 닿는다는 것. 동료들에게 할 말: 브리프를 질문으로만 줄이고, 이미지가 볼 가치가 없을 때 그렇다고 분명히 말해 줄 사람을 두세요.
