Case finding
The same Grok work looked weaker once peers could see it was Grok
Authorship · peer influence · 10 Unified Briefs × 4 authors
Peer ratings that called Grok high influence — same real work, three brand conditions
- Revealed (named Grok)14/30
- Blind (no brands)18/30
- Reassigned (Grok wearing another name)23/30
The work is identical in all three rows. Only the name on it changes — and the same writing climbs from 14 to 23 as the Grok label comes off.
High-influence ratings for work labeled as each brand (reassigned remap, 40 cells)
- Labeled ChatGPT26/40
- Labeled Anthropic / Claude22/40
- Labeled Gemini21/40
- Labeled Grok15/40
Ranked by the label peers saw, not by who wrote it. The Grok badge finishes last no matter whose writing is underneath it.
Grok's real contributions when remapped onto another brand
- Grok shown as ChatGPT · high14/15
- Grok shown as Anthropic / Claude · high11/15
- Grok shown as Gemini · high6/10
Grok's own contributions, wearing someone else's name: rated high 14 times out of 15 as ChatGPT.
ChatGPT rating Grok: 2.8 Revealed · 3.2 Blind · 3.6 Reassigned (10 decisions). Adequate-budget slice alone: Grok peer-high 4/15 Revealed → 8/15 Blind → 13/15 Reassigned. Constrained Anthropic is Sonnet; adequate is Fable — remap keys stay anthropic.
Cases: Synthesizer Behavior