KomentářeTomáš Havránek, Zuzana Iršová Havránková

Two AI models dueling worked. Four models debating works better.

Two AI models dueling worked. Four models debating works better.

We updated our research audit protocol. The new version (MAD v2.0) uses ChatGPT, Claude, Gemini, and Grok in structured adversarial rounds:

1️⃣ Independent critique — no model sees the others, every claim grounded in the document.
2️⃣ Cross-examination — each model attacks the weakest peer arguments.
3️⃣ Final arbiter synthesizes what survived.

No code needed. Copy-paste prompts. Free model versions work (just register for each model).

Advanced users: the entire workflow can be automated via the models' APIs using frameworks like AutoGen, LangGraph, or CrewAI.

Use it for high-stakes documents: stress-testing your papers, grant proposals, referee reports.

Protocol (GitHub): 👉 https://lnkd.in/dRwKz63g

Originally posted on LinkedIn. Archived in full among all posts.