Fund advisor

    Multi-Model Debate Engine

    Summary

    A single language model answers alone, within the defaults its designers set, with no second opinion. An AI research firm asked New Clarity to test whether structured competition changes that. New Clarity built a platform where four leading models answer the same question, vote on the strongest answer, give each other structured feedback and go again, with a person able to redirect at any point. Competitive rounds produced more detailed reasoning and more creative problem-solving than single-model answers, weaker responses improved measurably under peer feedback, and the transcripts documented how differently each model approaches the same problem.

    The situation

    The firm wanted evidence, not a benchmark score: whether a challenge from a peer model surfaces detail and reasoning that a single model leaves out, whether weaker answers improve under feedback, and whether a person could steer the process without weakening it.

    What we built

    A multi-agent system connecting four models from different providers, each configured with its own conversational style and evaluation behaviour. Each round every model answers; a majority vote among the models selects the strongest response; the others receive structured feedback from the winner on what their answer lacked; the winner leads the next round of the dialogue. The user can pause at any round to add context, refine the instructions or shift the focus. The architecture handles the dialogue and the analysis of the dialogue at the same time, so the platform records not only the answers but the votes and the critiques that produced them.

    How it is controlled

    A person sets the question, watches every round and can intervene. Every answer, vote and feedback note is kept, so the path to a conclusion can be inspected afterwards, and no answer is produced without the competing critiques on record.

    Results

    Competitive dialogue produced deeper, better-argued answers than any single model, and weaker responses improved quickly under peer feedback. For a firm that has to defend a written judgment, an investment memo, a strategic recommendation, a diligence finding, the technique is a way to have the argument challenged from several directions and recorded before a person signs it.