Close Menu
Tech Nova Mindset – Empower Innovation and Forward Thinking

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    A US-China AI Hotline Won’t Be Ready For a While

    September 23, 2026

    A congressional representative just proposed killing America’s border tower program

    September 23, 2026

    Parents Guide Kids to Top Engineering Skills In AI Age

    September 23, 2026
    Facebook X (Twitter) Instagram
    Trending
    • A US-China AI Hotline Won’t Be Ready For a While
    • A congressional representative just proposed killing America’s border tower program
    • Parents Guide Kids to Top Engineering Skills In AI Age
    • Meta’s Muse AI Assistant Rolled Out With a Serious Security Flaw
    • The Download: India’s smart glasses menace and AI’s trillion-dollar gamble
    • AT&T Is Automating Away Jobs—and Its Old Telecom Empire
    • Smart glasses are already causing havoc in India
    • Viture’s Vonder Glasses Are Meant to Map Your Mind
    Tech Nova Mindset – Empower Innovation and Forward Thinking
    • Home
    • Gadgets
    • Reviews
    • Tech News
    • Future Tech
    • AI & Robotics
    • How-To Guides
    • More
      • Cybersecurity
      • Startups & Innovation
    Tech Nova Mindset – Empower Innovation and Forward Thinking
    Home»AI & Robotics»An AI Council Just Aced the US Medical Licensing Exam
    AI & Robotics

    An AI Council Just Aced the US Medical Licensing Exam

    kirklandc008@gmail.comBy kirklandc008@gmail.comOctober 12, 2025No Comments3 Mins Read
    Facebook Twitter Pinterest LinkedIn Tumblr Email
    An AI Council Just Aced the US Medical Licensing Exam
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Despite their usefulness, large language models still have a reliability problem. A new study shows that a team of AIs working together can score up to 97 percent on US medical licensing exams, outperforming any single AI.

    While recent progress in large language models (LLMs) has led to systems capable of passing professional and academic tests, their performance remains inconsistent. They’re still prone to hallucinations—plausible sounding but incorrect statements—which has limited their use in high-stakes area like medicine and finance.

    Nonetheless, LLMs have scored impressive results on medical exams, suggesting the technology could be useful in this area if their inconsistencies can be controlled. Now, researchers have shown that getting a “council” of five AI models to deliberate over their answers rather than working alone can lead to record-breaking scores in the US Medical Licensing Examination (USMLE).

    “Our study shows that when multiple AIs deliberate together, they achieve the highest-ever performance on medical licensing exams,” Yahya Shaikh, from John Hopkins University, said in a press release. “This demonstrates the power of collaboration and dialogue between AI systems to reach more accurate and reliable answers.”

    The researchers’ approach takes advantage of a quirk in the models, rooted in the non-deterministic way they come up with responses. Ask the same model the same medical question twice, and it might produce two different answers—sometimes correct, sometimes not.

    In a paper in PLOS Medicine, the team describes how they harnessed this characteristic to create their AI “council.” They spun up five instances of OpenAI’s GPT-4 and prompted them to discuss answers to each question in a structured exchange overseen by a facilitator algorithm.

    When their responses diverged, the facilitator summarized the differing rationales and got the group to reconsider the answer, repeating the process until consensus emerged.

    When tested on 325 publicly available questions from the three stages of the USMLE, the AI council achieved 97 percent, 93 percent, and 94 percent accuracy respectively. These scores not only exceed the performance of any individual GPT-4 instance but also surpass the average human passing thresholds for the same tests.

    “Our work provides the first clear evidence that AI systems can self-correct through structured dialogue, with a performance of the collective better that the performance of any single AI,” says Shaikh.

    In a testament to the effectiveness of the approach, when the models initially disagreed, the deliberation process corrected more than half of their earlier errors. Overall, the council ultimately reached the correct conclusion 83 percent of the time when there wasn’t a unanimous initial answer.

    “This study isn’t about evaluating AI’s USMLE test-taking prowess,” co-author Zishan Siddiqui notes, also from John Hopkins, said in the press release. “We describe a method that improves accuracy by treating AI’s natural response variability as a strength. It allows the system to take a few tries, compare notes, and self-correct, and it should be built into future tools for education and, where appropriate, clinical care.”

    The team notes that their results come from controlled testing, not real-world clinical environments, so there’s a long way before the AI council could be deployed in the real world. But they suggest that the approach could prove useful in other domains as well.

    It seems like the old adage that two heads are better than one remains true even when those heads aren’t human.

    Aced Council Exam Licensing Medical
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    kirklandc008@gmail.com
    • Website

    Related Posts

    A US-China AI Hotline Won’t Be Ready For a While

    September 23, 2026

    A congressional representative just proposed killing America’s border tower program

    September 23, 2026

    Parents Guide Kids to Top Engineering Skills In AI Age

    September 23, 2026
    Leave A Reply Cancel Reply

    Top Posts

    Nothing CEO says phone prices are going to keep going up

    June 12, 20267 Views

    The best VPN routers of 2026: Expert tested and reviewed

    June 14, 20263 Views

    Google DeepMind Plans to Track AGI Progress With These 10 Traits of General Intelligence

    March 21, 20263 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Recent Posts
    • A US-China AI Hotline Won’t Be Ready For a While
    • A congressional representative just proposed killing America’s border tower program
    • Parents Guide Kids to Top Engineering Skills In AI Age
    • Meta’s Muse AI Assistant Rolled Out With a Serious Security Flaw
    • The Download: India’s smart glasses menace and AI’s trillion-dollar gamble

    A US-China AI Hotline Won’t Be Ready For a While

    September 23, 2026

    A congressional representative just proposed killing America’s border tower program

    September 23, 2026

    Parents Guide Kids to Top Engineering Skills In AI Age

    September 23, 2026

    Meta’s Muse AI Assistant Rolled Out With a Serious Security Flaw

    September 23, 2026
    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Privacy Policy
    • Terms and Conditions
    • Disclaimer
    © 2026 TechNovaMindset. Designed by By Pro.

    Type above and press Enter to search. Press Esc to cancel.