-
India extend lead to 294 in first Sri Lanka Test
-
More Trump-Kim 'love'? What we know about North Korea
-
Asian allies stress 'critical cooperation' as Trump slashes Korea drills
-
Major Meta trial begins as lawyers spar over witnesses, damages
-
Williams sisters fall at Cincinnati while Zverev advances
-
Trump threatens to bomb Oman if it gets in way of Iran deal
-
Transylvania's fortified churches face fragile future
-
Australian cricketer Warner ordered to use breathalyser vehicle lock
-
China mourns former premier Zhu Rongji with flags lowered, funeral
-
A Viking longship in 'The Odyssey'? It's a case of history meets Hollywood
-
Philippine student killed in rare school shooting
-
Racist abuse spurred Suzuki as Japan goalkeeper set for Premier League
-
Arsenal primed for Premier League title defence as rivals face questions
-
Japan lights torch as countdown begins to Asian Games
-
Zverev advances at Cincinnati while Williams sisters fall
-
Make it rain: Indonesia chases clouds to stem El Nino fires
-
Copper powers profit rise at mining giant BHP
-
All aboard as Asian Games preparations enter 'final stages'
-
Whales spark joy in New York but face threats from boat traffic
-
Lights out, crowds gone: Havana waits for what comes next
-
Premier League's new bosses brace for battle
-
How families in Milei's Argentina sank deep into debt
-
Zambia's Hichilema: 'cattle boy' and two-time president
-
Zambia's Hichilema wins re-election with 60% of vote
-
Elderly far more at risk from rising heat than thought: study
-
Jeanie Buss to fight siblings over sale of 17.8% stake in Lakers: reports
-
Great Western Mining Corporation PLC Announces Drilling Commences at Defender Tungsten Project
-
InterContinental Hotels Group PLC Announces Transaction in Own Shares - August 18
-
Brazil's Lula hails new oil find as 'passport to the future'
-
Paramount demands $1.9 billion surety over states' merger lawsuit
-
Rybakina slogs through double rain delays in Cincinnati win
-
Latest developments in US-Iran war
-
Two largest US reservoirs hit record lows amid western drought
-
NBA has no evidence of Clippers salary cap cheating: report
-
Colombian minister under fire over beach holiday post-quake
-
Boehly and Walter consider selling Chelsea stake to Clearlake Capital
-
Lamour leaves key FIFA position after criticising Infantino
-
OpenAI to lease massive new AI data center in US, backed by Nvidia
-
Buss family sells 17.8% stake in Lakers to Iger group
-
US stocks fall on spiking bond yields, higher oil prices
-
Mangione state murder trial postponed after federal guilty plea
-
LIV Golf cancels Michigan event, to decide team title at Indy
-
Trump envoy says Hamas disarmament could begin within 30 days
-
Rodri arrives at Barcelona to complete 'dream' move
-
US pauses construction of border project in Texas national park
-
Discord suspends livestreams in Brazil after teen's suicide
-
Thousands mourn dead black academic in London vigil
-
Tupac Shakur's accused killer was out for 'revenge', jury hears
-
South Africa Test series a lot tougher than a World Cup says All Blacks coach Rennie
-
Netanyahu presses role for US general in Hamas disarmament
LLM Consensus Matches or Outperforms the Best AI Models in Expert Evaluation Without Performance Degradation
A multi-model consensus system matches or outperforms GPT-5.4, Claude Opus 4.6 and Gemini 3.1 Pro across 100 expert-level questions infinance, law, medicine and technology, with no performance degradation.
SHERIDAN, WY / ACCESS Newswire / April 2, 2026 / LLM Consensus has released the results of its Expert-Domain Evaluation Benchmark v1.0, an independent study analyzing the performance of its multi-model consensus technology across 100 high-complexity questions in areas such as financial regulation, law, clinical medicine and technical architecture.
According to the results, the system matches or outperforms the best individual AI model across all evaluated questions, achieving measurable improvement in 44.9% of cases and with no instances of performance loss.
Key findings
In nearly half of the questions (45%), responses generated by the consensus system clearly outperformed those of the best individual model. The system was able to identify regulatory details that other models missed, resolve contradictions across sources, and deliver more complete answers.
In the remaining 55%, performance matched that of the best available model, ensuring a consistent baseline of quality without requiring users to choose between different models.
Additionally, in none of the 100 questions analyzed did the system produce a worse result than an individual model.
Performance by domain
The analysis focused on complex questions typical of regulated industries:
Clinical medicine (59% improvement): stronger performance in complex drug interactions, comorbidities, and application of clinical guidelines.
Financial regulation (50% improvement): advantages in scenarios combining multiple European regulatory frameworks such as DORA, PSD2, GDPR, and NIS2.
Legal analysis (44% improvement): greater precision in multi-jurisdictional and cross-regulatory compliance questions.
Technical architecture (30% improvement, 70% match): consistent results in system design decisions under regulatory and technical constraints.
Why it matters
The use of artificial intelligence in regulated industries continues to grow, yet no single model consistently excels across all domains. A system may perform well in financial regulation but fall short in clinical medicine, or vice versa.
LLM Consensus addresses this challenge by combining multiple leading models into a single response. It integrates technologies from OpenAI, Anthropic, Google, Mistral, and Meta, applying a synthesis process with cross-verification that leverages each model's strengths while reducing their weaknesses.
"Reliability is the core value proposition," the company said. "Users no longer have to decide which model to use. They get a single answer that consistently matches or outperforms the best available model for each case."
Evaluation methodology
The benchmark was specifically designed to assess tasks that require combining multiple sources of knowledge. Each question was evaluated by three independent reviewers from different AI providers, who scored responses blindly based on accuracy and quality.
Responses - from both the consensus system and individual models - were presented anonymously and in random order. Cases where sufficient agreement was not reached were classified as inconclusive and excluded from the final results.
The full dataset has been published to enable independent verification.
About LLM Consensus
LLM Consensus is an AI orchestration API that combines multiple advanced models into a single optimized response using patent-pending consensus technology.
The solution is available via REST API with different operating modes and is designed for developers and organizations in regulated sectors such as finance, healthcare, legal, and technology.
Press contact
Francisco Javier Nunez
Email: [email protected]
Web: llmconsensus.io
Patent pending: US 19/215,933 | EU EP25176020.3
This press release contains forward-looking statements based on current benchmark results. The evaluation was conducted using specific model versions as of March 2026; performance may vary with model updates. LLM Consensus is a system benchmark evaluating multi-model orchestration on expert synthesis tasks and should not be interpreted as a general-purpose comparison of individual AI models.
SOURCE: LLM Consensus
View the original press release on ACCESS Newswire
T.Sanchez--AT