-
WhiteBIT Launches Bitcoin Lightning Network Support, Powered by Voltage
-
ONAR Advances Nasdaq Listing Preparation Following Advertise Purple Acquisition and Financing
-
Oil surges, stocks sink as US reportedly eyes fresh Iran strikes
-
Thailand opens extradition process for reporter sought by China
-
Anne Carson: modern takes on Greek classics
-
France struggles to wean social housing off gas
-
Canada's Anne Carson wins Nobel literature prize
-
Germany lifts GDP forecast as economy withstands Iran war
-
Ethiopia drones hit Eritrean troops in Tigray
-
Sri Lanka appoints Kusal Mendis as Test captain
-
Russian strike on buses kills 30 in frontline Ukrainian city
-
Norris confident McLaren can shine in Singapore
-
Thailand begins extradition process against Chinese journalist
-
Alex Bodi pushes ahead with expansion: ALIX LASERS showroom planned for Bucharest
-
Quarantine, unease and rumours: Russian city grapples with plague scare
-
Fossil fuels off COP31 agenda as global output keeps rising
-
Taiwan says exports hit record in September on AI demand
-
Haze not 'a threat' to Singapore F1 race: organisers
-
Marc Marquez vows to 'attack' Martin's lead at Indonesian MotoGP
-
COP31 organisers say to seek 'ambitious' action with their text
-
Diplomats, schools in Saudi take safety measures after blasts heard: sources to AFP
-
Bartunkova thrashes Muchova to reach China Open semi-finals
-
In Beijing, EU trade envoy says deficit with China 'unsustainable'
-
Drone hits Yandex data centre, in first such attack on Russia
-
Sri Lanka's de Silva quits as Test captain
-
Oil prices surge and stocks sink on fresh inflation worries
-
Alcaraz feeling 'mentally fresh' ahead of Shanghai return
-
AI startup Manus raises $500m after Meta acquisition blocked by China
-
Israel drastically reduces British diplomatic presence in Jerusalem
-
From refuge to repatriations: Malaysia's Myanmar U-turn
-
Stocks fall further as oil surge fans fresh inflation worries
-
Morikawa says LIV players face 'business decision' after Rahm exit
-
North Korea demands South show 'utmost respect' to restore relations
-
Tenzing Norgay's son says Himalayas sending climate warning
-
Saudi says three dead in airport attacks claimed by Houthis
-
On your bike! Bottas opts for two wheels to reach Singapore GP
-
'Mad dream': New showcase for African film opens in London
-
Alcaraz, Sinner, Osaka commit to Australian Open '1 Point Slam'
-
All Blacks coach Rennie says Wallabies exit 'part of the game'
-
Dodgers beat Braves to advance in MLB playoffs, Rays oust Yankees
-
The growing protests against India's poll body - and PM Modi
-
'A sun as black as my lungs': how corruption crucified a polluted Albania city
-
Stocks fall further as oil spike fans fresh inflation worries
-
The British historian helping the French rediscover de Gaulle
-
Bad COP, worse COP? EU tempers hopes for UN climate summit
-
Dodgers beat Braves to advance, Guardians survive in MLB playoffs
-
Rookie Henry named in Wallabies midfield to face All Blacks
-
Latin American, Caribbean women top Nobel literature buzz
-
Housing crisis set to dominate Spain's snap election
-
US says Fiji-based Chinese official paid bribes for Beijing
LLM Consensus Matches or Outperforms the Best AI Models in Expert Evaluation Without Performance Degradation
A multi-model consensus system matches or outperforms GPT-5.4, Claude Opus 4.6 and Gemini 3.1 Pro across 100 expert-level questions infinance, law, medicine and technology, with no performance degradation.
SHERIDAN, WY / ACCESS Newswire / April 2, 2026 / LLM Consensus has released the results of its Expert-Domain Evaluation Benchmark v1.0, an independent study analyzing the performance of its multi-model consensus technology across 100 high-complexity questions in areas such as financial regulation, law, clinical medicine and technical architecture.
According to the results, the system matches or outperforms the best individual AI model across all evaluated questions, achieving measurable improvement in 44.9% of cases and with no instances of performance loss.
Key findings
In nearly half of the questions (45%), responses generated by the consensus system clearly outperformed those of the best individual model. The system was able to identify regulatory details that other models missed, resolve contradictions across sources, and deliver more complete answers.
In the remaining 55%, performance matched that of the best available model, ensuring a consistent baseline of quality without requiring users to choose between different models.
Additionally, in none of the 100 questions analyzed did the system produce a worse result than an individual model.
Performance by domain
The analysis focused on complex questions typical of regulated industries:
Clinical medicine (59% improvement): stronger performance in complex drug interactions, comorbidities, and application of clinical guidelines.
Financial regulation (50% improvement): advantages in scenarios combining multiple European regulatory frameworks such as DORA, PSD2, GDPR, and NIS2.
Legal analysis (44% improvement): greater precision in multi-jurisdictional and cross-regulatory compliance questions.
Technical architecture (30% improvement, 70% match): consistent results in system design decisions under regulatory and technical constraints.
Why it matters
The use of artificial intelligence in regulated industries continues to grow, yet no single model consistently excels across all domains. A system may perform well in financial regulation but fall short in clinical medicine, or vice versa.
LLM Consensus addresses this challenge by combining multiple leading models into a single response. It integrates technologies from OpenAI, Anthropic, Google, Mistral, and Meta, applying a synthesis process with cross-verification that leverages each model's strengths while reducing their weaknesses.
"Reliability is the core value proposition," the company said. "Users no longer have to decide which model to use. They get a single answer that consistently matches or outperforms the best available model for each case."
Evaluation methodology
The benchmark was specifically designed to assess tasks that require combining multiple sources of knowledge. Each question was evaluated by three independent reviewers from different AI providers, who scored responses blindly based on accuracy and quality.
Responses - from both the consensus system and individual models - were presented anonymously and in random order. Cases where sufficient agreement was not reached were classified as inconclusive and excluded from the final results.
The full dataset has been published to enable independent verification.
About LLM Consensus
LLM Consensus is an AI orchestration API that combines multiple advanced models into a single optimized response using patent-pending consensus technology.
The solution is available via REST API with different operating modes and is designed for developers and organizations in regulated sectors such as finance, healthcare, legal, and technology.
Press contact
Francisco Javier Nunez
Email: [email protected]
Web: llmconsensus.io
Patent pending: US 19/215,933 | EU EP25176020.3
This press release contains forward-looking statements based on current benchmark results. The evaluation was conducted using specific model versions as of March 2026; performance may vary with model updates. LLM Consensus is a system benchmark evaluating multi-model orchestration on expert synthesis tasks and should not be interpreted as a general-purpose comparison of individual AI models.
SOURCE: LLM Consensus
View the original press release on ACCESS Newswire
T.Sanchez--AT