-
Amid war, Ukrainian soldiers celebrate love at ritual wedding
-
Meloni government becomes Italy's longest-lasting
-
Venice to roll out red carpet for Clooney on opening night
-
Funerals in Nepal for missing, a week after deadly floods
-
Hong Kong democracy activist Wong pleads guilty to foreign collusion: AFP
-
US aircraft carrier arrives in Thailand after months at sea
-
Stocks sink, yields at multi-decade highs and oil spikes on Mideast flare-up
-
US aircraft carrier approaches Thai port after months at sea
-
US defends Trump Venezuela oil deal as energy secretary lands in Caracas
-
Nepal's young PM faces gravest test after flood catastrophe
-
Jubilant Monfils celebrates birthday with US Open win
-
What rainbow? LGBTQ Indonesians on edge as hostility mounts
-
The Japanese idealist behind the last deradicalisation NGO in Somalia
-
O'Connor eyes Wallabies World Cup swansong after Force return
-
Rybakina, Gauff advance as Zverev launches US Open campaign
-
Syria UXO causing death, injury every six hours: charity
-
Macron in Britain to help unveil tapestry, hold Burnham talks
-
Tiger Woods set to change not guilty plea in DUI case: court documents
-
China's Xi to hold talks with Sisi in rare visit to Egypt
-
Zelensky warns airlines Russian skies not safe due to Ukraine drones
-
Confident Gauff reaches US Open 2nd round with straight-set win
-
UK vows anti-settler measures as Israel warns of retaliation
-
US postal whistleblower warns ballot system could disrupt midterms
-
New pro Koivun among US Presidents Cup picks
-
John Ternus takes over as Apple enters era of AI, foldable phones
-
G20 finance leaders close talks marked by US welcoming Russia
-
Second-ranked Rybakina eases into US Open 2nd round
-
From Hollywood to big time tennis: teen makes US Open debut
-
Man City sign Fernandez, Ndiaye as Premier League clubs set new spending record
-
Man City secure Premier League record deal to sign Fernandez
-
Alcaraz aims to step it up in round two of US Open title defense
-
Kaepernick says NFL 'blackballing me' decade after protest
-
Honduran judge drops corruption charges against ex-leader Hernandez
-
Atletico loan David from Juve, send Lemar to Elche
-
Mudryk joins Spurs on loan from Chelsea after doping ban ends
-
Chelsea 'keeper Sanchez makes deadline-day move to Como
-
What's next after judge strikes down New York's landmark climate fund law
-
Kroenke to buy baseball's Los Angeles Angels: statement
-
Cobolli claws out five-set victory at rain-hit US Open
-
US defends Trump's Venezuela oil deal
-
Ndiaye seizes 'chance of a lifetime' in Man City move
-
OpenAI to launch new model with 'stronger safeguards' after hack
-
Global bond sell-off, surging oil prices send markets into the red
-
Fans mourn 'bittersweet' final Ariana Grande gig in London
-
Zuckerberg, Musk make plea at G20 for more AI data centers
-
Norway's King Haakon VIII swears allegiance to constitution
-
'Discomfort' at G20 finance talks as US hosts Russian minister
-
Syria site infrastructure 'consistent with nuclear reactor': IAEA
-
Man City set to sign Fernandez for record fee, Ndiaye on deadline day
-
Juventus sign Woltemade, Sarr on loan from Premier League
LLM Consensus Matches or Outperforms the Best AI Models in Expert Evaluation Without Performance Degradation
A multi-model consensus system matches or outperforms GPT-5.4, Claude Opus 4.6 and Gemini 3.1 Pro across 100 expert-level questions infinance, law, medicine and technology, with no performance degradation.
SHERIDAN, WY / ACCESS Newswire / April 2, 2026 / LLM Consensus has released the results of its Expert-Domain Evaluation Benchmark v1.0, an independent study analyzing the performance of its multi-model consensus technology across 100 high-complexity questions in areas such as financial regulation, law, clinical medicine and technical architecture.
According to the results, the system matches or outperforms the best individual AI model across all evaluated questions, achieving measurable improvement in 44.9% of cases and with no instances of performance loss.
Key findings
In nearly half of the questions (45%), responses generated by the consensus system clearly outperformed those of the best individual model. The system was able to identify regulatory details that other models missed, resolve contradictions across sources, and deliver more complete answers.
In the remaining 55%, performance matched that of the best available model, ensuring a consistent baseline of quality without requiring users to choose between different models.
Additionally, in none of the 100 questions analyzed did the system produce a worse result than an individual model.
Performance by domain
The analysis focused on complex questions typical of regulated industries:
Clinical medicine (59% improvement): stronger performance in complex drug interactions, comorbidities, and application of clinical guidelines.
Financial regulation (50% improvement): advantages in scenarios combining multiple European regulatory frameworks such as DORA, PSD2, GDPR, and NIS2.
Legal analysis (44% improvement): greater precision in multi-jurisdictional and cross-regulatory compliance questions.
Technical architecture (30% improvement, 70% match): consistent results in system design decisions under regulatory and technical constraints.
Why it matters
The use of artificial intelligence in regulated industries continues to grow, yet no single model consistently excels across all domains. A system may perform well in financial regulation but fall short in clinical medicine, or vice versa.
LLM Consensus addresses this challenge by combining multiple leading models into a single response. It integrates technologies from OpenAI, Anthropic, Google, Mistral, and Meta, applying a synthesis process with cross-verification that leverages each model's strengths while reducing their weaknesses.
"Reliability is the core value proposition," the company said. "Users no longer have to decide which model to use. They get a single answer that consistently matches or outperforms the best available model for each case."
Evaluation methodology
The benchmark was specifically designed to assess tasks that require combining multiple sources of knowledge. Each question was evaluated by three independent reviewers from different AI providers, who scored responses blindly based on accuracy and quality.
Responses - from both the consensus system and individual models - were presented anonymously and in random order. Cases where sufficient agreement was not reached were classified as inconclusive and excluded from the final results.
The full dataset has been published to enable independent verification.
About LLM Consensus
LLM Consensus is an AI orchestration API that combines multiple advanced models into a single optimized response using patent-pending consensus technology.
The solution is available via REST API with different operating modes and is designed for developers and organizations in regulated sectors such as finance, healthcare, legal, and technology.
Press contact
Francisco Javier Nunez
Email: [email protected]
Web: llmconsensus.io
Patent pending: US 19/215,933 | EU EP25176020.3
This press release contains forward-looking statements based on current benchmark results. The evaluation was conducted using specific model versions as of March 2026; performance may vary with model updates. LLM Consensus is a system benchmark evaluating multi-model orchestration on expert synthesis tasks and should not be interpreted as a general-purpose comparison of individual AI models.
SOURCE: LLM Consensus
View the original press release on ACCESS Newswire
P.Costa--AMWN