Close Menu
Alpha Leaders
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
What's On
‘I’m just stuck’: Meet the former OpenAI researcher sitting on 0K of equity he says is overvalued

‘I’m just stuck’: Meet the former OpenAI researcher sitting on $700K of equity he says is overvalued

30 July 2026
How AI Is Complicating Federal Reserve Interest-Rate Decisions

How AI Is Complicating Federal Reserve Interest-Rate Decisions

30 July 2026
Has OpenAI already quietly hit pause on some AI development?

Has OpenAI already quietly hit pause on some AI development?

30 July 2026
Facebook X (Twitter) Instagram
Facebook X (Twitter) Instagram
Alpha Leaders
newsletter
  • Home
  • News
  • Leadership
  • Entrepreneurs
  • Business
  • Living
  • Innovation
  • More
    • Money & Finance
    • Web Stories
    • Global
    • Press Release
Alpha Leaders
Home » This AI-Powered ‘Coach’ Catches Hallucinations In Other AI Models
Innovation

This AI-Powered ‘Coach’ Catches Hallucinations In Other AI Models

Press RoomBy Press Room11 July 20244 Mins Read
Facebook Twitter Copy Link Pinterest LinkedIn Tumblr Email WhatsApp
This AI-Powered ‘Coach’ Catches Hallucinations In Other AI Models

Generative AI models can produce tremendous results for certain applications–but they are also notorious for confidently, sometimes convincingly, making mistakes (or “hallucinations”) like suggesting people eat rocks five times a day or add glue to pizza. A new open source AI model called Lynx, developed by nascent AI evaluation company Patronus AI, aims to fix this problem. The model promises a faster, cheaper and more reliable way to detect such hallucinations without human help.

Cofounders Anand Kannappan and Rebecca Qian, both ex-Meta AI researchers, claim that the new model is more accurate than other leading AI systems, like OpenAI’s GPT models and Anthropic’s Claude 3 model, in detecting factual inaccuracies. To accomplish this, the company fine tuned Meta’s most advanced large language model, Llama 3, by showing it 2400 examples of hallucinations and their corresponding correct responses.

Before starting the company in September 2023, the duo spoke with about 60 company executives and found that their worst fear was launching an AI product and making headlines for the wrong reasons. CEO Kannappan hopes that Lynx can help assuage those concerns. He thinks of it as a “coach” for other AI models that can guide them to be more accurate. The goal is that its customers that are rolling out AI applications could use the Lynx to uncover hallucinations during development rather than fixing blunders after they have already launched.

“One of the reasons why Rebecca and I started the company was this concept called scalable oversight,” Kannappan said. “And it was about how humans can continue to supervise systems that far outperform them. And the only way you can do that is if you have a really, really powerful AI that evaluates AI.”

That’s a contrast to the way AI products are stress tested now before shipment, he continued, which involves a range of techniques. One of those is “red teaming,” which involves manually hacking AI models to expose vulnerabilities that might lead to mistakes. Other development teams use AI models like GPT-4 to catch hallucinations, Kannappan said, who criticized this approach as “literally GPT-4 evaluating GPT-4 itself.” That’s a problem, he explained, because general purpose models like GPT-4 weren’t specifically designed to catch errors. Lynx, on the other hand, was taught how to reason why an answer is wrong as it was fed more context, said Qian.

“We provided examples of incorrect answers and gave the specific financial calculation or medical citation that showed why the response is wrong,” she said. This is a more effective approach because the model is provided with additional background information to better catch similar mistakes.

The company has also released a new benchmark called HaluBench, which rates how well different AI models can catch hallucinations in model output, especially across legal, financial and medical domains. That benchmark shows that even Lynx isn’t perfect–it scores about 88% accuracy–but it does outperform most others, Kannappan said.

In March, Patronus AI also launched Copyright Catcher, a tool that detects when popular AI models like OpenAI’s GPT-4, Anthropic’s Claude 2 and Mistral AI’s Mixtral produce copyrighted content. The tool caught such models regurgitating entire paragraphs from books like Michelle Obama’s Becoming and John Green’s The Fault in Our Stars.

The company has also developed other tools that evaluate model performance across particular domains. For example, there’s FinanceBench, which is used to evaluate how well different LLMs answer financial queries; Enterprise PII, which helps companies use to detect whether AI models are exposing their sensitive and confidential information; and Simple Safety, which evaluates LLMs for safety risks such producing harmful responses related to suicide, child abuse and fraud.

All the work is geared towards the mission of the company—making sure LLMs aren’t producing bad results that people end up relying on. “When the model hallucinates, it still produces output that sounds plausible,” Qian said. “That ends up leading to misinformation.”

MORE FROM FORBES

AI Ai model Benchmark evaluation GPT-4 hallucination Lynx Meta Patronus AI red teaming
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link

Related Articles

How AI Is Complicating Federal Reserve Interest-Rate Decisions

How AI Is Complicating Federal Reserve Interest-Rate Decisions

30 July 2026
Friday, July 31 Clues And Answers

Friday, July 31 Clues And Answers

30 July 2026
Claude Makes Five AI Labs Publishing Private Chats To Google

Claude Makes Five AI Labs Publishing Private Chats To Google

30 July 2026
E-Day’ Dev Promises ‘No Battle Pass, No Bull—’ For Multiplayer

E-Day’ Dev Promises ‘No Battle Pass, No Bull—’ For Multiplayer

30 July 2026
‘Marathon’ Extends PvE Mode Trial As Its Pivot Ramps Up

‘Marathon’ Extends PvE Mode Trial As Its Pivot Ramps Up

30 July 2026
How Manufacturers Can Strengthen Resilience

How Manufacturers Can Strengthen Resilience

30 July 2026
Don't Miss
Exclusive: DeFi platform Azura launches after raising .9 million from Initialized

Exclusive: DeFi platform Azura launches after raising $6.9 million from Initialized

By Press Room22 October 2024

Azura, a new platform for decentralized finance, launched on Tuesday after raising $6.9 million in…

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

Unwrap Christmas Sustainably: How To Handle Gifts You Don’t Want

27 December 2024
Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

Sam Altman’s World Wants To Scan Your Eyes To Prove You’re Human

22 October 2024
Stay In Touch
  • Facebook
  • Twitter
  • Pinterest
  • Instagram
  • YouTube
  • Vimeo
Latest Articles
Nearly a third of workers admit to sabotaging their company’s AI—smaller paychecks may explain why

Nearly a third of workers admit to sabotaging their company’s AI—smaller paychecks may explain why

30 July 20261 Views
Claude Makes Five AI Labs Publishing Private Chats To Google

Claude Makes Five AI Labs Publishing Private Chats To Google

30 July 20261 Views
Ultra-rich are buying up  million mansions in London, with ‘Trump unease’ fueling the influx

Ultra-rich are buying up $49 million mansions in London, with ‘Trump unease’ fueling the influx

30 July 20261 Views
E-Day’ Dev Promises ‘No Battle Pass, No Bull—’ For Multiplayer

E-Day’ Dev Promises ‘No Battle Pass, No Bull—’ For Multiplayer

30 July 20261 Views

Recent Posts

  • ‘I’m just stuck’: Meet the former OpenAI researcher sitting on $700K of equity he says is overvalued
  • How AI Is Complicating Federal Reserve Interest-Rate Decisions
  • Has OpenAI already quietly hit pause on some AI development?
  • Friday, July 31 Clues And Answers
  • Nearly a third of workers admit to sabotaging their company’s AI—smaller paychecks may explain why

Recent Comments

No comments to show.
About Us
About Us

Alpha Leaders is your one-stop website for the latest Entrepreneurs and Leaders news and updates, follow us now to get the news that matters to you.

Facebook X (Twitter) Pinterest YouTube WhatsApp
Our Picks
‘I’m just stuck’: Meet the former OpenAI researcher sitting on 0K of equity he says is overvalued

‘I’m just stuck’: Meet the former OpenAI researcher sitting on $700K of equity he says is overvalued

30 July 2026
How AI Is Complicating Federal Reserve Interest-Rate Decisions

How AI Is Complicating Federal Reserve Interest-Rate Decisions

30 July 2026
Has OpenAI already quietly hit pause on some AI development?

Has OpenAI already quietly hit pause on some AI development?

30 July 2026
Most Popular
Friday, July 31 Clues And Answers

Friday, July 31 Clues And Answers

30 July 20261 Views
Nearly a third of workers admit to sabotaging their company’s AI—smaller paychecks may explain why

Nearly a third of workers admit to sabotaging their company’s AI—smaller paychecks may explain why

30 July 20261 Views
Claude Makes Five AI Labs Publishing Private Chats To Google

Claude Makes Five AI Labs Publishing Private Chats To Google

30 July 20261 Views

Archives

  • July 2026
  • June 2026
  • May 2026
  • April 2026
  • March 2026
  • February 2026
  • January 2026
  • December 2025
  • November 2025
  • October 2025
  • September 2025
  • August 2025
  • July 2025
  • June 2025
  • May 2025
  • April 2025
  • March 2025
  • February 2025
  • January 2025
  • December 2024
  • November 2024
  • October 2024
  • September 2024
  • August 2024
  • July 2024
  • June 2024
  • May 2024
  • April 2024
  • March 2024
  • February 2024
  • January 2024
  • December 2023
  • March 2022
  • January 2021
  • March 2020
  • January 2020

Categories

  • Blog
  • Business
  • Entrepreneurs
  • Global
  • Innovation
  • Leadership
  • Living
  • Money & Finance
  • News
  • Press Release
© 2026 Alpha Leaders. All Rights Reserved.
  • Privacy Policy
  • Terms of use
  • Advertise
  • Contact

Type above and press Enter to search. Press Esc to cancel.