Close Menu
TechTost
  • AI
  • Apps
  • Crypto
  • Fintech
  • Hardware
  • Media & Entertainment
  • Security
  • Startups
  • Transportation
  • Venture
  • Recommended Essentials
What's Hot

Port raises $100M valuation from $800M round to take on Spotify’s Backstage

India’s Spinny lines up $160m funding to acquire GoMechanic, sources say

OpenAI hits back at Google with GPT-5.2 after ‘code red’ memo.

Facebook X (Twitter) Instagram
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
Facebook X (Twitter) Instagram
TechTost
Subscribe Now
  • AI

    OpenAI hits back at Google with GPT-5.2 after ‘code red’ memo.

    14 December 2025

    Trump’s AI executive order promises ‘a rulebook’ – startups may find legal loophole instead

    13 December 2025

    Ok, so what’s up with the LinkedIn algo?

    12 December 2025

    Google Released Its Deepest Research AI Agent To Date — The Same Day OpenAI Dropped GPT-5.2

    12 December 2025

    Disney hits Google with cease and desist alleging ‘massive’ copyright infringement

    11 December 2025
  • Apps

    Google’s AI testing feature for clothes now only works with a selfie

    14 December 2025

    DoorDash driver faces felony charges after allegedly spraying customers’ food

    13 December 2025

    Google Translate now lets you listen to real-time translations on your headphones

    13 December 2025

    With iOS 26.2, Apple lets you bring back Liquid Glass again — this time on the lock screen

    12 December 2025

    World launches its ‘super app’, including payment encryption and encrypted chat features

    12 December 2025
  • Crypto

    New report examines how David Sachs may benefit from Trump administration role

    1 December 2025

    Why Benchmark Made a Rare Crypto Bet on Trading App Fomo, with $17M Series A

    6 November 2025

    Solana co-founder Anatoly Yakovenko is a big fan of agentic coding

    30 October 2025

    MoviePass opens Mogul fantasy league game to the public

    29 October 2025

    Only 5 days until Disrupt 2025 sets the startup world on fire

    22 October 2025
  • Fintech

    Coinbase starts onboarding users again in India, plans to do fiat on-ramp next year

    7 December 2025

    Walmart-backed PhonePe shuts down Pincode app in yet another step back in e-commerce

    5 December 2025

    Nexus stays out of AI, keeping half of its new $700M fund for India startup

    4 December 2025

    Fintech firm Marquis notifies dozens of US banks and credit unions of data breach after ransomware attack

    3 December 2025

    Revolut hits $75 billion valuation in new capital raise

    24 November 2025
  • Hardware

    Pebble founder unveils $75 AI smart ring to record short notes with the push of a button

    10 December 2025

    Amazon’s Ring launches controversial AI-powered facial recognition feature on video doorbells

    10 December 2025

    Google’s first AI glasses are expected next year

    9 December 2025

    eSIM adoption is on the rise thanks to travel and device compatibility

    6 December 2025

    AWS re:Invent was an all-in pitch for AI. Customers may not be ready.

    5 December 2025
  • Media & Entertainment

    Disney signs deal with OpenAI to allow Sora to create AI videos with its characters

    11 December 2025

    YouTube TV will launch genre-based subscription plans in 2026

    11 December 2025

    Founder of AI startup Tavus says users talk to AI Santa ‘for hours’ a day

    10 December 2025

    Spotify releases music videos in the US and Canada for Premium subscribers

    9 December 2025

    Amazon Music’s 2025 Delivered is now here to compete with Spotify Wrapped

    9 December 2025
  • Security

    The flaw in the photo booth manufacturer’s website exposes customers’ photos

    13 December 2025

    Home Depot exposed access to internal systems for a year, researcher says

    13 December 2025

    Security flaws in the Freedom Chat app exposed users’ phone numbers and PINs

    11 December 2025

    Petco takes down Vetco website after exposing customers’ personal information

    10 December 2025

    Petco’s security bug affected customers’ SSNs, driver’s licenses and more

    9 December 2025
  • Startups

    Port raises $100M valuation from $800M round to take on Spotify’s Backstage

    14 December 2025

    Eclipse Energy’s microbes can turn dormant oil wells into hydrogen factories

    13 December 2025

    Interest in Spoor’s AI bird tracking software is soaring

    13 December 2025

    Retro, a photo-sharing app for friends, lets you ‘time travel’ to your camera roll

    12 December 2025

    On Me Raises $6M to Shake Up the Gift Card Industry

    12 December 2025
  • Transportation

    India’s Spinny lines up $160m funding to acquire GoMechanic, sources say

    14 December 2025

    Inside Rivian’s big bet on self-driving with artificial intelligence

    13 December 2025

    Zevo wants to add robotaxis to its car-sharing fleet, starting with newcomer Tensor

    13 December 2025

    Driving aboard Rivian’s fight for autonomy

    12 December 2025

    Rivian goes big on autonomy, with custom silicon, lidar and a hint of robotaxis

    12 December 2025
  • Venture

    Runware raises $50 million in Series A to make it easier for developers to create images and videos

    12 December 2025

    Stanford’s star reporter understands Silicon Valley’s startup culture

    12 December 2025

    The market has “changed” and founders now have the power, VCs say

    11 December 2025

    Tiger Global plans cautious business future with new $2.2 billion fund

    8 December 2025

    Sources: AI-powered synthetic research startup Aaru raises Series A at $1B ‘headline’ valuation

    6 December 2025
  • Recommended Essentials
TechTost
You are at:Home»AI»Why RAG Won’t Solve the Problem of Genetic AI Hallucinations
AI

Why RAG Won’t Solve the Problem of Genetic AI Hallucinations

techtost.comBy techtost.com5 May 202405 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email
Why Rag Won't Solve The Problem Of Genetic Ai Hallucinations
Share
Facebook Twitter LinkedIn Pinterest Email

Illusions – the lies that artificial intelligence models are basically telling – are a big problem for businesses looking to integrate the technology into their operations.

Because models have no real intelligence and simply predict words, images, speech, music and other data according to a private schema, they sometimes get it wrong. Very wrong. In a recent article in the Wall Street Journal, a source recounts an example where Microsoft’s genetic AI invented meeting participants and implied that the conference calls were about topics that were not actually discussed on the call.

As I wrote a while back, illusions can be an intractable problem with today’s transformer-based model architectures. But a number of leading AI vendors are proposing to do so can to be more or less eliminated through a technical approach called recovery augmented production or RAG.

Here’s how one salesman, Squirro, he drops it:

At the core of the offering is the concept of Retrieval Augmented LLMs or Retrieval Augmented Generation (RAG) built into the solution… [our generative AI] is unique in its promise of zero hallucinations. Every piece of information it generates is traceable to a source, ensuring reliability.

Here is one similar step from SiftHub:

Using RAG technology and enhanced large language models with industry-specific knowledge training, SiftHub enables companies to create personalized responses with zero hallucinations. This guarantees increased transparency and reduced risk and inspires complete confidence in using AI for all their needs.

RAG was pioneered by data scientist Patrick Lewis, a researcher at Meta and University College London and lead author of the 2020 paper who coined the term. Applied to a model, RAG retrieves documents possibly related to a question — for example, a Wikipedia page about the Super Bowl — using what is essentially a keyword search, and then asks the model to generate answers to this additional context.

“When you interact with a generative AI model like ChatGPT or Llama and ask a question, the default is for the model to answer from its ‘parametric memory’ — that is, from the knowledge stored in its parameters as a result of training on huge data from the web,” explained David Wadden, a researcher at AI2, the research arm of the nonprofit Allen Institute, which focuses on artificial intelligence. “But as you are likely to give more accurate answers if you have a reference [like a book or a file] in front of you, the same applies in some cases to the models.’

RAG is undeniably useful — it allows one to attribute things a model generates to retrieved documents to verify their authenticity (and, as an added bonus, avoid potential copyright infringement). RAG also allows businesses that don’t want their documents used to train a model—say, companies in highly regulated industries like healthcare and law—to allow models to rely on those documents in a more secure and temporary way.

But RAG for sure slope stop a model from hallucinating. And it has limitations that many sellers ignore.

Wadden says RAG is most effective in “knowledge-intensive” scenarios where a user wants to use a model to address an “information need” — for example, to find out who won the Super Bowl last year. In these scenarios, the document that answers the question is likely to contain many of the same keywords as the question (eg “Super Bowl”, “last year”), making it relatively easy to find via keyword search .

Things get trickier with “reasoning-intensive” tasks like coding and math, where it’s harder to identify in a keyword-based search query the concepts needed to answer a query — much less identify which documents may be relevant.

Even with basic questions, models can be “distracted” by irrelevant content in documents, particularly long documents where the answer is not obvious. Or they may – for as yet unknown reasons – simply ignore the contents of retrieved documents, choosing instead to rely on their parametric memory.

RAG is also expensive in terms of the hardware required to implement it at scale.

This is because retrieved documents, whether from the web, an internal database, or somewhere else, must be stored in memory—at least temporarily—so that the model can refer to them. Another expense is accounting for the increased context a model must process before generating its response. For a technology already notorious for the amount of computation and electricity it requires even for basic functions, this is a serious consideration.

That’s not to say that RAG can’t be improved. Wadden noted several ongoing efforts to train models to make better use of documents retrieved from the RAG.

Some of these efforts include models that can “decide” when to make use of the documents, or models that can choose not to perform the retrieval in the first place if they deem it unnecessary. Others focus on ways to more efficiently index massive document datasets and improve search through better document representations—representations that go beyond keywords.

“We’re pretty good at retrieving documents based on keywords, but not so good at retrieving documents based on more abstract concepts, such as a proof technique needed to solve a math problem,” Wadden said. “Research is needed to create document representations and search techniques that can identify relevant documents for more abstract production tasks. I think that’s mostly an open question at this point.”

So RAG can help reduce a model’s hallucinations — but it’s not the answer to all of AI’s hallucination problems. Beware of any vendor who tries to claim otherwise.

All included Generative AI genetic hallucination problem hallucinations problem RAG solve wont
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleApple is adding more changes to its basic EU technology charge after criticism from developers
Next Article Climate tech investment returns with $8.1 billion start in 2024
bhanuprakash.cg
techtost.com
  • Website

Related Posts

OpenAI hits back at Google with GPT-5.2 after ‘code red’ memo.

14 December 2025

Trump’s AI executive order promises ‘a rulebook’ – startups may find legal loophole instead

13 December 2025

Google Translate now lets you listen to real-time translations on your headphones

13 December 2025
Add A Comment

Leave A Reply Cancel Reply

Don't Miss

Port raises $100M valuation from $800M round to take on Spotify’s Backstage

14 December 2025

India’s Spinny lines up $160m funding to acquire GoMechanic, sources say

14 December 2025

OpenAI hits back at Google with GPT-5.2 after ‘code red’ memo.

14 December 2025
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Fintech

Coinbase starts onboarding users again in India, plans to do fiat on-ramp next year

7 December 2025

Walmart-backed PhonePe shuts down Pincode app in yet another step back in e-commerce

5 December 2025

Nexus stays out of AI, keeping half of its new $700M fund for India startup

4 December 2025
Startups

Port raises $100M valuation from $800M round to take on Spotify’s Backstage

Eclipse Energy’s microbes can turn dormant oil wells into hydrogen factories

Interest in Spoor’s AI bird tracking software is soaring

© 2025 TechTost. All Rights Reserved
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer

Type above and press Enter to search. Press Esc to cancel.