Close Menu
TechTost
  • AI
  • Apps
  • Crypto
  • Fintech
  • Hardware
  • Media & Entertainment
  • Security
  • Startups
  • Transportation
  • Venture
  • Recommended Essentials
What's Hot

From Svedka to Anthropic, Brands Are Making Bold Plays With AI in Super Bowl Ads

Accel doubles down on Fibr AI as agents turn static websites into one-to-one experiences

SNAK Venture Partners raises $50 million in capital to support vertical acquisitions

Facebook X (Twitter) Instagram
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
Facebook X (Twitter) Instagram
TechTost
Subscribe Now
  • AI

    Benchmark raises $225 million in dedicated funds to double Cerebras

    7 February 2026

    How artificial intelligence is helping to solve the labor issue in treating rare diseases

    6 February 2026

    Amazon and Google are winning the AI ​​capital race — but what’s the prize?

    6 February 2026

    AWS revenue continues to grow as cloud demand remains high

    5 February 2026

    Sam Altman tested Claude’s Super Bowl commercials brilliantly

    5 February 2026
  • Apps

    EU says TikTok must disable ‘addictive’ features like infinite scrolling, fix recommendation engine

    7 February 2026

    Here’s how Roblox’s age controls work

    6 February 2026

    Meta is testing a standalone app for its AI-generated ‘Vibes’ videos

    6 February 2026

    Reddit sees AI search as the next big opportunity

    5 February 2026

    Tinder looks to AI to help fight dating app ‘fatigue’ and burnout

    5 February 2026
  • Crypto

    Hackers stole over $2.7 billion in crypto in 2025, data shows

    23 December 2025

    New report examines how David Sachs may benefit from Trump administration role

    1 December 2025

    Why Benchmark Made a Rare Crypto Bet on Trading App Fomo, with $17M Series A

    6 November 2025

    Solana co-founder Anatoly Yakovenko is a big fan of agentic coding

    30 October 2025

    MoviePass opens Mogul fantasy league game to the public

    29 October 2025
  • Fintech

    Stripe Alumni Raise €30M Series A for Duna, Backed by Stripe and Adyen Executives

    5 February 2026

    Fintech CEO and Forbes 30 Under 30 alum indicted for alleged fraud

    3 February 2026

    How Sequoia-backed Ethos went public while rivals lagged behind

    30 January 2026

    5 days left for TechCrunch Disrupt 2026 +1 pass with 50%

    26 January 2026

    50% off +1 ends | TechCrunch

    23 January 2026
  • Hardware

    Kindle Scribe Colorsoft is an expensive but beautiful color e-ink tablet with AI features

    6 February 2026

    Ring brings “Search Party” feature for finding lost dogs to non-Ring camera owners

    2 February 2026

    India offers zero taxes till 2047 to attract global AI workloads

    1 February 2026

    Microsoft won’t stop buying AI chips from Nvidia, AMD even after its own is released, says Nadella

    30 January 2026

    The iPhone just had its best quarter ever

    30 January 2026
  • Media & Entertainment

    From Svedka to Anthropic, Brands Are Making Bold Plays With AI in Super Bowl Ads

    7 February 2026

    “Industry” Season 4 captures tech fraud better than any show on TV right now

    7 February 2026

    Spotify’s new feature lets you explore the story behind the song you’re listening to

    6 February 2026

    The Washington Post retreats from Silicon Valley when it matters most

    6 February 2026

    Spotify is in the business of selling books and adding new audiobook features

    5 February 2026
  • Security

    Senator, who has repeatedly warned of secret US government surveillance, raises new alarm over ‘CIA activities’

    7 February 2026

    Substack confirms that the data breach affects users’ email addresses and phone numbers

    6 February 2026

    One of Europe’s biggest universities was offline for days after the cyber attack

    6 February 2026

    Cyber ​​tech giant Conduent’s hot air balloon data breach affects millions more Americans

    5 February 2026

    Hackers Release Personal Information Stolen During Harvard, UPenn Data Breach

    5 February 2026
  • Startups

    Accel doubles down on Fibr AI as agents turn static websites into one-to-one experiences

    7 February 2026

    ElevenLabs Raises $500M From Sequoia At $11B Valuation

    7 February 2026

    Fundamental raises $255 million in Series A with a new approach to big data analytics

    6 February 2026

    a16z VC wants founders to stop stressing about crazy ARR numbers

    6 February 2026

    Lunar Energy raises $232 million to develop home batteries that support the grid

    5 February 2026
  • Transportation

    Prince Andrew’s adviser suggested Jeffrey Epstein invest in EV startups like Lucid Motors

    7 February 2026

    Apeiron Labs Takes $9.5M to Flood Oceans with Autonomous Underwater Robots

    5 February 2026

    Uber appoints new CFO as its AV plans accelerate

    5 February 2026

    Skyryse lands another $300 million to make flying, even helicopters, simple and safe

    4 February 2026

    China is leading the fight against hidden car door handles

    3 February 2026
  • Venture

    SNAK Venture Partners raises $50 million in capital to support vertical acquisitions

    7 February 2026

    Reddit says it’s looking for more acquisitions in adtech and elsewhere

    7 February 2026

    Secondary sales are shifting from founders’ windfalls to employee retention tools

    6 February 2026

    Sapiom Raises $15M to Help AI Agents Buy Their Own Tech Tools

    6 February 2026

    What a16z actually funds (and what it ignores) when it comes to AI infra

    5 February 2026
  • Recommended Essentials
TechTost
You are at:Home»AI»Why RAG Won’t Solve the Problem of Genetic AI Hallucinations
AI

Why RAG Won’t Solve the Problem of Genetic AI Hallucinations

techtost.comBy techtost.com5 May 202405 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email
Why Rag Won't Solve The Problem Of Genetic Ai Hallucinations
Share
Facebook Twitter LinkedIn Pinterest Email

Illusions – the lies that artificial intelligence models are basically telling – are a big problem for businesses looking to integrate the technology into their operations.

Because models have no real intelligence and simply predict words, images, speech, music and other data according to a private schema, they sometimes get it wrong. Very wrong. In a recent article in the Wall Street Journal, a source recounts an example where Microsoft’s genetic AI invented meeting participants and implied that the conference calls were about topics that were not actually discussed on the call.

As I wrote a while back, illusions can be an intractable problem with today’s transformer-based model architectures. But a number of leading AI vendors are proposing to do so can to be more or less eliminated through a technical approach called recovery augmented production or RAG.

Here’s how one salesman, Squirro, he drops it:

At the core of the offering is the concept of Retrieval Augmented LLMs or Retrieval Augmented Generation (RAG) built into the solution… [our generative AI] is unique in its promise of zero hallucinations. Every piece of information it generates is traceable to a source, ensuring reliability.

Here is one similar step from SiftHub:

Using RAG technology and enhanced large language models with industry-specific knowledge training, SiftHub enables companies to create personalized responses with zero hallucinations. This guarantees increased transparency and reduced risk and inspires complete confidence in using AI for all their needs.

RAG was pioneered by data scientist Patrick Lewis, a researcher at Meta and University College London and lead author of the 2020 paper who coined the term. Applied to a model, RAG retrieves documents possibly related to a question — for example, a Wikipedia page about the Super Bowl — using what is essentially a keyword search, and then asks the model to generate answers to this additional context.

“When you interact with a generative AI model like ChatGPT or Llama and ask a question, the default is for the model to answer from its ‘parametric memory’ — that is, from the knowledge stored in its parameters as a result of training on huge data from the web,” explained David Wadden, a researcher at AI2, the research arm of the nonprofit Allen Institute, which focuses on artificial intelligence. “But as you are likely to give more accurate answers if you have a reference [like a book or a file] in front of you, the same applies in some cases to the models.’

RAG is undeniably useful — it allows one to attribute things a model generates to retrieved documents to verify their authenticity (and, as an added bonus, avoid potential copyright infringement). RAG also allows businesses that don’t want their documents used to train a model—say, companies in highly regulated industries like healthcare and law—to allow models to rely on those documents in a more secure and temporary way.

But RAG for sure slope stop a model from hallucinating. And it has limitations that many sellers ignore.

Wadden says RAG is most effective in “knowledge-intensive” scenarios where a user wants to use a model to address an “information need” — for example, to find out who won the Super Bowl last year. In these scenarios, the document that answers the question is likely to contain many of the same keywords as the question (eg “Super Bowl”, “last year”), making it relatively easy to find via keyword search .

Things get trickier with “reasoning-intensive” tasks like coding and math, where it’s harder to identify in a keyword-based search query the concepts needed to answer a query — much less identify which documents may be relevant.

Even with basic questions, models can be “distracted” by irrelevant content in documents, particularly long documents where the answer is not obvious. Or they may – for as yet unknown reasons – simply ignore the contents of retrieved documents, choosing instead to rely on their parametric memory.

RAG is also expensive in terms of the hardware required to implement it at scale.

This is because retrieved documents, whether from the web, an internal database, or somewhere else, must be stored in memory—at least temporarily—so that the model can refer to them. Another expense is accounting for the increased context a model must process before generating its response. For a technology already notorious for the amount of computation and electricity it requires even for basic functions, this is a serious consideration.

That’s not to say that RAG can’t be improved. Wadden noted several ongoing efforts to train models to make better use of documents retrieved from the RAG.

Some of these efforts include models that can “decide” when to make use of the documents, or models that can choose not to perform the retrieval in the first place if they deem it unnecessary. Others focus on ways to more efficiently index massive document datasets and improve search through better document representations—representations that go beyond keywords.

“We’re pretty good at retrieving documents based on keywords, but not so good at retrieving documents based on more abstract concepts, such as a proof technique needed to solve a math problem,” Wadden said. “Research is needed to create document representations and search techniques that can identify relevant documents for more abstract production tasks. I think that’s mostly an open question at this point.”

So RAG can help reduce a model’s hallucinations — but it’s not the answer to all of AI’s hallucination problems. Beware of any vendor who tries to claim otherwise.

All included Generative AI genetic hallucination problem hallucinations problem RAG solve wont
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleApple is adding more changes to its basic EU technology charge after criticism from developers
Next Article Climate tech investment returns with $8.1 billion start in 2024
bhanuprakash.cg
techtost.com
  • Website

Related Posts

Benchmark raises $225 million in dedicated funds to double Cerebras

7 February 2026

How artificial intelligence is helping to solve the labor issue in treating rare diseases

6 February 2026

Amazon and Google are winning the AI ​​capital race — but what’s the prize?

6 February 2026
Add A Comment

Leave A Reply Cancel Reply

Don't Miss

From Svedka to Anthropic, Brands Are Making Bold Plays With AI in Super Bowl Ads

7 February 2026

Accel doubles down on Fibr AI as agents turn static websites into one-to-one experiences

7 February 2026

SNAK Venture Partners raises $50 million in capital to support vertical acquisitions

7 February 2026
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Fintech

Stripe Alumni Raise €30M Series A for Duna, Backed by Stripe and Adyen Executives

5 February 2026

Fintech CEO and Forbes 30 Under 30 alum indicted for alleged fraud

3 February 2026

How Sequoia-backed Ethos went public while rivals lagged behind

30 January 2026
Startups

Accel doubles down on Fibr AI as agents turn static websites into one-to-one experiences

ElevenLabs Raises $500M From Sequoia At $11B Valuation

Fundamental raises $255 million in Series A with a new approach to big data analytics

© 2026 TechTost. All Rights Reserved
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer

Type above and press Enter to search. Press Esc to cancel.