Close Menu
TechTost
  • AI
  • Apps
  • Crypto
  • Fintech
  • Hardware
  • Media & Entertainment
  • Security
  • Startups
  • Transportation
  • Venture
  • Recommended Essentials
What's Hot

The biggest AI stories of the year (so far)

Travis Kalanick is launching a new company called Atoms that focuses on robotics

Founded by a father-son duo, Nyne gives AI agents the human context they’ve been missing

Facebook X (Twitter) Instagram
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
Facebook X (Twitter) Instagram
TechTost
Subscribe Now
  • AI

    ‘It wasn’t built right the first time’ — Musk’s xAI starts again, again

    14 March 2026

    Before quantum computing arrives, this startup wants businesses that are already working on it

    13 March 2026

    How to watch Jensen Huang’s Nvidia GTC 2026 keynote

    13 March 2026

    Ford’s new AI assistant will help fleet owners know if seat belts are being used

    12 March 2026

    AI ‘Actress’ Tilly Norwood Releases Worst Song I’ve Ever Heard

    12 March 2026
  • Apps

    Digg is laying off staff and shutting down the app as well as the company’s tools

    14 March 2026

    Truecaller now lets you hang up on scammers — on behalf of your family

    13 March 2026

    Channel Surfer lets you watch YouTube like it’s old-school cable TV

    13 March 2026

    Google Maps is getting an AI ‘Ask Maps’ feature and upgraded ‘immersive’ navigation

    12 March 2026

    Google Play adds new paid and PC games, game tests, community posts and more

    12 March 2026
  • Crypto

    Hackers stole over $2.7 billion in crypto in 2025, data shows

    23 December 2025

    New report examines how David Sachs may benefit from Trump administration role

    1 December 2025

    Why Benchmark Made a Rare Crypto Bet on Trading App Fomo, with $17M Series A

    6 November 2025

    Solana co-founder Anatoly Yakovenko is a big fan of agentic coding

    30 October 2025

    MoviePass opens Mogul fantasy league game to the public

    29 October 2025
  • Fintech

    India neobank Fi removes banking services on its platform

    11 March 2026

    X taps William Shatner to give invitations to his payment service, X Money

    4 March 2026

    Stripe wants to turn your AI costs into a profit center

    3 March 2026

    3 days left: Save up to $680 on your ticket to Disrupt 2026

    25 February 2026

    More startups surpass $10M ARR in 3 months than ever before

    24 February 2026
  • Hardware

    Ex-Apple Engineer Raises $5M for Note-Taking Locket That Only Records Your Voice

    12 March 2026

    Canopii seems to succeed where the old indoor farms failed

    11 March 2026

    Hyperscale Power is the latest startup to challenge 140-year-old transformer technology

    10 March 2026

    Whoop is launching a new blood test focused on women’s health

    10 March 2026

    Honor says its ‘Robot phone’ with moving camera can dance to music

    8 March 2026
  • Media & Entertainment

    Spotify will let you edit your taste profile to control your recommendations

    13 March 2026

    Disney+ launches TikTok-style short-form video stream ‘Verts’

    13 March 2026

    Substack launches an embedded recording studio

    12 March 2026

    TikTok now allows Apple Music subscribers to play entire songs without leaving the app

    12 March 2026

    WordPress debuts a private workspace that runs in your browser via a new service, my.WordPress.net

    11 March 2026
  • Security

    Law enforcement shuts down botnet consisting of tens of thousands of hacked routers

    12 March 2026

    The pro-Iranian hacktivist group says it is behind the attack on medical technology giant Stryker

    12 March 2026

    Salt Typhoon hacks the world’s phone and internet giants — here’s where they’ve been hit

    11 March 2026

    DOGE employee stole Social Security data and thumbed it, report says

    11 March 2026

    US military contractor likely built iPhone hacking tools used by Russian spies in Ukraine

    10 March 2026
  • Startups

    The biggest AI stories of the year (so far)

    14 March 2026

    Chinese brain interface startup Gestala raises $21 million just two months after launching

    13 March 2026

    Sales automation startup Rox AI hits $1.2 billion valuation, sources say

    13 March 2026

    When startups become a family business

    12 March 2026

    Ride-hailing inDrive acquires Pakistan’s Krave Mart to boost grocery delivery

    12 March 2026
  • Transportation

    Travis Kalanick is launching a new company called Atoms that focuses on robotics

    14 March 2026

    Kinetic robotics joins Uber’s Vegas app two years after major reset

    13 March 2026

    Why Rivian is holding onto the $45,000 R2 base model until ‘late 2027’

    13 March 2026

    Group14 opens factory to produce flash charge battery materials for EVs

    12 March 2026

    Nuro is testing its autonomous vehicle technology on the streets of Tokyo

    12 March 2026
  • Venture

    Founded by a father-son duo, Nyne gives AI agents the human context they’ve been missing

    14 March 2026

    Gumloop gets $50M from Benchmark to turn every worker into an AI agent builder

    13 March 2026

    This SpaceX Veteran Says The Next Big Thing In Space Is Satellites Returning To Earth

    10 March 2026

    Founders Fund is approaching $6 billion for its latest growth fund, sources say

    10 March 2026

    Robinhood’s startup fund stumbles in its NYSE debut

    7 March 2026
  • Recommended Essentials
TechTost
You are at:Home»AI»Openai’s research on AI models deliberate lies are wild
AI

Openai’s research on AI models deliberate lies are wild

techtost.comBy techtost.com19 September 202504 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email
Openai's Research On Ai Models Deliberate Lies Are Wild
Share
Facebook Twitter LinkedIn Pinterest Email

Every day, researchers in the largest technology companies fall a bomb. There was time that Google said that the latest quantum chip showed that there were multiple universes. Or when the man gave the Ai Claudius agent a snack sales machine to run and went amok, calling for safety to people and insisting he was human.

This week, it was Openai’s turn to increase our collective eyebrows.

Openai released Monday some survey explained How to stop AI models from “Scheming”. It is a practice in which an “AI behaves in a way on the surface while hiding its real goals”, Openai defined in his tweet for research.

In the document, conducted by Apollo’s research, the researchers went a little further, likening the AI ​​who was planning to a human stock market that breaks the law to make as much money as possible. The researchers, however, claimed that most AI “Scheming” were not so harmful. “The most common failures include simple forms of deception – for example, pretending to have completed a project without doing so,” they wrote.

The document was mostly published to show that the “continuing alignment”-the technique of the school that was tested-was well-tested.

But he also explained that AI developers have not found a way to train their models not to design. This is due to the fact that such education could really teach in the model how to design even better to avoid detection.

“An important way of failing to try to” train “Scheming is simply teaching the model to design more carefully and secretly,” the researchers wrote.

TechCrunch event

Francisco
|
27-29 October 2025

Perhaps the most amazing place is that if a model understands that it is being tested, it can pretend that it is not just planning to pass the test, even if it still is formed. “Models often become more aware that they are evaluated. This awareness of the situation can reduce Scheming itself, regardless of actual alignment,” the researchers wrote.

It’s not news that AI models will lie. So far most of us have experienced AI illusions, or the model with confidence, giving an answer to a prompt that is simply not true. But hallucinations show basic speculations with confidence, as the OpenAi survey released Earlier this month documented.

Scheming is something else. It is deliberate.

Even this revelation – that a model will deliberately mislead people – is not new. APOLLO research first Published a document in December documenting how the five models were formed when they were given instructions to achieve a “cost” goal.

The news here is really good news: The researchers saw significant reductions in the figure using “alignment”. This technique involves teaching the model a “protection specification” and then make the model review it before acting. It’s a bit like making young children repeat the rules before letting them play.

Openai researchers insist that the lies they have caught with their own models, or even Chatgpt, are not so serious. Like Openai co -founder Wojciech Zaremba, he told TechCrunch Maxwell Zeff for research: “This project has been done in the simulated environment and we think it represents future cases. Work. And this is just the lie.

The fact that AI models from many players deliberately deceive people is perhaps understandable. They were made by humans, to mimic people and (synthetic data) mostly trained in human -produced data.

They are also bonkers.

While we have all experienced the frustration of poor execution technology (who are thinking about you, printers in the house of the past), when was the last time the software that is not deliberately lied to you? Have your inbox ever built the emails on your own? Has the CMS that did not exist to place his numbers? Is your FinTech application its own banking transactions?

It is worth discussing, as the corporate world barrels to a future of AI where companies believe that agents can be treated as independent employees. The researchers of this document have the same warning.

“As AIs are assigned more complex tasks with real consequences and begin to seek more ambiguous, long-term goals, we expect the ability to be harmful to develop-so that our safeguards and ability to tasting strictly to grow respectively,” they wrote.

deliberate hallucinations lies models open OpenAIs Research Wild
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleThe concept launches agents for data analysis and automation
Next Article Dawn Capital’s Shamillah Bankiya is collapsing the status of business business markets
bhanuprakash.cg
techtost.com
  • Website

Related Posts

‘It wasn’t built right the first time’ — Musk’s xAI starts again, again

14 March 2026

Before quantum computing arrives, this startup wants businesses that are already working on it

13 March 2026

How to watch Jensen Huang’s Nvidia GTC 2026 keynote

13 March 2026
Add A Comment

Leave A Reply Cancel Reply

Don't Miss

The biggest AI stories of the year (so far)

14 March 2026

Travis Kalanick is launching a new company called Atoms that focuses on robotics

14 March 2026

Founded by a father-son duo, Nyne gives AI agents the human context they’ve been missing

14 March 2026
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Fintech

India neobank Fi removes banking services on its platform

11 March 2026

X taps William Shatner to give invitations to his payment service, X Money

4 March 2026

Stripe wants to turn your AI costs into a profit center

3 March 2026
Startups

The biggest AI stories of the year (so far)

Chinese brain interface startup Gestala raises $21 million just two months after launching

Sales automation startup Rox AI hits $1.2 billion valuation, sources say

© 2026 TechTost. All Rights Reserved
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer

Type above and press Enter to search. Press Esc to cancel.