Close Menu
TechTost
  • AI
  • Apps
  • Crypto
  • Fintech
  • Hardware
  • Media & Entertainment
  • Security
  • Startups
  • Transportation
  • Venture
  • Recommended Essentials
What's Hot

Are brain waves the next unlock for natural artificial intelligence?

Anthropic updates Claude voice mode with more capable models

Music streamer Deezer says more than 50% of daily uploads are generated by AI

Facebook X (Twitter) Instagram
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
Facebook X (Twitter) Instagram
TechTost
Subscribe Now
  • AI

    Are brain waves the next unlock for natural artificial intelligence?

    27 July 2026

    Librarians host viral ‘Avoid AI’ workshops for people fed up with big tech

    26 July 2026

    I tested OpenAI’s new AI keyboard — which will be fun for some coders and a little overwhelming for everyone else

    25 July 2026

    Anthropic Launches Opus 5 | TechCrunch

    24 July 2026

    How AI guardrails are hindering the work of aggressive cybersecurity researchers

    24 July 2026
  • Apps

    Anthropic updates Claude voice mode with more capable models

    27 July 2026

    Bluesky’s AI assistant Attie expands into an open social research tool

    26 July 2026

    Why Cognition bought Poke: AI personality becomes a competitive advantage

    26 July 2026

    Vietnam seeks to restrict social media for children. Here are the growing number of other countries doing the same

    25 July 2026

    India’s move against Jack Dorsey’s Bitchat sparks legal debate

    24 July 2026
  • Crypto

    Sam Altman’s biometrics startup World raises $52.5 million through crypto sale

    24 July 2026

    Venice AI goes unicorn with $65M Series A as first privacy AI platform takes off

    1 July 2026

    Crypto Exchange OKX wants AI agents to hire and pay each other

    30 June 2026

    Startup Battlefield 200 applications close today

    27 May 2026

    5 days left: Save up to $410 on Disrupt 2026 passes

    25 May 2026
  • Fintech

    TechCrunch Disrupt 2026’s new Smart Money Stage explores fintech, payments, artificial intelligence and everything

    25 July 2026

    Don’t want to invest in Elon Musk? Two new ETFs expressly exclude him

    10 July 2026

    India’s payments chief believes artificial intelligence will play a big part in the next era of digital payments development

    28 June 2026

    Early Bird pricing ends tonight for the Founder Summit

    26 June 2026

    4 days left to save up to $190 on Founder Summit 2026

    23 June 2026
  • Hardware

    AI chip startup Etched defies skeptics, hits $10.3 billion valuation from big-name investors

    24 July 2026

    After a shocking quarter, IBM insists that artificial intelligence is not killing the mainframe

    23 July 2026

    Light made a flip phone — it’s colorful and cheap

    22 July 2026

    Apple is partnering with Klarna to launch a rental program for iPhones, iPads and Macs

    22 July 2026

    The Xteink X4 Pro could be the tiny e-reader of your dreams

    21 July 2026
  • Media & Entertainment

    Music streamer Deezer says more than 50% of daily uploads are generated by AI

    27 July 2026

    Substack’s new tool lets you know who’s writing their newsletters with AI

    26 July 2026

    Kalshi demands Netflix take down trailer for ‘Prediction Games’ documentary.

    26 July 2026

    Amazon brings games to Prime Video

    24 July 2026

    SoundCloud acquires decentralized music platform Nina Protocol months after its shutdown

    23 July 2026
  • Security

    The hacker who humiliated spyware makers and was never caught

    25 July 2026

    Hugging Face confirms breach of internal datasets and credentials, prompts users to take action

    25 July 2026

    US accuses American of allegedly wiping his phone using a passcode ‘forcibly’ during border search

    24 July 2026

    If you pay a hacker’s ransom, chances are they’ll come back for more

    24 July 2026

    The US government says hackers linked to Iran are disrupting US water and energy providers

    23 July 2026
  • Startups

    Insurance startup Corgi reportedly raises more money to $4 billion – its third round in 8 weeks

    26 July 2026

    Build publicly, fail publicly: what it’s like to be a founder under 20 right now

    25 July 2026

    Prentis, new AI lab co-founded by Reid Hoffman and Mark Pincus in talks to raise $100 million

    25 July 2026

    Meet the judges who will crown Australia’s next startup

    24 July 2026

    AegisAI, founded by ex-Google security execs, raises $36M to stop AI-based spearfishing

    23 July 2026
  • Transportation

    TechCrunch Mobility: Uber is betting on its former CEO

    26 July 2026

    Volkswagen engineers charged with insider trading linked to the Rivian consortium

    25 July 2026

    SpaceX launches new V3 Starlink satellites but suffers another booster failure

    25 July 2026

    Tesla’s door handles may prompt new safety rules in the US

    24 July 2026

    Tesla’s robotaxis moves in reverse

    23 July 2026
  • Venture

    Edtech platform raises $4.5 million to help teach students how to code vibe

    23 July 2026

    Travis Kalanick’s robotics company raises $1.7 billion, led by a16z

    23 July 2026

    Cascade raises $3.5 million to help construction companies find and win projects

    22 July 2026

    StrictlyVC returns to New York on September 10 to celebrate a huge year for the city’s startup community

    21 July 2026

    Startup Inference Infinity raises $15 million from researchers Touring Capital, OpenAI and Anthropic

    20 July 2026
  • Recommended Essentials
TechTost
You are at:Home»AI»Anthropic wants to fund a new, more comprehensive generation of AI benchmarks
AI

Anthropic wants to fund a new, more comprehensive generation of AI benchmarks

techtost.comBy techtost.com2 July 202404 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email
Anthropic Wants To Fund A New, More Comprehensive Generation Of
Share
Facebook Twitter LinkedIn Pinterest Email

Anthropic is launching one program to fund the development of new types of benchmarks capable of evaluating the performance and impact of AI models, including production models like Claude’s.

Anthropic’s program, unveiled Monday, will make payments to third-party organizations that can, as the company puts it in a blog post, “effectively measure advanced capabilities in artificial intelligence models.” Interested parties may submit applications for evaluation on a rolling basis.

“Our investment in these assessments is intended to elevate the entire field of AI security, providing valuable tools that benefit the entire ecosystem,” Anthropic wrote on its official blog. “Developing high-quality, safety-relevant assessments remains a challenge, and demand outstrips supply.”

As we’ve pointed out before, AI has a benchmarking problem. The most commonly cited AI benchmarks today do a poor job of capturing how the average human actually uses the systems under review. There are also questions about whether some benchmarks, particularly those released before the dawn of modern genetic artificial intelligence, even measure what they are supposed to measure, given their age.

The very high-level, harder-than-it-sounds solution proposed by Anthropic creates challenging benchmarks with an emphasis on AI security and social impact through new tools, infrastructure and methods.

The company specifically requests tests that assess a model’s ability to perform tasks such as carrying out cyber attacks, “enhancing” weapons of mass destruction (e.g. nuclear weapons), and manipulating or deceiving people (e.g. via deepfakes or disinformation). For AI risks related to national security and defense, Anthropic says it’s committed to developing some kind of “early warning system” to identify and assess risks, though it didn’t reveal in the blog post what it might to imply such a system.

Anthropic also says it intends its new program to support benchmark research and “end-to-end” work that explores the potential of artificial intelligence to aid scientific study, converse in multiple languages, and mitigate entrenched biases. as well as toxicity self-censoring.

To achieve all this, Anthropic envisions new platforms that allow subject matter experts to develop their own assessments and large-scale model tests involving “thousands” of users. The company says it has hired a full-time coordinator for the program and may buy or expand projects it believes have the potential to scale.

“We offer a range of financing options tailored to the needs and stage of each project,” Anthropic writes in the post, though an Anthropic spokesperson declined to elaborate on those options. “Teams will have the opportunity to interact directly with Anthropic domain experts from the frontier red team, detail, trust and security and other relevant teams.”

Anthropic’s effort to support new AI benchmarks is commendable — assuming, of course, that there’s enough cash and manpower behind it. But given the company’s commercial ambitions in the AI ​​race, it may be hard to fully trust.

In the blog post, Anthropic is rather transparent about the fact that it wants some of the assessments it funds to align with AI Security Classifications the developed (with some input from third parties, such as the non-profit AI research organization METR). This is within the company’s prerogative. But it may also force applicants to the program to accept definitions of “safe” or “dangerous” AI with which they may not agree.

A portion of the AI ​​community is also likely to take issue with Anthropic’s references to “catastrophic” and “misleading” AI risks, such as the dangers of nuclear weapons. Many experts let’s just say there’s little evidence to suggest that AI as we know it will achieve global human-surpassing capabilities anytime soon, if ever. Claims of impending “superintelligence” only serve to draw attention away from pressing AI regulatory issues of the day, such as AI’s hallucinatory tendencies, these experts add.

In its post, Anthropic writes that it hopes its program will serve as a “catalyst for progress toward a future where comprehensive AI assessment is an industry standard.” This is a mission that many have opened, corporate-unaffiliated efforts to create better AI benchmarks can be identified. But it remains to be seen whether those efforts are willing to join forces with an AI vendor whose loyalty ultimately rests with shareholders.

All included Anthropic benchmarks comprehensive fund generation Humane reference points
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleApple is finally adding support for RCS in the latest iOS 18 beta
Next Article Japan’s SmartHR Raises $140M Series E as Strong Demand for HR Tech Boosts ARR to $100M
bhanuprakash.cg
techtost.com
  • Website

Related Posts

Are brain waves the next unlock for natural artificial intelligence?

27 July 2026

Anthropic updates Claude voice mode with more capable models

27 July 2026

Librarians host viral ‘Avoid AI’ workshops for people fed up with big tech

26 July 2026
Add A Comment

Leave A Reply Cancel Reply

Don't Miss

Are brain waves the next unlock for natural artificial intelligence?

27 July 2026

Anthropic updates Claude voice mode with more capable models

27 July 2026

Music streamer Deezer says more than 50% of daily uploads are generated by AI

27 July 2026
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Fintech

TechCrunch Disrupt 2026’s new Smart Money Stage explores fintech, payments, artificial intelligence and everything

25 July 2026

Don’t want to invest in Elon Musk? Two new ETFs expressly exclude him

10 July 2026

India’s payments chief believes artificial intelligence will play a big part in the next era of digital payments development

28 June 2026
Startups

Insurance startup Corgi reportedly raises more money to $4 billion – its third round in 8 weeks

Build publicly, fail publicly: what it’s like to be a founder under 20 right now

Prentis, new AI lab co-founded by Reid Hoffman and Mark Pincus in talks to raise $100 million

© 2026 TechTost. All Rights Reserved
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer

Type above and press Enter to search. Press Esc to cancel.