Close Menu
TechTost
  • AI
  • Apps
  • Crypto
  • Fintech
  • Hardware
  • Media & Entertainment
  • Security
  • Startups
  • Transportation
  • Venture
  • Recommended Essentials
What's Hot

This $9 key physically locks your most addictive apps

Claude Opus 5 went completely rogue when he was tasked with operating a vending machine

Sorry, haters. Ferrari’s first EV is doing just fine

Facebook X (Twitter) Instagram
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer
Facebook X (Twitter) Instagram
TechTost
Subscribe Now
  • AI

    Hint, a new AI startup co-founded by Martha Stewart, offers an AI assistant to homeowners

    29 July 2026

    Data centers may experience temporary power outages to prevent power outages across the larger US grid

    28 July 2026

    Are brain waves the next unlock for natural artificial intelligence?

    27 July 2026

    Librarians host viral ‘Avoid AI’ workshops for people fed up with big tech

    26 July 2026

    I tested OpenAI’s new AI keyboard — which will be fun for some coders and a little overwhelming for everyone else

    25 July 2026
  • Apps

    Google brings age proofing technology to Android developers around the world

    29 July 2026

    Apple sued after alleged App Store encryption scam cost users $1.8 million

    28 July 2026

    Anthropic updates Claude voice mode with more capable models

    27 July 2026

    Bluesky’s AI assistant Attie expands into an open social research tool

    26 July 2026

    Why Cognition bought Poke: AI personality becomes a competitive advantage

    26 July 2026
  • Crypto

    Sam Altman’s biometrics startup World raises $52.5 million through crypto sale

    24 July 2026

    Venice AI goes unicorn with $65M Series A as first privacy AI platform takes off

    1 July 2026

    Crypto Exchange OKX wants AI agents to hire and pay each other

    30 June 2026

    Startup Battlefield 200 applications close today

    27 May 2026

    5 days left: Save up to $410 on Disrupt 2026 passes

    25 May 2026
  • Fintech

    TechCrunch Disrupt 2026’s new Smart Money Stage explores fintech, payments, artificial intelligence and everything

    25 July 2026

    Don’t want to invest in Elon Musk? Two new ETFs expressly exclude him

    10 July 2026

    India’s payments chief believes artificial intelligence will play a big part in the next era of digital payments development

    28 June 2026

    Early Bird pricing ends tonight for the Founder Summit

    26 June 2026

    4 days left to save up to $190 on Founder Summit 2026

    23 June 2026
  • Hardware

    This $9 key physically locks your most addictive apps

    30 July 2026

    Apple launches ‘Upgrade’ device rental program in partnership with Klarna

    29 July 2026

    Ozlo’s Sleepbuds 2 builds on Bose’s legacy of sleep headphones

    29 July 2026

    AI chip startup Etched defies skeptics, hits $10.3 billion valuation from big-name investors

    24 July 2026

    After a shocking quarter, IBM insists that artificial intelligence is not killing the mainframe

    23 July 2026
  • Media & Entertainment

    Winamp is aiming for a comeback with a new music player powered by Deezer

    30 July 2026

    HBO Max embraces vertical video with a new “Shorts” stream.

    29 July 2026

    Music streamer Deezer says more than 50% of daily uploads are generated by AI

    27 July 2026

    Substack’s new tool lets you know who’s writing their newsletters with AI

    26 July 2026

    Kalshi demands Netflix take down trailer for ‘Prediction Games’ documentary.

    26 July 2026
  • Security

    US government bans new foreign-made humanoids, robot dogs and solar inverters, citing national security risks

    29 July 2026

    Microsoft launches its first cybersecurity model, as well as a new cyber security agency system

    29 July 2026

    PSA: The conversations and artifacts shared by Claude may have ended up on Google

    28 July 2026

    The hacker who humiliated spyware makers and was never caught

    25 July 2026

    Hugging Face confirms breach of internal datasets and credentials, prompts users to take action

    25 July 2026
  • Startups

    Claude Opus 5 went completely rogue when he was tasked with operating a vending machine

    30 July 2026

    Antares raises $470 million to build nuclear reactors for the US military

    27 July 2026

    Insurance startup Corgi reportedly raises more money to $4 billion – its third round in 8 weeks

    26 July 2026

    Build publicly, fail publicly: what it’s like to be a founder under 20 right now

    25 July 2026

    Prentis, new AI lab co-founded by Reid Hoffman and Mark Pincus in talks to raise $100 million

    25 July 2026
  • Transportation

    Sorry, haters. Ferrari’s first EV is doing just fine

    30 July 2026

    Rivian is suing the US government for ‘full refund’ of Trump tariffs

    27 July 2026

    TechCrunch Mobility: Uber is betting on its former CEO

    26 July 2026

    Volkswagen engineers charged with insider trading linked to the Rivian consortium

    25 July 2026

    SpaceX launches new V3 Starlink satellites but suffers another booster failure

    25 July 2026
  • Venture

    Europe got its own TBPN-style live show and everyone is looking for a guest spot

    28 July 2026

    Edtech platform raises $4.5 million to help teach students how to code vibe

    23 July 2026

    Travis Kalanick’s robotics company raises $1.7 billion, led by a16z

    23 July 2026

    Cascade raises $3.5 million to help construction companies find and win projects

    22 July 2026

    StrictlyVC returns to New York on September 10 to celebrate a huge year for the city’s startup community

    21 July 2026
  • Recommended Essentials
TechTost
You are at:Home»AI»Meta releases Llama 3, it claims to be one of the best open models out there
AI

Meta releases Llama 3, it claims to be one of the best open models out there

techtost.comBy techtost.com18 April 202407 Mins Read
Share Facebook Twitter Pinterest LinkedIn Tumblr Email
Meta Releases Llama 3, It Claims To Be One Of
Share
Facebook Twitter LinkedIn Pinterest Email

Meta has was released the latest entry in Llama’s line of open AI models: Llama 3. Or, more accurately, the company debuted two models in its new Llama 3 family, with the rest coming at an unspecified future date.

Meta describes the new models — Llama 3 8B, which contains 8 billion parameters, and Llama 3 70B, which contains 70 billion parameters — as a “big leap” compared to the previous generation Llama models, Llama 2 8B and Llama 2 70B . performance. (Parameters essentially determine an AI model’s ability at a problem, such as text parsing and generation; models with a higher number of parameters are, generally speaking, more capable than models with a lower number of parameters.) In fact, Meta says that, for their respective parameter measurements, Llama 3 8B and Llama 3 70B — trained on two custom-built 24,000 GPU clusters — are is among the best performing AI models available today.

That’s quite a claim. So how does Meta support it? Well, the company notes the scores of the Llama 3 models on popular AI benchmarks like MMLU (which attempts to measure cognition), ARC (which attempts to measure skill acquisition), and DROP (which tests reasoning of a model in chunks of text). As we’ve written before, the usefulness — and validity — of these benchmarks is up for debate. But for better or worse, they remain one of the few standardized ways that AI players like Meta evaluate their models.

Llama 3 8B outperforms other open models such as Mistral’s Mistral 7B and Google’s Gemma 7B, which contain 7 billion parameters, in at least nine benchmarks: MMLU, ARC, DROP, GPQA (a set of biology, physics and chemistry-related questions), HumanEval (a code generation test), GSM-8K (math word problems), MATH (another math benchmark), AGIEval (a set of problem-solving tests), and BIG-Bench Hard (an assessment of joint reasoning logic).

Now, the Mistral 7B and Gemma 7B aren’t exactly on the cutting edge (Mistral 7B was released last September), and in some of the benchmarks reported by Meta, the Llama 3 8B scores only a few percentage points higher than the two. But Meta also claims that its higher-spec Llama 3 model, the Llama 3 70B, is competitive with flagship AI production models, including the Gemini 1.5 Pro, the latest in Google’s Gemini series.

Image Credits: After

The Llama 3 70B beats the Gemini 1.5 Pro in MMLU, HumanEval and GSM-8K, and — while not competing with Anthropic’s most efficient model, the Claude 3 Opus — the Llama 3 70B scores better than the second weakest model in the Claude 3 series , Claude 3 Sonnet, on five benchmarks (MMLU, GPQA, HumanEval, GSM-8K and MATH).

Meta Llama 3

Image Credits: After

For what it’s worth, Meta also developed its own test suite that covers use cases ranging from coding and creative writing to reasoning to summarization and — surprise! — Llama 3 70B beat Mistral’s Mistral Medium model, OpenAI’s GPT-3.5, and Claude Sonnet. Meta says it banned its modeling teams from accessing the set to maintain objectivity, but obviously—given that Meta devised the test itself—the results should be taken with a grain of salt.

Meta Llama 3

Image Credits: After

More qualitatively, Meta says users of the new Llama models should expect more “directionality,” a lower likelihood of refusing to answer questions, and greater accuracy on trivia questions, questions related to history, and STEM fields like engineering and science and general coding recommendations. This is due in part to a much larger data set: a collection of 15 trillion tokens, or a staggering ~750,000,000,000 words — seven times the size of the Llama 2 training set. (In the AI ​​field, “tokens” refers to subdivided bits raw data, such as the syllables “fan,” “tas,” and “tic” in the word “fantastic.”)

Where did this data come from? Good question. Meta didn’t say, only disclosing that it was pulling from “public sources”, included four times more code than in the Llama 2 training dataset, and that 5% of that set has non-English data (in ~30 languages) to improve performance in languages ​​other than English. Meta also said it used synthetic data – e.g. AI generated data – to create larger documents for Llama 3 models for training, a somewhat controversial approach because of the potential performance disadvantages.

“While the models we release today are only tuned for English results, the increased data diversity helps the models better recognize nuances and patterns and perform strongly on a variety of tasks,” Meta writes in a blog post shared on TechCrunch.

Many AI makers see training data as a competitive advantage and thus keep it and the information related to it close to the chest. But the details of the training data are also a potential source of intellectual property lawsuits, another disincentive to disclose much. Recent report revealed that Meta, in an effort to keep pace with AI competitors, at one point used copyrighted AI training e-books despite warnings from the company’s lawyers. Meta and OpenAI are the subject of an ongoing lawsuit by authors, including comedian Sarah Silverman, over the vendors’ alleged unauthorized use of copyrighted data for education.

So what about toxicity and bias, two other common problems with genetic AI models (including Llama 2)? Does Llama 3 improve in these areas? Yes, claims Meta.

Meta says it has developed new data filtering pipelines to boost the quality of its model training data and has updated its pair of productive AI security suites, Llama Guard and CybersecEval, to try to prevent misuse and unwanted text generation from Llama 3 models and more. The company is also releasing a new tool, Code Shield, designed to detect code from artificial intelligence models being created that might introduce security vulnerabilities.

However, filtering isn’t foolproof — and tools like Llama Guard, CyberSecEval, and Code Shield only go so far. (See: Llama 2’s tendency to they answer questions and leak private health and financial information.) We’ll have to wait and see how the Llama 3 models perform in the wild, including testing by academics on alternative benchmarks.

Meta says Llama 3 models – which are available to download now and power Meta’s Meta AI assistant on Facebook, Instagram, WhatsApp, Messenger and the web – will soon be hosted in managed form on a wide range of platforms cloud, including AWS. Databricks, Google Cloud, Hugging Face, Kaggle, IBM’s WatsonX, Microsoft Azure, Nvidia’s NIM and Snowflake. In the future, versions of the models optimized for hardware from AMD, AWS, Dell, Intel, Nvidia and Qualcomm will also be available.

Llama 3 models may be widely available. But you’ll notice we use “open” to describe them as opposed to “open source”. And that’s because, besides Meta’s claims, the Llama model family isn’t as close-knit as people would think. Yes, they are available for both research and commercial applications. However, Meta forbids Developers do not use Llama models to train other production models, and developers of apps with more than 700 million monthly users must request special permission from Meta which the company will — or will not — grant at its discretion.

More capable Llama models are on the horizon.

Meta says it’s currently training Llama 3 models with more than 400 billion parameters — models that can “conversate in multiple languages,” take in more data, and understand images and other formats as well as text, which will bring the Llama 3 series in line with with open releases like Hugging Face’s Idefics2.

Meta Llama 3

Image Credits: After

“Our goal in the near future is to make Llama 3 multilingual and multimodal, have a larger framework, and continue to improve overall performance across the core [large language model] capabilities like reasoning and coding,” Meta writes in a blog post. “More to come.”

Actually.

after All included blade 3 claims Generative AI Llama Meta models open open source releases
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
Previous ArticleMeta adds its AI chatbot, powered by Llama 3, to the search bar in its apps
Next Article Ibotta’s IPO opens sharply higher, hinting at growing public market interest in tech stocks
bhanuprakash.cg
techtost.com
  • Website

Related Posts

Hint, a new AI startup co-founded by Martha Stewart, offers an AI assistant to homeowners

29 July 2026

Data centers may experience temporary power outages to prevent power outages across the larger US grid

28 July 2026

Are brain waves the next unlock for natural artificial intelligence?

27 July 2026
Add A Comment

Leave A Reply Cancel Reply

Don't Miss

This $9 key physically locks your most addictive apps

30 July 2026

Claude Opus 5 went completely rogue when he was tasked with operating a vending machine

30 July 2026

Sorry, haters. Ferrari’s first EV is doing just fine

30 July 2026
Stay In Touch
  • Facebook
  • YouTube
  • TikTok
  • WhatsApp
  • Twitter
  • Instagram
Fintech

TechCrunch Disrupt 2026’s new Smart Money Stage explores fintech, payments, artificial intelligence and everything

25 July 2026

Don’t want to invest in Elon Musk? Two new ETFs expressly exclude him

10 July 2026

India’s payments chief believes artificial intelligence will play a big part in the next era of digital payments development

28 June 2026
Startups

Claude Opus 5 went completely rogue when he was tasked with operating a vending machine

Antares raises $470 million to build nuclear reactors for the US military

Insurance startup Corgi reportedly raises more money to $4 billion – its third round in 8 weeks

© 2026 TechTost. All Rights Reserved
  • About Us
  • Contact Us
  • Privacy Policy
  • Terms and Conditions
  • Disclaimer

Type above and press Enter to search. Press Esc to cancel.