onlyTrustedInfo.comonlyTrustedInfo.comonlyTrustedInfo.com
Font ResizerAa
  • News
  • Finance
  • Sports
  • Life
  • Entertainment
  • Tech
Reading: Meta’s vanilla Maverick AI model ranks below rivals on a popular chat benchmark
Share
onlyTrustedInfo.comonlyTrustedInfo.com
Font ResizerAa
  • News
  • Finance
  • Sports
  • Life
  • Entertainment
  • Tech
Search
  • News
  • Finance
  • Sports
  • Life
  • Entertainment
  • Tech
  • Advertise
  • Advertise
© 2025 OnlyTrustedInfo.com . All Rights Reserved.
Tech

Meta’s vanilla Maverick AI model ranks below rivals on a popular chat benchmark

Last updated: April 11, 2025 6:46 pm
OnlyTrustedInfo.com
Share
2 Min Read
Meta’s vanilla Maverick AI model ranks below rivals on a popular chat benchmark
SHARE

Earlier this week, Meta landed in hot water for using an experimental, unreleased version of its Llama 4 Maverick model to achieve a high score on a crowdsourced benchmark, LM Arena. The incident prompted the maintainers of LM Arena to apologize, change their policies, and score the unmodified, vanilla Maverick.

Turns out, it’s not very competitive.

The unmodified Maverick, “Llama-4-Maverick-17B-128E-Instruct,” was ranked below models including OpenAI’s GPT-4o, Anthropic’s Claude 3.5 Sonnet, and Google’s Gemini 1.5 Pro as of Friday. Many of these models are months old.

The release version of Llama 4 has been added to LMArena after it was found out they cheated, but you probably didn’t see it because you have to scroll down to 32nd place which is where is ranks pic.twitter.com/A0Bxkdx4LX

— ρ:ɡeσn (@pigeon__s) April 11, 2025

Why the poor performance? Meta’s experimental Maverick, Llama-4-Maverick-03-26-Experimental, was “optimized for conversationality,” the company explained in a chart published last Saturday. Those optimizations evidently played well to LM Arena, which has human raters compare the outputs of models and choose which they prefer.

As we’ve written about before, for various reasons, LM Arena has never been the most reliable measure of an AI model’s performance. Still, tailoring a model to a benchmark — besides being misleading — makes it challenging for developers to predict exactly how well the model will perform in different contexts.

In a statement, a Meta spokesperson told TechCrunch that Meta experiments with “all types of custom variants.”

“‘Llama-4-Maverick-03-26-Experimental’ is a chat optimized version we experimented with that also performs well on LMArena,” the spokesperson said. “We have now released our open source version and will see how developers customize Llama 4 for their own use cases. We’re excited to see what they will build and look forward to their ongoing feedback.”

You Might Also Like

Apple Unveils Lower-Priced iPhone 16e With A.I. Features

From Secretive to Spectacle: How Transparent Vet Care at Turtle Back Zoo Redefines the Visitor Experience

Trump’s 25% Chip Tariff Is Just the Opening Salvo in a U.S. vs. Asia Silicon War

UK financial regulator partners with Nvidia in AI ‘sandbox’

Cloud Confidence Shaken: What the Microsoft Azure & 365 Outage Reveals About the Future of Digital Reliability

Share This Article
Facebook X Copy Link Print
Share
Previous Article Selena Gomez Returns To ‘Only Murders’ Set, Leaves Justin Bieber Noise Behind Selena Gomez Returns To ‘Only Murders’ Set, Leaves Justin Bieber Noise Behind
Next Article Wolverhampton Wanderers vs Tottenham Hotspur Prediction and Betting Tips Wolverhampton Wanderers vs Tottenham Hotspur Prediction and Betting Tips

Latest News

PFL Brussels 2026: Why the Odds Are Stacked Against the Underdogs in a Night of Dominant Favorites
PFL Brussels 2026: Why the Odds Are Stacked Against the Underdogs in a Night of Dominant Favorites
Sports May 23, 2026
Ja Morant Spotted at WNBA’s Dream vs. Wings: What His Presence Means for the NBA Star and Women’s Basketball
Ja Morant Spotted at WNBA’s Dream vs. Wings: What His Presence Means for the NBA Star and Women’s Basketball
Sports May 23, 2026
WWE Clash in Italy: Rhea Ripley vs. Jade Cargill Rematch Confirmed—Why This Title Showdown Matters
WWE Clash in Italy: Rhea Ripley vs. Jade Cargill Rematch Confirmed—Why This Title Showdown Matters
Sports May 23, 2026
Gerrit Cole’s Triumphant Return: 6 Shutout Innings After 569-Day Absence, But Yankees Fall to Rays
Gerrit Cole’s Triumphant Return: 6 Shutout Innings After 569-Day Absence, But Yankees Fall to Rays
Sports May 23, 2026
//
  • About Us
  • Contact US
  • Privacy Policy
onlyTrustedInfo.comonlyTrustedInfo.com
© 2026 OnlyTrustedInfo.com . All Rights Reserved.