Llama Drama: Meta Used An ‘Experimental’ AI Model To Climb Leaderboards, Raising Questions About Fairness, Transparency, And What Users Actually Get To Use

Meta launched two new versions of its Llama 4 AI over the weekend, including a smaller model called Scout and a mid-sized model called Maverick. The company claimed that its latter model outperformed ChatGPT-4o and Gemini 2.0 Flash on many popular tests, but it appears that there is something the company did not tell the testers, or did it? Meta faces backlash for using a custom-tuned AI model in public benchmarks, prompting accusations of misleading performance claims Meta's Maverick gained the second spot on LMArena soon after its launch, climbing the leaderboard in an attempt to take the throne for […]