Meta’s GAMUT benchmark evaluates AI factual completeness across 1,813 multimodal questions. Top-scoring Gemini 3.1 Pro managed just 58.7% among
The post Meta introduces GAMUT benchmark to measure factual completeness in AI appeared first on Crypto Briefing.






