Meta announced on 𝕏 that its AI model achieved gold or gold-level results in all five STEM Olympiads, with perfect scores in the theoretical exams of the Asian Physics Olympiad (APhO) and the International Physics Olympiad (IPhO). In the report card, IMO is explicitly listed as 'gold,' while IChO and RMM are listed as 'gold-level performance.' This wording typically implies that the results were not officially evaluated by the competition committees but were self-assessed against the year's gold medal cutoff scores. Meta also thanked the contestants and committees for their support, suggesting that participation in some events was coordinated with the committees, which is more formal than OpenAI's self-scoring approach in 2025.
Meta specifically emphasized that to test pure reasoning ability, the model was disabled from using all tools: no search, no coding, no calculators. This constraint is significant because in the past year, many top Olympiad results relied on agent pipelines of 'strong model + verifier + multi-round refinement.' For example, a researcher used Gemini 3.1 Pro to build a simple agent that achieved perfect scores on IPhO 2025 theoretical problems five times, but the author also noted that data contamination could not be ruled out. Meta deliberately abandoned this path, highlighting its pure reasoning positioning. However, physics competitions include experimental exams, and Meta's report likely covers only the theoretical portion.