ChinaTalk launched a $75,000 contest to crowdsource evaluation protocols for AI models used in diplomatic and national-security decision support
Benchmark result Provisional 74% confidence first seen
ChinaTalk announced a rolling contest with $75,000 in total prize money to develop and run evaluations for frontier AI models aimed at diplomatic and national-security decision-making. The contest solicits proposals by a September 1 deadline, which are intended to produce concrete benchmarking work, including how models perform on strategic and high-stakes questions over time, with input from listed researchers and labs.