AI’s capabilities may be exaggerated by flawed tests, study says

AI’s capabilities may be exaggerated by flawed tests, study says

Researchers behind a new study say that the methods used to evaluate AI systems’ capabilities routinely oversell AI performance and lack scientific rigor.

The study, led by researchers at the Oxford Internet Institute in partnership with over three dozen researchers from other institutions, examined 445 leading AI tests, called benchmarks, often used to measure the performance of AI models across a variety of topic areas.

AI developers and researchers use these benchmarks to evaluate model abilities and tout technical progress , referencing them to make claims on topics ranging from software engineering performance to abstract-reasoning capacity . However, the paper, released Tuesday, claims these fundamental tests might not be reliable and calls into question the validity of m

See Full Page

Interests (0)

Settings

AI’s capabilities may be exaggerated by flawed tests, study says

Wyoming Cows Go High-Tech With Electronic Collars And ‘Virtual Fencing’

MTG’s Boyfriend Sets Record Straight on Her Presidential Plans

Report: WA counties fail to keep guns from alleged abusers

DOJ tells Republicans that Epstein files even worse for Trump than they thought: report

Photos: Kim Kardashian Soaks up Family Fun at Kendall Jenner’s Birthday

Former NFL players join a new team, the Jacksonville Sheriff’s Office

Alex Ovechkin scores his 900th NHL goal with the Washington Capitals

Amanda Anisimova beats Iga Swiatek to join Elena Rybakina in last four at WTA Finals

Second Trump-designated holiday is next week: What's the significance?

US ends protected status for South Sudanese nationals

Jonathan Bailey calls People's Sexist Man Alive title 'the honour of a lifetime'

Food Distribution Event Draws Large Crowd Amid SNAP Shutdown

Supreme Court Questions Legality of Trump's Tariffs

Trump pressures GOP senators to end the government shutdown, now the longest ever

Democrat Jay Jones wins race to be Virginia attorney general despite texts endorsing violence

Dick Cheney, powerful former US vice president who pushed for Iraq war, dies at 84

Boebert's Halloween Costume Draws Criticism and Backlash

Mother of Missing Girl Allegedly Switched License Plates