Sound Research WIKINDX

WIKINDX Resources

Raji, I. D., Bender, E. M., Paullada, A., Denton, E., & Hanna, A. (2021). AI and the everything in the whole wide world benchmark. arXiv preprint arXiv:2111.15366. 
Added by: alexb44 (26/08/2026, 11:52)   Last edited by: alexb44 (26/08/2026, 12:03)
Resource type: Journal Article
Published
BibTeX citation key: Raji2021
Email resource to friend
View all bibliographic details
Categories: AI/Machine Learning
Keywords: Artificial General Intelligence, Artificial Intelligence
Creators: Bender, Denton, Hanna, Paullada, Raji
Collection: arXiv preprint arXiv:2111.15366
Views: 10/15
Abstract
There is a tendency across different subfields in AI to valorize a small collection of influential benchmarks. These benchmarks operate as stand-ins for a range of anointed common problems that are frequently framed as foundational mile- stones on the path towards flexible and generalizable AI systems. State-of-the-art performance on these benchmarks is widely understood as indicative of progress towards these long-term goals. In this position paper, we explore the limits of such benchmarks in order to reveal the construct validity issues in their framing as the functionally “general” broad measures of progress they are set up to be.
Added by: alexb44  
WIKINDX 6.17.0 | Total resources: 1551 | Username: -- | Bibliography: WIKINDX Master Bibliography | Style: American Psychological Association (APA) | Time Zone: Europe/Copenhagen (+02:00)