AutoBencher: Creating Salient, Novel, Difficult Datasets for Language Modelsarxiv.org 2Garcia982ydiscuss
Air Force Official's Story of Killer AI Was a Hypothetical Situationbusinessinsider.com 1Garcia983ydiscuss
Chain-of-Thought Hub: Measuring LLMs' Reasoning Performancegithub.com/FranxYao 114Garcia983y26 comments
Spain calls snap election after conservative and far-right wins in local pollstheguardian.com 2Garcia983ydiscuss
Impossible Distillation: From Low-Quality Model to High-Quality Dataset & Modelarxiv.org 1Garcia983ydiscuss
Microsoft Build brings AI tools to the forefront for developersblogs.microsoft.com 13Garcia983y3 comments