Authors

Yidi Wu

Publications 1

Preprint Open access

ArchitectureIQ: On the Measure of Training Intuition

Top researchers have good intuition, but do language models have as good intuition about model training as top AI researchers? To measure model intuition of LLMs and humans, we introduce the ArchitectureIQ benchmark. Each question presents a synthetic dataset and several training recipes, and the test-taker is asked to …

Co-authors