نسخة أولية وصول مفتوح
GNN-CB: A Graph Neural Network Competition Benchmark for Human and LLM Evaluation
Large language models (LLMs) have demonstrated strong performance on coding and reasoning benchmarks; however, their ability to solve graph-structured machine learning problems remains largely unexplored. In particular, no benchmark currently evaluates whether LLMs can autonomously solve end-to-end Graph Neural Network …