الملخص

Which forms of test-time compute improve the predictions of strong pretrained tabular foundation models (TFMs)? We systematically study this along three axes: adaptation, aggregation, and context construction. Our evaluation spans modern TFMs across the TabArena benchmark, supplemented by experiments on wide and large-scale tables from OpenML. For adaptation, we introduce DiagScale, a diagonal query-key similarity update. It trains only 0.003-0.03% of model parameters and achieves gains comparable to full fine-tuning across three independently pretrained backbones. For aggregation, both pool composition and selection strategy matter. TabPFN-3 already averages predictions from different preprocessing variants of the same data, and adding more such predictions yields diminishing returns. With a broader pool of 96 configurations, greedy selection reduces error by 2.4% relative to the default predictor, but uniform averaging increases error. For context construction, attention-guided retrieval improves TabPFN-3's predictions on some large tables and supports source pools beyond the full context memory limit. The context expansion methods we test yield no consistent improvement. Taken together, our results suggest that adaptation and selective aggregation yield consistent benchmark-level gains. The benefits of context construction depend more on the task and data regime. Adaptation and aggregation over the same backbone yield further gains when combined, but require substantially more computation than default inference. These trade-offs motivate choosing strategies according to the available computation budget. Code is available at https://github.com/kanghui-learning/test-time-compute-for-tabular-foundation-models.

الكلمات المفتاحية

الموضوع

بيانات النشر

المجلة
غير متاح
وصول مفتوح
وصول مفتوح أخضر

اقتبس هذه المقالة

APA 7

Ning, K., Biloš, M., Wilson, J. T., Zhang, Y., Rasul, K., Song, D., Schneider, A., & Nevmyvaka, Y. (2026). Test-Time Compute for Tabular Foundation Models: Mechanisms, Gains, and Limits. https://omanscience.com/ar/articles/test-time-compute-for-tabular-foundation-models-mechanisms-gains-and-limits

MLA 9

Ning, Kanghui, et al. "Test-Time Compute for Tabular Foundation Models: Mechanisms, Gains, and Limits." https://omanscience.com/ar/articles/test-time-compute-for-tabular-foundation-models-mechanisms-gains-and-limits.

شيكاغو (المؤلف–التاريخ)

Ning, Kanghui, Marin Biloš, James T. Wilson, Yilang Zhang, Kashif Rasul, Dongjin Song, Anderson Schneider, and Yuriy Nevmyvaka. 2026. "Test-Time Compute for Tabular Foundation Models: Mechanisms, Gains, and Limits." https://omanscience.com/ar/articles/test-time-compute-for-tabular-foundation-models-mechanisms-gains-and-limits.

هارفارد

Ning, K., Biloš, M., Wilson, J. T., Zhang, Y., Rasul, K., Song, D., Schneider, A. and Nevmyvaka, Y. (2026) 'Test-Time Compute for Tabular Foundation Models: Mechanisms, Gains, and Limits', Available at: https://omanscience.com/ar/articles/test-time-compute-for-tabular-foundation-models-mechanisms-gains-and-limits.

فانكوفر

Ning K, Biloš M, Wilson JT, Zhang Y, Rasul K, Song D, et al. Test-Time Compute for Tabular Foundation Models: Mechanisms, Gains, and Limits. https://omanscience.com/ar/articles/test-time-compute-for-tabular-foundation-models-mechanisms-gains-and-limits

IEEE

K. Ning, M. Biloš, J. T. Wilson, Y. Zhang, K. Rasul, D. Song, A. Schneider, and Y. Nevmyvaka, "Test-Time Compute for Tabular Foundation Models: Mechanisms, Gains, and Limits," https://omanscience.com/ar/articles/test-time-compute-for-tabular-foundation-models-mechanisms-gains-and-limits.