الباحثون

Kelly Wan

المنشورات 2

نسخة أولية وصول مفتوح

Capability Scaling-Down Laws for LLM Compression

Xueqi Cheng, Liang Wu, Kelly Wan وآخرون · 2026

LLM compression reduces inference costs and memory requirements, but selecting a method and configuration remains largely empirical because comparable resource reductions can produce different capability losses. We systematically investigate capability scaling-down laws for LLM compression across pruning, quantization, …

نسخة أولية وصول مفتوح

AnyJev Technical Report

Jiamu Zhang, Tianze Yang, Yucheng Shi وآخرون · 2026

A typed decision is a choice among a fixed set of options, returned as a probability rather than as text. Systems that need typed decisions today use models trained for that purpose. This report describes AnyJev, which reads a typed decision from one prefill of a pretrained instruction-tuned language model. The readout …

المؤلفون المشاركون