الباحثون

Yanmeng Wang

المنشورات 2

نسخة أولية وصول مفتوح

PIVOT: Perplexity-Informed KD-to-RL Transition Scheduling for Vertical-Domain Few-Shot Distillation

Heng Li, Yong Zhang, Ning Cheng وآخرون · 2026

Vertical-domain few-shot classification remains challenging for small language models, as limited supervision makes it difficult to acquire domain-specific decision knowledge. On-Policy Distillation (OPD) can improve teacher-guided adaptation by supervising student-generated rollouts, while GRPO-based reinforcement lea …

نسخة أولية وصول مفتوح

Ask Without Telling: Local SLMs Consult Cloud LLMs Without Revealing Task Intent

Yanmeng Wang, Yunxuan Li, Shilong Fan وآخرون · 2026

As local small language models (SLMs) increasingly collaborate with more capable cloud large language models (LLMs), a natural privacy question arises: Can a local SLM obtain cloud LLM guidance while protecting user privacy? Existing privacy-preserving SLM-LLM frameworks primarily hide sensitive values while preserving …

المؤلفون المشاركون