الباحثون

Michael Witbrock

المنشورات 2

نسخة أولية وصول مفتوح

Loop-Free Inverse Reinforcement Learning via Sequential Value Recovery with Q-Score Matching

Yang Chen, Yitan Zhang, Michael Witbrock وآخرون · 2026

Inverse Reinforcement Learning (IRL) aims to recover a reward function that explains expert demonstrations. Existing IRL methods typically rely on a bi-level optimization procedure that alternates between reward learning and policy optimization, leading to substantial computational burden and training instability. In t …

نسخة أولية وصول مفتوح

PunGraph: Retrieval-Enhanced Phonetic-Semantic Graph Reasoning for Pun Understanding

Yuchen Su, Zijian Huang, Yaotian Shi وآخرون · 2026

Puns are a challenging form of figurative language that exploit phonetic similarity and semantic ambiguity to convey multiple meanings. Although large language models (LLMs) demonstrate strong language understanding capabilities, they still struggle with pun reasoning due to limited phonetic modeling and uncontrolled e …

المؤلفون المشاركون