الباحثون

Daniel Khashabi

المنشورات 5

نسخة أولية وصول مفتوح

Accurate but Not Humble: Evaluating Epistemic Humility in LLM Agents under Knowledge Conflict

When retrieved evidence contradicts an agent's prior beliefs, does it revise its answer, acknowledge uncertainty, or persist with an incorrect conclusion? Existing evaluations of agentic systems focus primarily on task success, offering limited insight into how agents handle such conflicts. We propose to evaluate agent …

نسخة أولية وصول مفتوح

CrossWeave: Bridging Perspectives Across Online Communities with a Dual-Pane Design

Fei Fang, Reva Hirave, William Jurayj وآخرون · 2026 · 10.1145/3785651.3831421

Social media systems typically display conversations among already familiar contributors, which can be predictable and one-sided. In civic discourse, this design narrows discussion, reinforces divides, and distorts the perception of public opinion. To encourage cross-community engagement, we present CrossWeave, an AI-p …

نسخة أولية وصول مفتوح

JEPA-TTT: Persistent Test-Time Training of Latent World Models for Planning under Dynamics Shifts

Zheyuan Zhang, Suyu Ye, Nakul Agarwal وآخرون · 2026

World models enable agents to plan by predicting future states of the environment, but their predictions can become unreliable when test-time dynamics differ from those seen during training. We present JEPA-TTT, which adapts the latent dynamics predictor of a pretrained action-conditioned Joint-Embedding Predictive Arc …

نسخة أولية وصول مفتوح

Harness Learning Enables Generalizable Test-Time Adaptation

Alvin Zhang, Xuecheng Liu, Zixuan Wang وآخرون · 2026

A language-model agent is jointly defined by its model and its harness, the executable program that organizes model calls, tool use, and information flow. Because different tasks call for different ways of organizing these operations, the harness needs to be adapted using feedback from the task at hand. We introduce ha …

المؤلفون المشاركون