الباحثون

Eunho Yang

المنشورات 4

نسخة أولية وصول مفتوح

Learning Beyond What You Sample: Off-Policy-Aware Cross-Model Trajectory Exchange for RLVR

Doohyuk Jang, Yoonsik Park, Gyouk Chu وآخرون · 2026

Reinforcement Learning with Verifiable Rewards (RLVR) methods such as GRPO rely on successful self-generated trajectories, but finite rollout budgets can produce all-fail groups with no reward-based policy-gradient signal. While additional rollouts improve the chance of success at higher cost, successful trajectories m …

نسخة أولية وصول مفتوح

SentZero: An Enhanced Sentence-Centric Vision-Language Pretraining for Multi-Task Zero-Shot Chest X-Ray Analysis

Hangyul Yoon, Hyungyung Lee, Edward Choi وآخرون · 2026

Vision-language (VL) pretraining using paired chest X-ray (CXR) images and radiology reports has shown strong potential for medical image understanding. However, existing methods often remain dependent on task-specific finetuning because radiology reports are lengthy, clinically dense, and difficult to align with simpl …

نسخة أولية وصول مفتوح

FlowTool: Controlling Tool Parameter in Image Retouching via Flow Matching

Tool-based image editing (image retouching) is commonly formulated with autoregressive multimodal large language models (MLLMs) that sequentially generate reasoning, tool selections, and parameter values. In this work, we present a novel approach to tool-based image editing by framing the task as a flow matching proble …

نسخة أولية وصول مفتوح

Looks the Same, Answers Differently: Flip-Direction Steering for Robust Vision-Language Reasoning

Yeonsung Jung, Joonhyun Jeong, Hoang Pham وآخرون · 2026

Vision-language models (VLMs) achieve strong visual reasoning performance, yet subtle changes from routine image capture and processing can alter their reasoning trajectories even when images appear nearly identical. In long-horizon generation, the resulting activation shifts may accumulate across decoding steps, progr …

المؤلفون المشاركون