Authors

Jaedong Hwang

Publications 2

Preprint Open access

Mid-Training Language Models on Raw Video

Multimodal large language models learn mostly from paired image-text data or annotated video, and raw web video is rarely used to further train an existing language model. We study whether raw video, with no captions and no text loss, can serve as mid-training data for a pretrained language model. Frames are encoded in …

Preprint Open access

Sensor Geometry as a Flow-Matching Prior for Multi-Channel Brain Signals

Jaedong Hwang · 2026

Flow-matching models start from an isotropic Gaussian source, the standard choice when the correlation structure of the data is unknown in advance. For multi-channel brain recordings, however, part of this structure is known in advance. Electrodes sit at fixed positions on the head, and volume conduction through the sk …

Co-authors