Skip to main content

Research timeline

Related research and updates

Public articles linked to the same research event.

arXiv

ECHO-k uses a foundation model's internal representations as self-supervised proxy targets to select modalities sequentially under budget and improve downstream performance when the task is unknown

The work introduces ECHO-k, a task-agnostic and self-supervised principle for modality acquisition that uses a deep model's internal pretrained representations (e.g., from a foundation model) as proxy targets summarizing cross-modal information, provides theoretical guarantees in a stylized linear setting that motivate a reinforcement learning policy for sequential modality selection, and, across task-agnostic and label-free acquisition baselines, consistently improves budgeted downstream performance across diverse foundation-model backends.