Ambient @ EgoLongQA 2026: Distilling Long-Video perception into a Sub-2B Model
We describe our entry to the EgoLongQA track of the Wearable-AI Challenge in ECCV 2026, which placed first in
- 用途
- 技術検証・論文読解補助
- 難易度
- Easy
- コスト
- High
「qa」の検索結果
9 件We describe our entry to the EgoLongQA track of the Wearable-AI Challenge in ECCV 2026, which placed first in
Reacting to sudden physical hazards (catching a slipping plate, dodging a falling knife) is both a meaningful
Video understanding has rapidly evolved toward video large language models (VideoLLMs): systems that couple vi
The KV cache is a primary bottleneck for Transformer decoding: its memory footprint and cache-read traffic gro
Large Vision-Language Models (LVLMs) have achieved strong performance on diverse visual tasks, yet their abili
Sequential memory agents process long documents by reading chunks one after another while maintaining a compac
Recursive Super-Resolution (SR) extends fixed-scale SR to extreme magnification by repeatedly feeding predicti
Data policies for reinforcement learning with verifiable rewards (RLVR) determine which rollouts are used, how
Recent advances in wearable sensing enable continuous monitoring of physiological and behavioral signals, yet