Benchmark Radar: A Living Database and Search Engine for AI Benchmarks and Evaluation
Benchmark researchers and developers of large language models (LLMs) and other AI systems need to find relevan
- 用途
- 技術検証・論文読解補助
- 難易度
- Easy
- コスト
- High
「retrieval」の検索結果
9 件Benchmark researchers and developers of large language models (LLMs) and other AI systems need to find relevan
Objects in post-fire environments often undergo irreversible physical transformations that change their geomet
Late-interaction retrieval is the state-of-the-art for visual document search, but it pays for its accuracy in
Multimodal entity linking grounds entity mentions in text and images to knowledge-base entries. These systems
Video understanding has rapidly evolved toward video large language models (VideoLLMs): systems that couple vi
Text-to-motion (T2M) generation maps natural language to human joint movements, aiding gaming, VR, and robotic
Ensuring the safety of autonomous driving is a critical challenge. Scenario-based testing is a systematic proc
High-quality structured organic reaction data are essential for developing artificial intelligence for chemist
Large language model (LLM) agents increasingly rely on external skills, but routing user requests over large s