Agent as Policy for Robotic Manipulation
We demonstrate that a general-purpose agent can directly drive a physical robot throughout task execution with
- 用途
- 技術検証・論文読解補助
- 難易度
- Easy
- コスト
- High
「supervised」の検索結果
15 件We demonstrate that a general-purpose agent can directly drive a physical robot throughout task execution with
Reinforcement learning with verifiable rewards is typically performed on-policy, keeping training data close t
We introduce StepAudio 3 Gen, a general-purpose audio generation model that supports zero-shot text-to-speech
We launch SenseNova-U1.5, an 8B-MoT native unified multimodal model that understands, reasons about, and gener
Iterative reference-conditioned image editing can introduce grid-like and granular textures, commonly describe
Diffusion large language models (dLLMs) achieve high decoding efficiency through block-parallel, arbitrary-ord
We study how model post-training and test-time inference design affect natural-language proof generation for h
Training capable cyber agents is often treated primarily as a problem of model scale, yet open-weight post-tra
Image tokenizers define the ``visual language'' of unified multimodal models, yet are commonly studied through
As LLMs are increasingly used for pre-submission self-review, there is growing demand for feedback that not on
SWE-Bench Pro has emerged as a standard benchmark for evaluating software engineering agents on challenging re
Invasive coronary angiography (CAG) is the gold standard for diagnosing coronary artery disease, but interpret
Large language models (LLMs) are often post-trained on pre-collected reasoning trajectories to improve their r
Recursive Super-Resolution (SR) extends fixed-scale SR to extreme magnification by repeatedly feeding predicti
Grounding natural-language instructions into reliable and executable actions remains a fundamental challenge f