Training-Free Task Vectors for LLM Behavioral Control
Task vectors enable post-training model editing by identifying semantically meaningful directions in weight sp
- 用途
- 技術検証・論文読解補助
- 難易度
- Hard
- コスト
- High
「Fine-tuning」の検索結果
82 件Task vectors enable post-training model editing by identifying semantically meaningful directions in weight sp
Dense self-attention treats all token pairs as equally plausible before learning, an interaction-isotropic pri
We apply transfer learning (TL) to construct data-driven models of inclusive electron-nucleus cross sections.
Large Language Models demonstrate remarkable proficiency in static reasoning, yet training them as autonomous
We present Miles v0.1, a full-stack, production-ready system for frontier post-training. Building upon the cle
Several geometry-aware approaches to low-rank adaptation have emerged for parameter-efficient fine-tuning of l
Table detection is a core task in document analysis, supporting downstream applications such as information re
Monocular depth estimation is a ubiquitous yet highly ill-posed computer vision task, with downstream applicat
Large Language Models (LLMs) are increasingly deployed in settings where rare but severe harmful generations c
Agent harnesses (the system prompt, tool set, execution hooks, and context-management scaffolding around a mod
Language-model checkpoints are commonly selected by pretraining loss or benchmark scores, assuming that the hi
Ensuring the safety of autonomous driving is a critical challenge. Scenario-based testing is a systematic proc
As Large Language Models (LLMs) are increasingly integrated into human society, aligning them with pluralistic
Training capable cyber agents is often treated primarily as a problem of model scale, yet open-weight post-tra
Automating filament tracing in Cryo-Electron Microscopy (Cryo-EM) is essential for 3D helical reconstruction b
Learning direct current circuit concepts requires learners to connect invisible physical quantities, such as c
Vision-language models (VLMs) augmented with retrieval-augmented generation (RAG) benefit from access to exter
We introduce OntologyBench, a tiered biomedical retrieval benchmark comprising 471,854 training and 125,744 ev
Mixture-of-Experts (MoE) pretraining relies on an auxiliary load-balancing loss (LBL) to drive per-expert util
Preference-based fine-tuning methods such as RLHF and DPO require substantial compute and large preference dat
As LLMs are increasingly used for pre-submission self-review, there is growing demand for feedback that not on
Social interaction is central to children's language learning, but the effects of different forms of caregiver
Recursive self-improvement (RSI) requires a concrete mechanism through which an AI system observes its capabil
Updating a language model's knowledge through fine-tuning is essential for keeping its outputs current, yet ca
While 3D Gaussian Splatting (3DGS) has emerged as a powerful representation for real-time novel view synthesis
Despite recent advances in surgical vision-language models (VLMs), temporal reasoning remains limited because
Micro-actions are subtle, low-intensity non-verbal behaviors that provide cues to fine-grained human states, i
Reconstructing a three-dimensional left-ventricular (LV) endocardial surface from cardiac magnetic resonance (
Generalist robot policies carry broad manipulation priors from large-scale data, but specializing them to a ne
Long-horizon target navigation requires a robot to sustain task execution across evolving observations, decisi
Acidophilic proteins that remain stable and functional under highly acidic conditions, are important for indus
Localized dimensionality reduction improves the scalability of operator learning for high-dimensional partial
Semantic and goal-oriented communication is increasingly studied for 6G, but generalization beyond seen data r
Automatic modulation classification (AMC) models are frequently trained and validated on synthetic or channel-
Full-parameter fine-tuning of large language models has substantial memory costs because backpropagation store
Tool-augmented language agents are vulnerable to indirect prompt injection (IPI). Unlike direct prompt injecti
Large language models (LLMs) have demonstrated strong performance in table understanding. However, they typica
Large language models (LLMs) have recently shown promise for historical entity linking, but preference optimiz
Large language models (LLMs) are often post-trained on pre-collected reasoning trajectories to improve their r
Voice phishing detection faces three critical challenges: real criminal recordings are unavailable due to priv
Automated drone surveillance has become increasingly important for public safety, critical infrastructure prot
Accurate 3D plant organ segmentation is fundamental to automated phenotyping. Existing approaches rely on anno
Agentic systems can interpret user requests, search the live web, and use external tools, but their ability to
While Reinforcement Learning (RL) effectively incentivizes reasoning in Large Language Models, current pipelin
Single-view 3D reconstruction, also known as image-to-3D, is a persistently challenging task due to the extrem
Text spotting requires both accurate text recognition and precise spatial localization. Current specialised sp
Compared with relying solely on initial observations and language instructions, predicting goal images with ge
High-quality demonstration data is becoming a central bottleneck for training general-purpose humanoid robots.
Contact-rich assembly remains challenging because it requires submillimeter spatial accuracy and reliable inte
Visual-Inertial (VI) fusion is fundamental to accurate and robust state estimation, where camera and IMU measu
Training strategy, namely whether to retrain from scratch or fine-tune from the previous checkpoint, is an ove
Temporal relation extraction determines whether an event occurs before, after, or simultaneously with another
The In-context learning (ICL) paradigm aids large language models (LLMs) to adapt to new tasks without need fo
Although professional workflows leverage large language models widely, the interpretation for auditing unconst
Existing approaches to visual attribute value extraction (AVE) primarily rely on static product images, failin
Although highly effective in vision and language domains, applying in-context learning to robotics remains cha
Behavior-cloned visuomotor policies can remain accurate near their training distribution yet fail when object
We present LANTERN, a closed-loop benchmark for temporally grounded cooperative warnings. LANTERN separates wa
Existing unlearning approaches typically rely on post hoc weight adaptation or distillation, leading to duplic
Vision-language-action (VLA) models adapted through supervised fine-tuning (SFT) inherit a structural asymmetr
Humanoid loco-manipulation requires accurate whole-body motion tracking in the world frame for physical intera
Human demonstrations contain rich manipulation knowledge, but it remains unclear what information can be trans
Can interactive vision-and-language agents learn not just what to say but also \textbf{\textit{when}} to say i
World-action models (WAM) predict future states to guide robot actions, enabling learning from both action-fre
Vision-language-action (VLA) models can execute short manipulation skills, but remain brittle in long-horizon
The autonomous localization of fugitive gas emissions using small Unmanned Aircraft Systems (sUAS) constitutes
Autonomous underwater robots are widely used for exploration, monitoring, and inspection, where safe navigatio
Causal inference is the practice of estimating the effect of a treatment or intervention from data. It traditi
Learned graph simulators provide an efficient alternative to high-fidelity solvers for granular dynamics. Howe
The Rashomon effect is a machine learning phenomenon where equally accurate models produce different predictio
Emerging edge AI workloads increasingly require arithmetic units that can trade computational accuracy for eff
Industrial recommenders give new content initial views through budgeted exploration, then use early performanc
Automated bidding (autobidding) is a core component of modern online advertising systems. Within this componen
Searchless chess networks reach human master strength from a single forward pass by imitating a stronger teach
Discrete diffusion models have become a strong, widely adopted class of generators for sequence data, and stee
Automatic Speech Recognition (ASR) technologies have achieved remarkable performance in recent years through t
Population optimizers such as CMA-ES, DE, and multi-objective evolutionary algorithms drive search mainly thro
Zeroth-order (ZO) optimization estimates gradients using only forward-pass evaluations, making it suitable for
Gradient injection helps Particle Swarm Optimization (PSO) only when the swarm has identified a basin with smo
LLMが外部アクションを取り続けている場合、エージェントのメモリーが古くなったり誤った情報を持ったりする可能性があります。この問題を解決するために、この研究ではSafeCommitという技術を提案しています。SafeCo
この研究では、記憶を維持し更新する能力を強化するため、繰り返し神経ネットワークを用いて記憶の特性を分析しました。記憶を維持するための神経計算の一種として、ダイショビュールノーマリゼーション(Divisive Normal
ディナミカルシステムを化学的なパラフレーズで説明する手法を提案し、システムの