Do Reasoning Representations Help Humans Evaluate LLM Outputs?
Reasoning representations are increasingly used as explanations for large language model outputs. Yet they are
- 用途
- 検出
- 難易度
- Hard
- コスト
- High
「detection」の検索結果
135 件Reasoning representations are increasingly used as explanations for large language model outputs. Yet they are
The local dark matter density determines the strength of the signal expected in direct-detection experiments,
Industrial fraud detection often relies on costly expert-crafted features that overlook graph-structured relat
When an external reference set (an anchor) is used to decompose an LLM-judge panel's error into a quality sign
Reinforcement Learning with Verifiable Rewards (RLVR) has been central to the recent success of Large Reasonin
Chagas disease is a major cause of cardiomyopathy in Latin America. Cardiac magnetic resonance (CMR) imaging c
Working entirely on topologically anonymized embeddings, we perform fraud detection using iterative rounds of
Test-time adaptation (TTA) has emerged as a prominent strategy for adapting vision-language models to distribu
Table detection is a core task in document analysis, supporting downstream applications such as information re
To mitigate the scalability bottleneck in the radio access network (RAN) in federated edge learning (FEEL), ov
We introduce HoneyRoute, an inference-serving layer that detects whether an incoming request is malicious and,
Token-level text anomaly detection, as an emerging trend of text anomaly detection, moves beyond coarse-graine
Monocular depth estimation is a ubiquitous yet highly ill-posed computer vision task, with downstream applicat
Open vocabulary 3D semantic segmentation methods typically lift CLIP features into 3D. This embeds points in a
Long-form instructional videos require automatic chaptering to support browsing, navigation, and knowledge acc
Onboard object detection in Earth observation is constrained by limited computational resources and the absenc
Egocentric 4D interaction forecasting aims to anticipate both where future interactions will occur in 3D and h
In persistent interactions, long contexts may encode an evolving process rather than a fixed record: later eve
Moving-object perception must decide which image regions correspond to real motion and keep every instance ide
Object detection and tracking are fundamental components of perception systems for autonomous driving. Achievi
Automating filament tracing in Cryo-Electron Microscopy (Cryo-EM) is essential for 3D helical reconstruction b
Hallucination detection is crucial for large language models (LLMs), as hallucinated content creates significa
Because dense frame-level annotation of colonoscopy videos is costly, we propose WSPolypNet, a weakly supervis
Speech deepfakes can mimic a speaker's voice convincingly enough to deceive listeners and automated systems. T
Large Language Models (LLMs) are increasingly deployed with hierarchical instructions, yet they remain vulnera
Large Language Models are increasingly deployed as information intermediaries, yet measuring their political b
Remote sensing multimodal large language models (RS-MLLMs) have advanced scene understanding and visual questi
Oncology care operates at constant pressure of absorbing rapidly evolving evidence base in biomedicine. The Am
Solving repository-level code tasks requires LLM-based agents to use code search tools to navigate large codeb
We describe the Snugi-AI-v2 submission to eRisk 2026 Task 2, the second edition of contextualized early depres
Autonomous 3D active mapping requires a space robot to choose where to sense while building the geometry neede
State-of-the-art vision models process images in their entirety, lacking the ability to selectively zoom in on
Spherical observations provide global visual context for 3D scene understanding. However, visual information i
Pixel-level annotation of fixed traffic-camera imagery is expensive, while crosswalk models trained from stree
Cancer segmentation models can fail silently, generating plausible but incorrect masks that risk missed findin
Reasoning segmentation converts an implicit linguistic conclusion into a precise mask, requiring both semantic
We present FIRE3D, a unified framework that takes a single RGB image or casual RGB video and transforms it int
Preoperative evaluation of trigeminal neuralgia (TN) often requires joint interpretation of structural MRI, wh
Table Structure Recognition (TSR) aims to extract the bounding boxes of cells and table structure (e.g., HTML)
A simultaneous localization and mapping (SLAM) method using a monocular camera and a low-cost inertial measure
Out-of-distribution (OOD) detection is crucial for safe deployment of medical AI systems, where domain shifts
Floorplans are compact, appearance-invariant maps ideal for indoor localization, yet existing methods rely on
Video Large Language Models (VideoLLMs) are increasingly deployed in safety-critical applications such as cont
Objective assessment of Freezing of Gait (FoG) in Parkinson's disease (PD) relies predominantly on wearable In
Multi-modal large language models (MLLMs) have demonstrated significant potential in image quality assessment
While multimodal large language models (MLLMs) achieve remarkable performance on generic image captioning, the
Object detectors have shown remarkable performance in various fields, among these medical imaging, surveillanc
Multi-object tracking (MOT) is an essential computer vision task that simultaneously tracks multiple objects i
Timestamp-supervised action segmentation aims to segment and classify actions in untrimmed videos with a rando
Low-rank tensor modeling has become an effective tool for hyperspectral anomaly detection. However, existing m
Tensegrity robots offer lightweight, compliant mobility over challenging terrain but remain difficult to model
Soft robots offer safe and adaptive interaction with humans and unstructured environments through their inhere
In multi-robot collaboration, task handovers rely on downstream verifiers performing remote attestation, which
Physics-Informed Neural Networks (PINNs) have recently emerged as a promising approach for solving Partial Dif
Deep learning is an efficient technique to monitor the real time welding process, reducing post-welding repair
Physics-informed neural networks (PINNs) struggle on PDEs whose governing physics varies across the domain. We
Recurrent detectors such as bidirectional long short-term memory (Bi-LSTM) networks are low-complexity alterna
AI agents sometimes act aligned when they infer they are being tested, and differently when not. We argue this
How can vision-language models help video anomaly detection (VAD) when surveillance data remain distributed, w
This study investigates the classification of individuals as healthy or at risk of Parkinson's disease using m
Smart healthcare monitoring systems require precise action recognition to ensure well-being and timely interve
LiDAR-based 3D Single Object Tracking (3D SOT) is critical for robotic perception and navigation and aims to l
Medical imaging artificial intelligence (AI) is commonly developed as separate mappings from radiographs to di
The growing realism and accessibility of manipulated and generated faces threaten the trustworthiness of digit
Even though backdoors in LLMs have been a growing concern, their inner workings are still under heavy scrutiny
LLM assistants often produce more answers than humans can review before users see them. Most evaluations ask w
Retrieval-augmented generation (RAG) enables large language models (LLMs) to answer questions by accessing ext
Voice phishing detection faces three critical challenges: real criminal recordings are unavailable due to priv
Large language models (LLMs) increasingly generate Markdown that is consumed by renderers, agents, code extrac
Reasoning over language instructions in embodied tasks such as robotics often requires understanding spatial r
Image restoration is commonly applied before object detection under adverse conditions, yet a visually improve
In this paper, we propose ReactVAU, a Slow-Fast Decoupled Framework for real-time streaming Video Anomaly Unde
Automated drone surveillance has become increasingly important for public safety, critical infrastructure prot
In real-world applications, pedestrian trajectory prediction models rely on inputs from detection and tracking
Bridging the simulation-to-reality gap in roadside LiDAR requires addressing several coupled discrepancies, in
We present our solution to the LUMPI track of the UCF UrbanTwin Sim2Real LiDAR Challenge at the 6th DriveX Wor
360° salient object detection (SOD) aims to accurately segment salient regions across a full field of view. Ho
Multi-object tracking (MOT) has advanced rapidly in urban surveillance and autonomous driving, yet many tracke
Infrared small target detection (ISTD) is an important research direction in image processing. However, existi
Standard Convolutional Neural Networks (CNNs) exhibit severe performance degradation due to a strong inductive
Post-hoc explanation methods are widely used to inspect image classifiers, but their reliability depends on de
Anticipating whether a person will interact from one's own perspective is a highly intuitive task for humans,
Modern Visual Place Recognition (VPR) methods excel on standard benchmarks yet remain brittle in feature-poor
We introduce DF26, a novel benchmark for detecting AI-generated videos containing fully synthetic clips produc
Climate change is increasing the severity and unpredictability of natural disasters. In time-critical crises s
Morphology of the sub-basal nerve plexus (SNP) reflects peripheral nerve health, and corneal confocal microsco
Segmentation of complex structures in X-ray tomographic data is a fundamental task in biomedical research, but
Vision-language models offer a promising approach for zero-shot anomaly detection (ZSAD). However, due to obje
Out-of-distribution (OOD) detection is critical for safe deployment of medical AI systems. Recently, test-time
Deep learning models for CT scan analysis are often limited by the scarcity of precise pixel-level annotations
Text spotting requires both accurate text recognition and precise spatial localization. Current specialised sp
Graphic designs, such as posters, advertisements, and infographics, are an important medium for communicating
Localizing a font into new languages is a highly intricate task requiring precise design adaptation of glyphs,
Accurate segmentation of pulmonary lesions is essential for effective clinical diagnosis and treatment strateg
Progress in video anomaly understanding (VAU) has long been limited by inherent deficiencies of real-world ano
Robot foundation models are trained and evaluated predominantly in English, and robot demonstration corpora do
Autonomous mobile robots performing person-following tasks often suffer from temporary occlusions and sensor t
Trust-Hub reuses host circuits: several files differ mainly in the inserted Trojan. When gates from sibling va
LLM agents operate in workflows where unsafe actions can have real consequences. Existing safety evaluations o
Authorship signals matter in settings where writing style carries identity: digital forensics, plagiarism anal
Randomized benchmarking can hide classical temporal correlations because its Clifford-twirled response is even
Vision-language-action (VLA) models adapted through supervised fine-tuning (SFT) inherit a structural asymmetr
Optimization-based simultaneous localization and mapping (SLAM) makes it possible to reduce accumulated naviga
We give a necessary and sufficient condition for the existence of power-one sequential tests in an i.i.d. comp
World-action models (WAM) predict future states to guide robot actions, enabling learning from both action-fre
Human-robot interaction (HRI) enables intuitive and intelligent collaboration between humans and robots in rea
Achieving robust SLAM in large-scale underground coal mines with complex structures and severe degeneracies re
The autonomous localization of fugitive gas emissions using small Unmanned Aircraft Systems (sUAS) constitutes
Mixture-of-experts (MoE) architectures increase model capacity by combining a collection of expert predictors
Cortical neurons fire sparsely -- often fewer than one spike per sensory window -- making rate coding insuffic
Understanding the composition of large-scale autonomous driving datasets is essential for safety, robustness,
Accurate identification of weld seam geometries is essential for automated robotic post processing operations
For robots to operate reliably in real-world environments, they need to perceive their surroundings, act, and
Reinforcement learning typically optimizes average reward. For generative policies, the average can hide an im
Dynamic networks are being applied in many domains, from social media to logistics systems, each with their ow
We model insider threat detection as a dynamic Bayesian game in which a platform coordinates a committee of st
Post-hoc out-of-distribution detectors are fitted on a finite reference set, so every score they produce is an
Automated repair of Hardware Description Language (HDL) designs remains challenging due to the large search sp
Symbolic Regression (SR) seeks to find succinct mathematical expressions that represent the fundamental relati
Machine learning systems are increasingly corrected while they run, and the decision of when to intervene is i
Global goodness-of-fit and discrepancy statistics can establish that a sample departs from a reference distrib
We propose a multivariate extension of the pseudo-Voigt profile-a weighted convex combination of Gaussian and
Point-cloud data routinely captured by modern imaging and sensor technologies provide detailed geometric descr
Bug localization is a labor-intensive task, particularly in large software systems. When abnormal behavior occ
Deep neural networks often exploit spurious associations, a failure known as shortcut learning. Auditing for s
An intuitive method for dimensionality reduction is proposed, which is highly effective for finding interestin
Predictive Coding (PC) is a neural learning paradigm that enables parallelizable neural network layer updates.
Current approaches to simulating biological neural circuits, whether on general-purpose hardware or dedicated
Streaming systems that maintain a pool of expert models must repeatedly decide whether to reuse an existing ex
Persistent acoustic monitoring can detect machine faults without physical contact, but always-on inference is
Sinkhole attacks in large-scale wireless sensor networks (WSNs) pose a serious threat to network functionality
We introduce a repeated dynamic incentive framework for characterizing when "compliance", or full-effort hones
Spiking Transformers provide a promising paradigm for efficient visual processing with spike-driven computatio
EEG foundation-model gains may depend on cohort, montage, or probe design. We evaluated five models on five ta
We present a regression-based approach to Arabic dialect geolocation that models dialectal variation as a cont