Beyond Sufficiency: Time Series Explanation with Counterfactual Necessity
時系列データ分析技術のための新しいアプローチであるTimePNS(Time Series Explanation with Counterfactual Necessity)を提案しました。TimePNSは、時系列データ
- 用途
- 時系列データ分析技術の開発
- 難易度
- Hard
- コスト
- Low
「classification」の検索結果
141 件時系列データ分析技術のための新しいアプローチであるTimePNS(Time Series Explanation with Counterfactual Necessity)を提案しました。TimePNSは、時系列データ
Cognitive impairment (CI) is a growing public health concern. Early and accurate diagnosis is critical for ena
The original ALPHA benchmark introduced a taxonomy-aware penalty for evaluating CWE-level vulnerability predic
Automated detection of vision impairing retina-based ocular conditions from fundus images is important for ear
Ordinal Classification (OC) deals with classification tasks where the classes follow a natural order. Despite
ダブル量子ドットのCharge Stateを分析するためのMachine Learning方法を提案した研究で、この方法により、量子ドットのCharge Stateが効率的に分析できる。
Artificial Epanorthosisは、大規模言語モデルが古典的なルレチックの表現を使用する傾向に注目した。結果は、モデルのトレーニングデータの形状がこの傾向に影響していることができた。
The rise of human-AI collaborative writing has created a growing need for fine-grained detection methods that
Instance-level explanations aim to reveal the rationale behind a model's decisions for a specific graph. Previ
Phonetic forced alignment is a key technique in phonetic research, yet existing alignment systems lack special
この研究では、大規模言語モデルを使用して、basketボールの動的理解に基づいて、プレイヤーへの関わりや時間境界を推測するモデルを開発しました。
表現的推論問題とは、与えられた現象に対して説明が見つかることを目的とした問題です。この研究では、解決方法を最適化するために、代表集合 (representative sets) の概念を活用します。
知識重視の質問応答システム (KI-VQA) を分析するために、新しい評価基準を提案します。これらの基準では、VLMの各タスクを個別に評価することができます。
Multimodal large language models (MLLMs) have achieved impressive performance in multimodal emotion recognitio
Zero-shot summarization using Large Language Models (LLMs) has significantly advanced the abstractive summariz
Large vision-language models are becoming increasingly dominant in 3D medical image interpretation, but we rar
Body-based emotion recognition is important for real-time affective systems, but graph-based skeleton models c
The electrocardiogram (ECG) is a cornerstone of cardiac as- sessment, yet clinical deployment of deep learning
この論文では、DONDO と呼ばれるアフリカ諸国向けの音声認識ベースモデル (ASR)が構築されました。これらのモデルは、自律学習型スピーチエンコーダーであるw2v-BERT 2.0を使用して構築されています。このエンコ
We present VibeVoice-ASR-BitNet, a compressed variant of VibeVoice-ASR optimized for real-time inference on ed
アイス認識の精度を向上させるための方法を提案し、視覚認識におけるアイス認識タスクの課題を分析した。
Quantitative drug-induced sleep endoscopy (DISE) requires reliable airway boundaries at specific anatomical le
Conventional face recognition relies on static appearance cues and degrades in unconstrained settings with exp
大規模言語モデルはさまざまな画像をテキストに変換する上で優れた性能を示しているが、発生するホログラフィックな診断にはまだ解決策が必要です。この研究では、主流の粗い検出方法の欠点を補うため、細部の診断方法を提案しています。
We present HyperImageNet, a large-scale benchmark for fine-grained hyperspectral land-cover understanding. The
Deepfake detection is moving beyond binary classification decisions toward systems that can also explain the v
Recently, cross-domain few-shot facial expression recognition (CF-FER) has received considerable attention. Ho
この研究では、都市ウォークビデオを分析するために、4つのモダリティの表現(スペース時領域情報、時間平均画像、オーディオ符号化、テキストベースの表現)を使用しました。
この論文では、finger vein画像から年齢と性別を推測するためのMulti-InstanceAge and Gender Estimation(MAGE-Vein)モデルを提案します。
この論文では、Webly Supervised Multi-Label Recognition(WS-MLR)という手法を提案します。WS-MLRは、web画像データセットを使用して、多ラベルを解釈します。
Stress is a dynamic process characterized by significant individual variability in facial expression. Traditio
Scalar metrics are often used to evaluate clusterings against known classes, but they can obscure a fundamenta
For two decades, the standard remedy for class-imbalanced learning has been to fabricate synthetic minority ex
Multimodal learning is a robust approach to improve predictive performance in applications such as medical pro
Machine learning (ML) models are increasingly integrated into optical network automation frameworks to support
The robustness of machine learning techniques across heterogeneous network domains remains an open challenge i
Unique and rapid classification of knots and links is an open mathematical problem that is relevant to a range
Federated learning (FL) enables multiple clinical institutions to collaboratively train a shared disease class
この研究では、IBM Quantum ハードウェア上で実行される4キュビットの量子カーネルについて、スルーハードウェアで状態ベクトルを使用して、グラム行列における幾何学的情報の実行時の正確性を検証した。
この研究では、ナノポア測定器から得られる複雑な信号を分析するために、多モーダル変換ニューラルネットワーク (Multi-modal Transformer) を提案し、信号分類の精度を向上させた。
Medical image encoders from different groups are increasingly treated as interchangeable, on the assumption th
Neural network misclassifications exhibit characteristic spectral instability in internal activations that is
ユーザー センター化されたトランザクション シーケンスのモデリングを目指す新しいアプローチを提案し、Contrastive Representation Learning と State Space Models の組み
生分子注文を扱う研究、Plausibility-Driven Prioritization を用いて生分子注文を提案する。
エッジコンピューティング用ニューロモーフィッククラッサを提案する。
複数プロジェクト間の欠陥予測を扱う研究、Multi-stage Dynamic Selection を用いて複数プロジェクト間の欠陥予測を提案する。
流動画像生成を扱う研究、HeadCast を用いて流動画像生成を提案する。
協同学習を扱う研究、Autonomous Collaborative Learning を用いて協同学習を提案する。
この研究では、抗原特異性抗体を設計するために、抗原および抗体の間でエピトープレベルでのペアリングが必要であることを考慮した、抗原特異性の抗体多モーダルファンデーションモデル(AAMFM)を提案しました。
Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting t
この研究では、深層人工神経ネットワーク(DNN)の検証コストを削減するために、ニューラル崩壊 instabilitiyを用いたテストケース優先順位付け方法を提案しました。
この研究では、Androidマルウェアの検出に使用されるデープラーニングモデルをOptimizeする方法を提案しました。
この研究では、人間の皮膚から放出される光を検出し、その測定値から健康状態の予測を可能にする手法を提案しました。
この研究では、人工知能の研究者と神経科学者の間の分野を結びつけるために、脳のシステム構造を研究し、その研究から導かれた新しいアプローチを提案しました。
Adversarial robustness is commonly evaluated with predefined attack ensembles, such as AutoAttack, at a single
量子合成関数を使用した分類問題を解決するための方法を提案している。この方法は、学習可能な量子関数を使用し、訓練データサイズの線形スケーリングを実現している。
Classifier-free guidance (CFG) is the default mechanism for conditional generation in diffusion models, but th
この研究では、オンライン予測の再調整アルゴリズムを提案します。このアルゴリズムは、適切な損失に対して誤差が小さい新しい予測を生成でき、過去の予測に比べて性能が向上した結果を確保することができます。
This paper examines the question of whether artificial intelligence (AI) systems can be creative, approached f
Concept Bottleneck Models provide interpretable-by-design predictions by mediating diagnosis through human-und
Optical Character Recognition (OCR) for Persian remains substantially less mature than for Latin-script langua
LLMは、状況から価値観を判断できるかどうか、という研究が調査されました。LLMは、状況に応じて真の価値観を推測することができました。
大きな言語モデルは、ユーザーの信念と事実的な正しさを合わせる傾向があるが、これらの傾向は多様であることを明らかにしました。
職業コード付けは、職業タイトルから職業分類を識別することであり、二つのステップで実行される二つのアプローチのうちのどちらかが最も効果的であることを示しました。
This paper presents the second edition of the TalentCLEF Challenge, which will run as an evaluation lab as par
Detecting media bias automatically is difficult because biased framing is often subtle, yet in domains such as
人名と場所の関係を抽出するタスクは、歴史的ニュース記事の解釈において重要です。従来の方法では、言語モデルの前処理が必要でしたが、Lightweightアルゴリズムは、依存グラフと近接特性を使って、歴史的ニュース記事から人
Virtual reality (VR) headsets (e.g., Meta Quest, Apple Vision Pro) provide a seamless user experience due to t
Housing-level urban physical examination is essential for identifying residential building problems and suppor
OffNadirLocは地学化におけるオフナジアムの視点を考慮するための基準セットを提案します。これにより、ドローンと衛星画像の交差視点地学化プロセスでは重要な構造的シーン理解と内部ドメイン間の関係的制約に重点を置くこと
Understanding instrument-tissue interactions is essential for context-aware surgical AI and autonomous robotic
Thermal-to-visible face translation presents fundamental challenges including geometric discontinuities, seman
この研究では、地象性AIにおける物理的知識を使用してポーラリメトリック合成開口ラダール画像を分類するための新しいモデルを提案しました。このモデルは、ラダール画像を物理的なプロセスと関連付けることができます。
3D Gaussian Splatting (3DGS) は 3D セグメント間の接着を実行するために使用され、テキスト ドライブ の 3D シーン エディットには不可欠です。現行の方法では、固定位置撮影から 2D ディ
Pain is a complex and pervasive phenomenon affecting a large percentage of the population, and accurate assess
Noisy and corrupted points can substantially degrade point cloud recognition performance, especially under cha
分類タスクの性能を最適化するために、分布化されたクラスター間の分離を考慮したアルゴリズムが提案されていました。
A walking fly steers toward a goal direction, held as a bump of activity across the FC2 neurons of the fan-sha
Instruction tuning is meant to make language models follow user requests, yet it is unclear whether small mode
人間が与えるラベル相違を研究した研究では、主に分異議のあるデータを選んで分析したところ、非上昇漸近性オペレータを持つ仮説がChaosNLIで低いラベル相互に関連性があると結論づけた。しかしながら、データが選択されていない
Clinical NLP evaluation remains dominated by multiple-choice question answering (MCQA), which scores only fina
この研究では、エージェントメモリーのワークロードは直接的事実検索、関係連鎖や現在の状態の推論、長時間の履歴上に関係がある合成を組み合わせて、Supra Cognitive Modes を開発しました。このアーキテクチャで
人間の耳は最高の聴覚能力をもつものであると考えられており、音声認識では人間の聴覚機能を上回るようなシステムが作りだされるのを待っている。しかし、このようなシステムは実現しておらず、人間は音声認識システムの基準作成の参考と
Multimodal humor in memes, cartoons, and comics remains difficult for AI systems because intended meaning depe
We present AutoJourn, a demonstration system for multi-perspective news generation and bias-aware evaluation u
アラビア語の発音記号化は重要な問題だが、データが不足していることが難点の一つである。この問題を解決するために、ここでは「Connectionist Temporal Classification (CTC)」を使った制約
Automatic speech recognition (ASR) for African languages is constrained by orthographic inconsistency, annotat
Large Language Models (LLMs) are increasingly fine-tuned for critical-domain Question-Answering (QA), yet choo
The allocation of visual attention by pathologists during cancer diagnosis is a highly selective process that
記述情報に従って画像や動画データを混ぜ合わせる「対数混合法」を拡張する方法、InstructMixupを提案する。これにより、データを拡張しながらデータの内容とラベルが維持される。
計算機ビジョンと画像認識では、画像の視覚的なリッチネスを評価するために有用な指標が求められるが、これまでの指標は制限があった。この問題を解決するために、チャンネル空間の分散を利用した指標を提案する。
With growing surgeon shortages, automating surgical sub-tasks such as tissue dissection offers a promising ste
行動モーター特徴は、社会認知や人間ロボットインターフェースなどの行動認識の核心です。人間ロボットのNICO用に、2段階のアーキテクチャを提案します。1段階目では、腕の移動を学習するSOMと、手の移動を学習するSOMを使用
Robot-assisted minimally invasive surgery (RMIS) offers major benefits over open and conventional laparoscopic
Combinatorial markets provide a general framework for trading bundles of indivisible goods. Building on the co
Language models produce probabilities over words, but professional decisions require uncertainty over meaningf
Dynamic graph learning aims to capture evolving structural and semantic patterns in real-world systems, such a
Purpose: Understanding how much of routine policing involves vulnerable people could inform resourcing, traini
Automated analyses of privacy policies enable large-scale assessments of transparency in digital ecosystems, y
Pruning long context for coding agents has been a vital technology for efficient context management. While exi
エッジロボチクスでの画像認識精度を安定させ、その安定性を確保するために、量化後のパフォーマンスを向上させ、分散型データ量化を実現し、分布シフトの影響を緩和する、新しい機械学習アプローチを提案します。
共役ロボットは、人間オペレータと同梱するワークスペースを共有し、機械手のハンドオーバーなどの安全性の高いマイクロイベント頻繁に発生します。但し、従来の静的なハンドオーバーは、非対称の産業工具を取り扱う際、不自然な抓を持つ
採掘ロボットの性能向上を目指したSeg2Graspを構築し、セグメンテーション、グレイシング、クラスフィルタリングの3つのモジュールで構成されます。セグメンテーションモジュールではTransformerを利用したオブジェ
Humanoid robots have become increasingly popular in applications such as social interaction, education, and se
How deep does a graph neural network need to be on a sparse graph? We study its purest statistical form: node
LiDAR place recognition supports loop closure, relocalization, and multi-agent map management. As robotic plat
Modern safety-critical systems increasingly rely on human-robot interaction to reduce disaster risk and suppor
脳のニューロン同士のつながりを分析する方法を提案する。この方法では、神経伝達の構造を考慮しながら、ニューロン間のつながりを分析できる。
noisyラベルを扱うための学習アプローチを提案し、それをテストした。
distillationにおける予測のみを扱う学習アプローチを提案し、それをテストした。
高次元カテゴリデータを可視化するため、hierarchical optimizing linear assignment (HOMALS)を使用し、可視化に役立つ関連表
この研究では、共有ニューロンを使用して時系列グラフを学習する方法、NeuronSoup を開発しました。NeuronSoup では、各パスの信号は、変数数の間のニューロンを通過する途中で、共有ニューロンを使用して伝票され
この研究では、ウェアラブルデバイスで電気生理学記録(ECG)を分析するために使用される深層学習アルゴリズムを開発することを目的としています。このアルゴリズムは、エネルギー効率が高く、小型化が可能であるため、心臓の病気の検
Nonlinear thermodynamic computers based on Langevin dynamics exploit thermal fluctuations as a physical substr
Boosting is one of the most successful learning techniques for standard classification and regression tasks. I
Interfacing with Biological Neural Networks (BNNs) requires encoding information into stimulation patterns tha
Spiking Neural Networks (SNNs) trained through unsupervised Spike-Timing-Dependent Plasticity (STDP) have been
We introduce Sticky Jump Diffusions (SJDs), continuous-time Markov processes on $\mathbb R^d$ whose discrete a
Which discrete symmetry groups can arise from strategic interaction? We tile the plane with copies of a bimatr
spiking neural networkの設計とシミュレーションを行うためのフレームワークを提案する
この研究では、アルツハイマー病の前期診断と生物学的マーカーの検出にAI技術を適用します。AIモデルをトレーニングするために、電気エイセフィログラム(EEG)データを使用し、精度を高めます。また、AIモデルが得た情報を分析
Strategic value can fall when an option becomes visible. A route, signal, bet, or opportunity may be attractiv
マルチアジェント最適化を使用して、クラスター選択とモデル調整のためのMMAOクラスの実現を提案しました。
Lightweight neuromorphic computing offers a promising route to efficient AI, with particular benefits for reso
Biological neural circuits obey Dale's principle: each neuron's synapses are uniformly excitatory or inhibitor
Humans facing algorithmic decision systems have been found to ``game'' them by altering their input data (at a
Continual learning (CL), where a model is trained on a sequence of data tasks, is increasingly being adopted a
Modern machine learning applications employ deep neural networks training with the error backpropagation algor
Algorithmic developments in Strategic Classification have been mostly limited to linear classifiers in setting
Self-organized criticality (SOC), a dynamical regime associated with maximal information processing, offers a
Traditional evaluation of machine learning (ML) models typically focuses on achieving the maximum possible acc
Continual learning that is gradient-free, local, online, and append-only is attractive for edge and streaming
機械学習モデルを評価する手法を提案。既存の評価方法ではモデルが誤った結果を出してしまうため、これによりモデルが正確に評価できる。
神経網路の設計を目指す本研究では、ANNとSNNを組み合わせたハフマン式設計法
この論文では、スパイクニューロンの生物学的合理性を評価するためのオプティマイズフレームワークを提案します。このフレームワークは、Izhikevichの生物学的合理性の定義に基づいており、スパイクニューロンをモデル化するた
Efficient processing of continuous audio streams remains a key challenge for real-time and resource-constraine
この研究では、Controlled Dynamics Attractor Transformer (CDAT)を提案しました。このTransformerは、Self-Attention MechanismとAssocia
Learning in biological multilayer neuronal networks offers insights that extend beyond the classical weighted-
Evolutionary optimization of spiking neural networks (SNNs) becomes increasingly difficult as task complexity
スパイクニューラルネットワークは、エネルギー効率のよいAIモデルです。この研究では、スパイクニューラルネットワークのアクセラレータを実装し、その性能をテストしました。
Safe rehabilitation is an interaction-dynamics problem: the controller must regulate a prescribed motion while