label-studio — Label Studio is a multi-type data labeling and annotation tool with standardized output format
データラベル化と注釈化を行うためのツールです。
- 用途
- データラベル化ツール
- 難易度
- Easy
- コスト
- Low
「detection」の検索結果
36 件データラベル化と注釈化を行うためのツールです。
ultralyticsはYOLO(You Only Look Once)の技術を使用したオブジェクト検出ライブラリで、高い精度を提供している。
supervisionは、機械学習技術を活用して、ユーザー独自のコンピュータビジョンツールを作成することができる。
CVATは、機械学習用の業界標準のデータエンジンです。さまざまなスケールのチームが使用し、さまざまなスケールのデータに対応しています。
FiftyOneは、データセットの精査とAIモデル可視化を支援するライブラリです。このライブラリは、データセットの品質を高め、AIモデルを可視化するのを支援するために使用できます。
このプロジェクトは2Dおよび3D顔の分析を実現するための基盤プロジェクトであり、最先端の技術を導入して顔の分析を実現します。
電気生理信号から表現を学習し、脳コンピューターインターフェースの開発を支援する。
presidioは、テキスト、画像、構造化データを含む敏感データを検出、削除、マスク、アノニマイズするオープンソースフレームワークです。自然言語処理、パターンマッチング、カスタマイズ可能なパイプラインをサポートします。
Test-time adaptation (TTA) has emerged as a prominent strategy for adapting vision-language models to distribu
Table detection is a core task in document analysis, supporting downstream applications such as information re
Token-level text anomaly detection, as an emerging trend of text anomaly detection, moves beyond coarse-graine
Automating filament tracing in Cryo-Electron Microscopy (Cryo-EM) is essential for 3D helical reconstruction b
Remote sensing multimodal large language models (RS-MLLMs) have advanced scene understanding and visual questi
We describe the Snugi-AI-v2 submission to eRisk 2026 Task 2, the second edition of contextualized early depres
Preoperative evaluation of trigeminal neuralgia (TN) often requires joint interpretation of structural MRI, wh
While multimodal large language models (MLLMs) achieve remarkable performance on generic image captioning, the
LiDAR-based 3D Single Object Tracking (3D SOT) is critical for robotic perception and navigation and aims to l
Retrieval-augmented generation (RAG) enables large language models (LLMs) to answer questions by accessing ext
Multi-object tracking (MOT) has advanced rapidly in urban surveillance and autonomous driving, yet many tracke
Infrared small target detection (ISTD) is an important research direction in image processing. However, existi
Out-of-distribution (OOD) detection is critical for safe deployment of medical AI systems. Recently, test-time
Text spotting requires both accurate text recognition and precise spatial localization. Current specialised sp
Authorship signals matter in settings where writing style carries identity: digital forensics, plagiarism anal
We present Cadence, an error-bounded lossy compressor for numeric time series pairing a 330M-parameter time-se
Agents can turn shared infrastructure into a channel for coordinated intrusion. The Hugging Face incident and
We present 'RenderFormer-V2', a unified learned transformer-based neural rendering model, complementary to mod
Understanding the composition of large-scale autonomous driving datasets is essential for safety, robustness,
The ambition of the 2025 PNPL competition (Landau et al., 2025) was to launch a multi-year curriculum for non-
Personalized assistants should not only comply with user requests but also assess whether those requests are a
We present FlashRender, a few-step generative rendering framework that retakes a source video along a target c
Visual fluency in generated video does not imply physical reliability, and a scalar quality score alone is inc
Ensuring factuality remains a critical challenge for deploying LLMs in high-stakes settings. Existing hallucin
YOLOv5という物体検出アルゴリズムをPyTorchから他の言語に変換できるライブラリ。
Reinforcement learning with verifiable rewards (RLVR) substantially improves single-sample accuracy (pass@1) b
Change data synthesis provides a cost-effective solution for expanding training data and improving the perform
このライブラリは、コンピューター ビジョンのための高度なAI解釈と可視化ソリューションです。このライブラリは、CNN、ビジョン トランスフォーム、分類、物体検出、分割、画像類似度など、さまざまなコンピューター ビジョンの