LLM Fine-tuning の論文・実装まとめ

wandb — The AI developer platform. Use Weights & Biases to train and fine-tune models, and manage models from experimentation to production.

Weights & Biasesは、AI開発を支援するプラットフォームです。このプラットフォームは、モデル開発から生産準備までを支援し、コストをコントロールし、モデルとデータへのアクセスを管理します。

Score 99

great_expectations — Always know what to expect from your data.

データの期待値を把握するためのフレームワークです。

Score 99

peft — 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

パラメータ効率の向上のための最先端のフィネチュニングフレームワークです。

深層学習Transformer

Score 94

prompts.chat — f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.

prompts.chatは、コミュニティが共有したChatGPT用のプロンプットを発見・収集できる場所で、無料でオープンソースで提供されている。

Score 94

ray — Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

rayは、core分布ランタイムとAIライブラリで構成されたAI計算エンジンで、スケーラブルなAI計算をサポートする。

Score 94

rig — ⚙️🦀 Build modular and scalable LLM Applications in Rust

Rustを使ってモジュラーLLMアプリケーションを構築することができるライブラリです。

自然言語処理大規模言語モデル生成

新着記事

wandb — The AI developer platform. Use Weights & Biases to train and fine-tune models, and manage models from experimentation to production.

用途: AI開発プラットフォーム
難易度: Easy
コスト: Medium

great_expectations — Always know what to expect from your data.

データの期待値を把握するためのフレームワークです。

用途: データの期待値を把握する
難易度: Easy
コスト: Medium

peft — 🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.

パラメータ効率の向上のための最先端のフィネチュニングフレームワークです。

深層学習Transformer

用途: パラメータ効率の向上
難易度: Easy
コスト: Medium

prompts.chat — f.k.a. Awesome ChatGPT Prompts. Share, discover, and collect prompts from the community. Free and open source — self-host for your organization with complete privacy.

prompts.chatは、コミュニティが共有したChatGPT用のプロンプットを発見・収集できる場所で、無料でオープンソースで提供されている。

用途: チャットGPT用のプロンプトを共有
難易度: Easy
コスト: High

ray — Ray is an AI compute engine. Ray consists of a core distributed runtime and a set of AI Libraries for accelerating ML workloads.

rayは、core分布ランタイムとAIライブラリで構成されたAI計算エンジンで、スケーラブルなAI計算をサポートする。

用途: AI計算
難易度: Easy
コスト: High

rig — ⚙️🦀 Build modular and scalable LLM Applications in Rust

Rustを使ってモジュラーLLMアプリケーションを構築することができるライブラリです。

自然言語処理大規模言語モデル生成

用途: モジュラーLLMアプリケーション作成
難易度: Easy
コスト: High

自然言語処理大規模言語モデルテキスト音声マルチモーダル

screenpipe — YC (S26) | Record your screen 24/7 and plug into your agents. Local, private, secure. Connect to OpenClaw, Hermes agent and 100+ apps

ユーザーの行動を認識し、オートエージェントを構築するためのツール。

用途: オートエージェント構築
難易度: Easy
コスト: High

unsloth — Unsloth is a local UI for training and running Gemma 4, Qwen3.6, DeepSeek, Kimi, GLM and other models.

Unsloth Studioは、オープンモデルのトレーニングと実行を支援するWebUIです。このライブラリは、Gemma4、Qwen3.5などのオープンモデルのテストとトレーニングを支援するために使われます。

自然言語処理大規模言語モデルテキスト音声

用途: オープンモデルのトレーニングと実行
難易度: Easy
コスト: High

machine-learning-for-trading — Code for Machine Learning for Trading, 3rd edition — from data sourcing to live execution.

LLMの推論 Transparency を高めるために、DiffusionGemmaの計算を分離しVariable Transparency とAlgorithmic Transparencyを評価します。

強化学習

用途: LLMの透明性、誤用、過度安定化を理解する
難易度: Easy
コスト: High

自然言語処理大規模言語モデルテキストマルチモーダル

ai-agent-book — 《深入理解 AI Agent：设计原理与工程实践》（李博杰著）开源主仓库：全书正文、编译版 PDF 与按章配套代码

この論文では、現在のVision-Language-Benchmark（VLB）を超える、MLLMがアクティブな観察を実演できるようにするためのバenchmark、ActiveVisionを提案する。このActiveVi

用途: 弁論の実際的な対象を形成するためにAIが活用される
難易度: Easy
コスト: High

stable-baselines3 — PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.

このリポジトリでは、LLMベースのエージェントアプリケーションのための強化学習の橋渡しを提供しています。

強化学習

用途: 強化学習を簡素化させる橋渡し
難易度: Easy
コスト: High

ART — Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!

ARTは、多段強化学習トレーナーです。このトレーナーは、GRPOを使用して、現実世界のタスクに対して、多段強化学習を行うことができます。

自然言語処理大規模言語モデル強化学習

用途: 多段強化学習トレーナー
難易度: Easy
コスト: High