AILP — AI 学習許可プロトコル
AI 学習許可プロトコル
AI はこれから学習してよいのか?どの深さまで、どの用途のために、どんな義務の下で?
01 · 定義
AILP operationalizes the AIRS spectrum for the specific question of learning. It distinguishes acts that binary crawler rules collapse together: being read is not being learned from; being retrieved is not being trained on; fine-tuning is not distillation; verbatim memorization is different from statistical internalization. AILP lets a publisher declare, in a single machine-readable file, per-dimension permissions — access, indexing, inference input, embedding, training, fine-tuning, distillation, memory, output, attribution and compensation — each as allowed, denied, or license-required. An AILP grant answers "may AI learn from this content?" only — it does not answer "may this Agent invoke a tool, write to a database, or publish?" (AARS), nor "who authorized this Agent?" (AADP).
02 · 目的
- 「AI はこれから学習してよいか?」という問いに、クロールビットではなく次元レベルの精度で答えます。
- 学習許可(モデルへの内在化)をアクセス許可(ページの取得)から分離します。
- 補償モデルをサポートします:無料、非商用、要ライセンス、収益分配。
- 開放性を機械実行可能にします — 善意が法的不確実性として捨てられないように。
03 · 範囲
access / indexing — 取得と意味的索引化
inference_input — 推論時のコンテキストとしての利用(RAG)
embedding — ベクトル化と検索インデックス
training / fine_tuning / distillation — モデル内在化の階層
verbatim_memory — 逐語的な再現が許可されるか
attribution / compensation — 引用義務と支払い条件
04 · 機械可読の例
学習許可宣言 — /ai/rights-spectrum.json(AILP プロファイル)
{
"version": "0.1",
"protocol": "AILP",
"publisher": "example.org",
"default": {
"access": "allowed",
"indexing": "allowed",
"inference_input": "allowed",
"embedding": "allowed",
"training": "license_required",
"fine_tuning": "license_required",
"distillation": "denied",
"verbatim_memory": "denied",
"attribution": "required",
"compensation": "contact"
},
"contact": "licensing@example.org"
} 05 · 限界
- 学習許可のセマンティクスは技術的な検証が最も難しい領域です — AILP は意図を宣言するものであり、遵守を証明することはまだできません。
- 次元リストは草案であり、訓練・ファインチューニング・蒸留の境界はなお議論の途上です。
- 学習許可の法的な承認は法域によって異なり、未解決です。
- v0.1.1: AILP(Resource,ContentUse)=Allow does not imply AARS(Actor,Action)=Allow, and does not imply AADP(Principal,Actor,Authority)=Valid — a system exposing both content and tools must evaluate them separately.