Skip to content

feature: speaker identification — SPEAKER_N 對映 names[] 真名(兩階段之二,blocked by #25) #26

Description

@kiki830621

Problem

兩階段裁決的第二階段(sibling of #25):#25 交付 diarization(SPEAKER_1/2 聲學分離)後,本 issue 追蹤 speaker identification——對照已知聲紋辨識出「具體是誰」,把 SPEAKER_N 自動映射成 context names[] 真名。

Type

feature

Expected

使用者可為 context names[] 的人提供 enrollment 聲紋樣本;轉寫時 SPEAKER_N 自動對映真名(含信心門檻與未知說話者 fallback)。

Actual

(依 #25 交付後)僅有匿名 SPEAKER_N 標記;真名對齊靠 srt-proofread 語意推測。

Impact

Blocked by #25

Clarity Surface(idd-clarify run 2026-07-03T07:05:00Z)

Type Source Suggested canonical Status
ambiguity 「提供 enrollment 聲紋樣本」的介面 (a) 專用 CLI 指令(新 surface + 狀態管理)vs (b) context-dir 慣例資料夾 voices/<name>.wav(零新指令、沿 context-calibration 資料夾哲學、三層 precedence 免費繼承) resolved @ 2026-07-03T07:05:00Z (reason: 採 (b)——issue 本文即綁定 context names[];design 決策、與使用者既有裁決一致)
ambiguity 聲紋「儲存」形式 原始音檔(使用者自管、可讀可刪)vs 預計算 embedding(另一格式的生物特徵、需失效管理) resolved @ 2026-07-03T07:05:00Z (reason: 使用者放原始音檔、embedding 運行時計算——local-only 鐵律下音檔最透明;快取為後續可選)

Current Status

Metadata

Metadata

Assignees

No one assigned

    Labels

    enhancementNew feature or request

    Type

    No type

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions