stt

Imports Clova Note transcripts and transcribes Apple Voice Memos or exported recordings locally on your Mac with whisper.cpp, saving everything as Markdown notes for a vault. Apple recordings are never uploaded.

#stt#transcription#clovanote#whisper#markdown#obsidian#plugin
Install

Installation

/plugin marketplace add hjsh200219/stt /plugin install stt@stt /reload-plugins
Requirements
  • Claude Code (CLI)
  • 플러그인 설치만으로 전사 엔진·모델은 설치되지 않음 — Apple 전사에는 ffmpeg · whisper-cli(whisper.cpp) · GGML 모델 필요
  • 클로바노트: 네이버 로그인과 `~/.clovanote/.env` 설정, 로그인용 playwright. 첫 로그인(`--seed`)에서 보호조치·CAPTCHA는 사람이 1회 해제
  • 클로바노트는 공개 API가 없어 웹 SPA의 내부 API를 로그인 세션 쿠키로 재현 — 서비스 변경 시 동작하지 않을 수 있음
  • Apple 전사는 타임스탬프만 있고 화자 분리·요약은 없으며, Whisper 오인식·무음 환각 가능성이 있어 원음 대조 필요
Capabilities

What it does

  • 01`/stt:clovanote` — 클로바노트 노트를 vault(`CLOVANOTE_OUT`)에 적재, `list`로 목록만, `auth`로 세션 로그인·갱신
  • 02클로바노트는 서비스에 이미 생성된 전사·화자 구분을 가져옴 (세션을 만든 뒤에는 브라우저 없이 동작)
  • 03`/stt:apple` — Apple 음성 메모·내보낸 음성 파일을 whisper.cpp로 로컬 전사해 Markdown으로 저장, `list`·`doctor`로 녹음 수·엔진·모델 점검
  • 04지원 오디오 — m4a·wav·mp3·aac·aiff·flac·caf (ffmpeg로 임시 WAV 변환). 이미 처리한 녹음은 중복 적재 방지
  • 05Obsidian 등 Markdown 기반 vault와 함께 사용 가능
About

About

stt is a Claude Code plugin that turns spoken recordings into Markdown notes you can drop into a knowledge vault such as Obsidian. It has one skill per source: /stt:clovanote imports transcripts and speaker labels that Naver Clova Note has already produced, and /stt:apple transcribes Apple Voice Memos or exported audio files locally with whisper.cpp. Apple transcription never uploads the recording. Installing the plugin does not install the transcription engine or model, so ffmpeg, whisper-cli and a GGML model have to be set up separately. Clova Note has no public API; the plugin replays the web app's internal API with a logged-in session, and the first login needs a person to clear Naver's protection step once. Transcripts carry timestamps but no speaker separation or summary, and Whisper can mishear or hallucinate on silence, so important passages should be checked against the audio.