Project overview
audio-to-action transcribes a recording, cleans speaker sections, classifies one of five scenes, and selects a matching report template for meetings, discussions, memos, or progress updates. Tasks and message drafts carry timestamp evidence and explicit uncertainty, while external sending always waits for user approval. The current version has no real speaker diarization, automatic splitting of very long audio, cross-meeting memory, streaming, or automatic sending, and needs a configured ASR endpoint; authentication, when required, comes from an environment variable.
Repository facts
- Primary language
- Python
- License
- MIT
- Repository updated
- May 7, 2026
- Default branch
- main
Resource types
General skill
Use cases
Productivity and officeCommunication and collaborationMedia, audio and video
Platforms
Claude Code, Codex, and more
Capabilities
Transcription