Extract structured intelligence from audio using the AssemblyAI API with sentiment analysis, entity detection, topic modeling, and auto-chapter generation.
Diagnose and fix AssemblyAI common errors and exceptions. Use when encountering AssemblyAI errors, debugging failed transcriptions, or troubleshooting streaming and LeMUR issues.
Execute AssemblyAI primary workflow: async transcription with audio intelligence. Use when transcribing audio/video files, enabling speaker diarization, sentiment analysis, entity…
Execute AssemblyAI streaming transcription and LeMUR workflows. Use when implementing real-time speech-to-text, live captions, voice agents, or LLM-powered audio analysis with…
Optimize AssemblyAI costs through model selection, feature budgeting, and usage monitoring. Use when analyzing AssemblyAI billing, reducing transcription costs, or implementing…
Create a minimal working AssemblyAI transcription example. Use when starting a new AssemblyAI integration, testing your setup, or learning basic transcription patterns.
Optimize AssemblyAI API performance with caching, parallel processing, and model selection. Use when experiencing slow transcriptions, implementing caching strategies, or…
Execute AssemblyAI production deployment checklist and rollback procedures. Use when deploying AssemblyAI integrations to production, preparing for launch, or implementing go-live…
Streams audio from Twilio Media Streams over WebSocket to AssemblyAI real-time transcription, extracting speaker-diarized transcripts with word-level timestamps.
Transcribes audio and generates auto-chapters with summaries using AssemblyAI's /v2/transcript endpoint with auto_chapters=true.
Transcribe audio/video with AssemblyAI (local upload or URL), plus subtitles + paragraph/sentence exports.
Implement AssemblyAI webhook handling for transcription completion events. Use when setting up webhook endpoints, handling transcription callbacks, or processing async…
Use when when you have aligned ATAC-seq BAM files and need to quantify Tn5 transposase insertion patterns around specific genomic coordinates (motif sites, peaks, regulatory…
Use when you have aligned ATAC-seq BAM files and need to detect transcription factor binding sites via footprint analysis.
Use when you have raw ATAC-seq BAM files and need to perform footprinting analysis to detect transcription factor binding through Tn5 insertion patterns.
음성 합성(TTS), 알림음, 진동, 접근성 서비스 통합. Expo Speech, 사운드 재생, 화면 읽기, 움직임 감소 설정. Use when: (1) TTS 음성 안내, (2) 알림음/진동 구현, (3) 접근성 기능 개발, (4) VoiceOver/TalkBack 지원.
Audio analysis with Tone.js and Web Audio API including FFT, frequency data extraction, amplitude measurement, and waveform analysis.
Comprehensive audio analysis with waveform visualization, spectrogram, BPM detection, key detection, frequency analysis, and loudness metrics.
Generate a concise, audio-ready executive briefing script from data across your systems. Use when the user wants a morning brief, situation update, or a summary they can listen to…
Convert audio files between formats (MP3, WAV, FLAC, OGG, M4A) with bitrate and sample rate control. Batch processing supported.
Automatically helps debug Web Audio API issues, audio playback problems, pitch preservation, and caching issues in the VSSK-shadecn music practice app
Làm sạch bản ghi giọng nói WAV/MP3 theo workflow 2-phase semantic — AI viết lại nội dung không lặp vào TOML, sau đó căn keep flag từng token để render audio cuối.
Comprehensive onset diagnosis - determine if onset is noise or real kalimba note
Expert in digital signal processing for audio applications. Validates biquad filter implementations, frequency response calculations, and audio algorithms.
Master the essential audio post-production techniques—normalization, compression, EQ, and noise reduction—using the correct processing order to achieve professional-quality audio.
FFmpeg audio processing, batch editing, normalization, mixing, and automated audio production workflows.
Create standard SuperCollider audio effects for Bice-Box (delays, reverbs, filters, distortions). Provides templates, ControlSpecs, common patterns, and MCP workflow for safely…
Audio engineering — mastering, mixing, EQ, compression, loudness standards, synthesis, podcast production, music theory, spectrum analysis.
Audio production concepts, DSP fundamentals, mixing/mastering techniques, and DAW workflows. Bridges modular synthesis philosophy with practical audio engineering.
Remove background noise, enhance clarity, and improve audio quality. Perfect for podcasts, videos, and calls. No API keys needed. $2 FREE credits to start.
从视频文件中提取音频。Use when user wants to 提取音频, 抽取音频, 视频转音频, 导出音频, extract audio, video to audio, get audio from video, 把视频的声音提取出来.
ffmpeg patterns for extracting audio from video files and transcoding between formats
You are the audio architecture expert ensuring Leavn's complex audio pipeline stays coherent.
You are the audio fingerprinting and pattern detection specialist for Modcaster's content analysis.
Identifies audio content using Chromaprint/AcoustID fingerprinting, Shazam API recognition, and ACRCloud monitoring.
통합 오디오 생성 스킬. ElevenLabs MCP 기반 TTS(32개국어), 보이스 클로닝(1분 샘플), 다국어 더빙(립싱크), 효과음 생성을 지원. "목소리 생성", "TTS", "음성 합성", "보이스 클로닝", "더빙", "나레이션", "효과음", "AI 음성" 요청 시 사용.
Use whenever the user asks to install, configure, uninstall, snooze, mute, test, troubleshoot, or change settings for the claude-code-audio-hooks audio notification system.
Test Bob The Skull with virtual audio injection instead of speaking. Use when testing wake word detection, STT accuracy, full conversation pipeline, or automated testing.
Audio generation skill — jingles, beds, voiceover, and sound effects. Routes music requests to Suno V5 / Udio / Lyria, speech to MiniMax TTS / FishAudio / ElevenLabs V3, and SFX…
Create memorable sonic logos using design principles from Intel, Netflix, and McDonald's—crafting 2-5 second audio signatures that achieve instant brand recognition.
You are the on-device audio ML specialist for Modcaster's AI-driven audio processing.
Use when writing songs, generating music or sound with AI, preparing Suno/HeartMuLa prompts, or analyzing audio features and spectrograms.
Use when asked to normalize audio volume, match loudness, or apply peak/RMS normalization to audio files.
Audio playback using Tone.js including players, transport, scheduling, and loading audio. Use when implementing background music, sound effects, audio synchronization, or timed…
Audio ingestion, analysis, transformation, and generation (Transcribe, TTS, VAD, Features).
Converts and processes audio files using ffmpeg. Supports format conversion, sample rate changes, mono/stereo conversion, and segment splitting.
Professional audio production for music, podcasts, and sound design. Use when working with audio recording, mixing, mastering, or sound design for any medium.
Analyze audio recording quality - echo detection, loudness, speech intelligibility, SNR, spectral analysis.
Analyze the WaveCap-SDR audio stream to assess tuning quality, detect silence, noise, proper audio, or distortion.
Binding audio analysis data to visual parameters including smoothing, beat detection responses, and frequency-to-visual mappings.
Generate audio replies using TTS. Trigger with "read it to me [URL]" to fetch and read content aloud, or "talk to me [topic]" to generate a spoken response.
Router for audio domain including playback, analysis, and audio-reactive visuals. Use when implementing any audio functionality including music, sound effects, visualizers, or…
Separates audio tracks into individual stems (vocals, drums, bass, other) using Meta's Demucs neural network model via the demucs Python package.
팟캐스트 대본작가(scriptwriter)와 쇼노트편집자(shownote-editor)가 사용하는 오디오 스토리텔링 전문 스킬. 귀로만 듣는 매체에서 청취자의 몰입을 극대화하는 서사 구조, 페이싱, 사운드 연출 방법론을 제공한다.
Implements audio systems including sound management, music systems, positional audio, and audio effects.
Game audio systems, music, spatial audio, sound effects, and voice implementation. Build immersive audio experiences with professional middleware integration.
Turn a dialogue, a story with dialogue, or a one-line idea into finished multi-character audio - a radio drama, a two-host podcast, or clean lip-sync voice clips - 100% locally…
Turn creator audio into clean text captions for ecommerce content and reuse. Use when teams need fast transcript-to-caption workflows.
End-to-end audio production workflow with stems, effects, archiving, and verification
Step-by-step audio production with per-stem verification, timing alignment, and incremental quality gates