Claude Code Skills·Claude Skills·The open SKILL.md registry for Claude
ClaudSkillsContent › Audio Podcast › Page 9

Audio Podcast (Page 9 of 12)

688 Claude Code skills in the Audio Podcast sub-category of Content.

688 skills · updated 2026-08-26 · showing 481–540 of 688 by quality score

For the full experience including quality scoring and one-click install features for each skill — upgrade to Pro.

Use when analyzing transcription factor (TF) regulatory networks using Dorothea database. Input gene list, identify regulating transcription factors, generate TF-Target network…
Generate, convert, clean, and prepare audio assets for Three.js browser games using ElevenLabs. Use for sound effects, looping ambience, UI sounds, impact/weapon/vehicle audio,…
Primary entrypoint for complete Three.js browser game creation, premium iteration, and automatic phase orchestration.
Produces word- and segment-timed transcripts with WhisperX or faster-whisper, optional Demucs vocal stems, then SRT, VTT, ASS karaoke, or JSON for lip-sync.
Use when you have ATAC-seq BAM files and a set of genomic coordinates (e.g., transcription factor motif sites, peak regions) and need to quantify the spatial distribution of Tn5…
Audio signal processing library for PyTorch. Covers feature extraction (spectrograms, mel-scale), waveform manipulation, and GPU-accelerated data augmentation techniques.
Transcribes video audio using WhisperX, preserving original timestamps. Creates JSON transcript with word-level timing. Use when you need to generate audio transcripts for videos.
Produce timestamped transcript sidecars for acquired audio/video with hashes, source metadata, speaker labels when available, and explicit degraded plans when STT tooling is…
Transcribe YouTube videos and local audio/video files with speaker diarization. Use when user asks to transcribe a YouTube URL, podcast, video, or audio file.
Get transcripts from any YouTube video — for summarization, research, translation, quoting, or content analysis.
Transcribes audio and video files through case.dev with speaker diarization. Supports MP3, WAV, M4A, FLAC, OGG, WEBM, MP4 up to 5GB.
Automate audio/video transcription, meeting notes, subtitle generation, and content processing
Use when after bias-correcting ATAC-seq cutsite signal (using ATACorrect or equivalent) when you have a set of genomic regions of interest (e.
Use when you have a chromVARDeviations object with precomputed bias-corrected deviations and z-scores for multiple annotation sets (e.
Use when after bias-corrected ATAC-seq signal tracks (bigWig files) have been generated and you need to quantify transcription factor binding strength within open chromatin…
Use when you have ATAC-seq BAM files and want to detect transcription factor footprints—characteristic depletion patterns of Tn5 insertions around protein-bound motif sites.
Use when you have corrected ATAC-seq footprint scores (from ATACorrect and ScoreBigwig) at open chromatin regions and a motif database (e.g., JASPAR PWMs), and you need to…
Transform podcast transcripts into multiple content assets—blog posts, social snippets, newsletters, and SEO-optimized landing pages—using systematic repurposing workflows.
Use when you have RNA-seq read counts (FPKM or similar) for multiple cell lines or biological samples, a genome-scale metabolic model with GPR associations, and you need to…
One-off transcription of local audio or video files to text or subtitle files using Transloadit via the official `@transloadit/node` CLI.
Use when the user invokes /transkrip-youtube or asks to summarize/transcribe a YouTube video. Uses yt-dlp to stream subtitles directly to stdout via a tempdir with trap-based…
Explicit-entry skill for Gemini TTS audio. Invoked deliberately via the /tts-duet command (generation) and /tts-duet-setup (configuration); not auto-triggered.
Intégration Tabletop Simulator (TTS) de mc4db-2.0 — export JSON côté backend (decks, packs, scénarios) et import Lua côté TTS (scripts dans mc_tools).
Atomic reference for @panda-video-generator/tts-node: pnpm tts, cli.ts, processNarrationFile — Edge-TTS narration → audio.mp3 + audio.vtt; env vars TTS_*, EDGE_TTS_*, ffmpeg.
Multi-engine text-to-speech skill. Supports Qwen3-TTS local voice cloning, VoiceCraft online TTS, and OpenAI TTS.
Send high-quality text-to-speech voice messages on WhatsApp in 40+ languages with automatic delivery
面向团队与企业用户的 WhatsApp 语音消息工具(专业版)。核心能力: - 涵盖免费版全部能力(Piper TTS、40+ 语言、单条发送) - 群组广播:发送到 WhatsApp 群组 - 批量发送:联系人列表群发 - 定时发送:cron 任务自动发送 - 消息模板:变量替换与个性化 - 多语言批量:一次任务多语言消息 - 发送队列 — from…
基于 Piper TTS 的 WhatsApp 语音消息发送工具,支持 40+ 语言,适合个人用户发送语音消息。Use when 需要文本翻译、多语言转换、本地化处理时使用。不适用于专业医学法律翻译认证。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。
面向团队与企业用户的 WhatsApp 语音消息工具(专业版)。核心能力: - 涵盖免费版全部能力(Piper TTS、40+ 语言、单条发送) - 群组广播:发送到 WhatsApp 群组 - 批量发送:联系人列表群发 - 定时发送:cron 任务自动发送 - 消息模板:变量替换与个性化 - 多语言批量:一次任务多语言消息 - 发送队列 — from…
Diagnose and fix TwinMind common errors and exceptions. Use when encountering transcription errors, debugging failed requests, or troubleshooting integration issues.
Execute TwinMind primary workflow: Meeting transcription and summary generation. Use when implementing meeting capture, building transcription features, or automating meeting…
Create your first TwinMind meeting transcription and AI summary. Use when starting with TwinMind, testing your setup, or learning basic transcription and summary patterns.
Incident response for TwinMind failures: transcription not starting, audio not captured, sync failures, and calendar disconnect.
Install and configure TwinMind Chrome extension, mobile app, and API access. Use when setting up TwinMind for meeting transcription, configuring calendar integration, or…
Set up local development workflow with TwinMind API integration. Use when building applications that integrate TwinMind transcription, testing API calls locally, or developing…
Monitor TwinMind transcription quality, meeting coverage, action item extraction rates, and memory vault health.
Optimize TwinMind transcription accuracy and speed with Ear-3 model configuration, audio quality tuning, and caching strategies.
Handle TwinMind meeting events including transcription completion, action item extraction, and calendar sync notifications.
Fetch Evolutionary Conservation scores (phyloP, phastCons) and Transcription Factor Binding Sites (TFBS) from the UCSC Genome Browser.
Queries the UniBind database for experimentally validated transcription factor (TF) binding sites. Use when retrieving direct TF-DNA interaction datasets, downloading binding site…
Upload local files or user attachments to the Gradio server and get back public URLs. Use when: upload image, upload XML, upload attachment, file to URL, host file, upload for…
Use when the user wants local voice transcription instead of OpenAI Whisper API. Switches to whisper.cpp running on Apple Silicon. WhatsApp only for now.
Async music / audio-track generation via Venice. Covers the /audio/quote + /audio/queue + /audio/retrieve + /audio/complete lifecycle, lyrics vs instrumental, voice selec — from…
Generate speech from text via POST /audio/speech. Covers TTS models (Kokoro, Qwen 3, xAI, Inworld, Chatterbox, Orpheus, ElevenLabs Turbo, MiniMax, Gemini Flash), voices p — from…
Transcribe audio files to text via POST /audio/transcriptions. Covers supported models (Parakeet, Whisper, Wizper, Scribe, xAI STT), supported formats (wav/flac/m4a/aac/m — from…
Async music / audio-track generation via Venice. Covers the /audio/quote + /audio/queue + /audio/retrieve + /audio/complete lifecycle, lyrics vs instrumental, voice selec — from…
Generate speech from text via POST /audio/speech. Covers TTS models (Kokoro, Qwen 3, xAI, Inworld, Chatterbox, Orpheus, ElevenLabs Turbo, MiniMax, Gemini Flash), voices p — from…
Transcribe audio files to text via POST /audio/transcriptions. Covers supported models (Parakeet, Whisper, Wizper, Scribe, xAI STT), supported formats (wav/flac/m4a/aac/m — from…
Convert text to speech audio using OpenAI TTS-1-HD through the verging.ai proxy API. Supports multiple voices, playback speed control, and various audio output formats.
Internationales Handelsrecht: Fortschritts-Dashboard und nächste Schritte für komplexe internationale Handelsfälle.
Use whenever a hand-written or scanned answer PDF needs transcription to markdown for /grade. Three tiers — Claude native vision (default, no extra install), local Qwen3-VL 8B via…
vLLM non-chat inference surfaces — text embeddings (`/v1/embeddings`, `/v2/embed`), reranking/scoring (`/rerank`, `/score`), speech-to-text (`/v1/audio/transcriptions`,…
Generate audio and speech with vLLM-Omni using Qwen3-TTS, Fish Speech S2 Pro, CosyVoice3, MiMo-Audio, and Stable-Audio models.
Integrate a new text-to-speech model into vLLM-Omni from HuggingFace reference implementation through production-ready serving with streaming and CUDA graph acceleration.
Use when the user has recorded vocals and wants a processing chain set up. Examples - "set up my vocal chain", "process this vocal", "make the vocal sit in the mix", "give me a…
Handles voice-to-voice conversations on WhatsApp. Automatically transcribes incoming audio and responds with local TTS audio.
Voice AI — text-to-speech (ElevenLabs, OpenAI TTS), speech-to-text (Whisper), voice cloning, real-time voice
AI voice agent for handling incoming calls, appointment scheduling, lead qualification, and 24/7 customer service without human intervention.
Build voice-based AI agents for phone calls, meetings, customer support, and sales qualification using Vapi,
Expert in building voice AI applications - from real-time voice agents to voice-enabled apps. Covers OpenAI Realtime API, Vapi for voice agents, Deepgram for transcriptio — from…
All Content skills →
More in ContentStorytelling (1,078) · Video (848) · Translation (654) · Editorial (422) · Writing (350) · Image Design (294)