面向团队与企业的本地离线文字转语音引擎. 核心能力: 批量合成、自定义音色训练、多语言支持、SSML 标记、API 服务化、语音后处理、跨平台部署. 适用场景: 内容批量配音、多语言客服、有声书制作、无障碍服务、企业通知语音化. 差异化: 专业版在免费版基础上新增批量处理与音色定制,兼容免费版合成命令与音色模型.
基于 Piper 的本地离线文字转语音工具,零云端调用、零 API 费用,适合个人单条语音生成。Use when 需要API集成、接口对接、Webhook配置、系统连接时使用。不适用于逆向工程闭源API。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。
Podcast knowledge workflows powered by Podwise CLI: search podcasts and episodes by keyword, monitor followed shows for new releases, find popular episodes, ask questions and…
Pollinations.ai API for AI generation - text, images, videos, audio, and analysis. Use when user requests AI-powered generation (text completion, images, videos, audio,…
Transcribe audio or video through the PopiArt runtime baseline for local-first speech-to-text. Use this when the user wants a PopiArt-managed STT path for transcripts, captions,…
Convert text to speech through the PopiArt runtime baseline for multi-model TTS. Use this when the user wants one general text-to-speech entry point without handling upstream…
Use when you have ATAC-seq BAM alignments with classified motif sites (bound vs. unbound based on chromatin accessibility or binding thresholds) and wish to detect and visualize…
转录后调控 workflow skill。用于 alternative splicing、differential splicing、Ribo-seq、small RNA/miRNA、CLIP-seq/RBP binding、m6A/epitranscriptomics、RNA modification calling 和 RNA 调控可视化交接。
Ordered quality gate checklist to run after every code change in react-native-audio-api. Covers formatting, linting, type checking, C++ tests, JS tests, and enum sync validation.
Build an Apple Push to Talk channel experience with system controls, ephemeral PTT pushes, audio-session ownership, and an app-owned communication backend.
pyannote.audio is an open-source Python toolkit for speaker diarization built on PyTorch. It provides state-of-the-art pretrained models and pipelines for speech activity…
pydub is a Python library that provides a simple, high-level interface for manipulating audio files. It supports slicing, concatenation, volume adjustment, crossfading, format…
Complete SDK for controlling Reachy Mini robot - head movement, antennas, camera, audio, motion recording/playback.
Preprocesses photographed sheets of many business cards — slicing each into overlapping high-resolution tiles and de-glaring them with container tooling (OpenCV/ImageMagick) —…
Real-time audio playback patterns for macOS Apple Silicon. TRIGGERS - audio jitter, tts choppy, sounddevice, afplay jitter, audio architecture, playback glitch, GIL contention…
Deep research and fresh perspective on current blockers. Use when stuck, for methodology validation, or auto-triggered during long optimization runs.
Ressortaufgaben AA: typische Legistik-Aufgaben im Geschäftsbereich Auswaertiges Amt. Klaert Vorhabenart; Begruendungspflichten; Verbaendeanhoerung nach GGO Paragraf 47;…
Ressortaufgaben BMAS: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Arbeit und Soziales.
Ressortaufgaben BMBFSFJ: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Bildung; Familie; Senioren; Frauen und Jugend.
Ressortaufgaben BMDS: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Digitales und Staatsmodernisierung.
Ressortaufgaben BMDS: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Digitales und Staatsmodernisierung.
Ressortaufgaben BMF: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium der Finanzen. Klaert Vorhabenart; Begruendungspflichten; Verbaendeanhoerung nach GGO Paragraf…
Ressortaufgaben BMFTR: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Forschung; Technologie und Raumfahrt.
Ressortaufgaben BMG: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Gesundheit.
Ressortaufgaben BMI: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium des Innern. Klaert Vorhabenart; Begruendungspflichten; Verbaendeanhoerung nach GGO Paragraf…
Ressortaufgaben BMJV: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium der Justiz und für Verbraucherschutz.
Ressortaufgaben BMLEH: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Landwirtschaft; Ernaehrung und Heimat.
Ressortaufgaben BMUKN: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Umwelt; Klimaschutz; Naturschutz und nukleare Sicherheit.
Ressortaufgaben BMV: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Verkehr. Klaert Vorhabenart; Begruendungspflichten; Verbaendeanhoerung nach GGO — from…
Ressortaufgaben BMV: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Verkehr. Klaert Vorhabenart; Begruendungspflichten; Verbaendeanhoerung nach GGO — from…
Ressortaufgaben BMVg: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium der Verteidigung.
Ressortaufgaben BMWE: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Wirtschaft und Energie.
Ressortaufgaben BMWSB: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für Wohnen; Stadtentwicklung und Bauwesen.
Ressortaufgaben BMZ: typische Legistik-Aufgaben im Geschäftsbereich Bundesministerium für wirtschaftliche Zusammenarbeit und Entwicklung.
Nutze diesen Skill, wenn Rücktritts- oder Widerrufsfolgen neben Bereicherungsrecht stehen. Normen: §§ 346 bis 359 BGB; § 812 BGB; §§ 355 bis 361 BGB.
Use Sarvam AI for Indian language Text-to-Speech (TTS), Speech-to-Text (STT), Translation, and Chat.
Give Claude the ability to see your screen, watch it in real-time with audio transcription, analyze video files, and extract text via OCR.
Query the user's screen recordings, audio, UI elements, and usage analytics via the local Screenpipe REST API at localhost:3030.
Reference skill for Zoom AI Services Scribe. Use after routing to a transcription workflow when handling uploaded or stored media, Build-platform JWT auth, fast mode…
Use the moment you're about to tell the user you can't do something — or about to suggest they sign up for, get an API key for, or go to an external tool, site, or API to — from…
Use Speaches when an agent stack expects OpenAI-style audio endpoints but you want a self-hosted speech backend for transcription, translation, and text-to-speech instead of a…
Text-to-Speech using SiliconFlow API (CosyVoice2). Supports multiple voices, languages, and dialects.
Single-molecule FISH spot detection and per-cell transcript counts for Miller-Jensen lab bursting / HIV-transcription work.
Music and audio file analysis agent. Produces reproducible Python/Shell analysis pipelines (librosa, pyloudnorm, essentia, madmom, mutagen, ffprobe) for BPM/key/time-signature…
Songwriting craft, AI music generation prompts (Suno focus), parody/adaptation techniques, phonetic tricks, and lessons learned. These are tools and ideas, not rules.
Sound designer senior. Foley, ambient, SFX, mixing, mastering, podcast, video, game.
Transcribe speech to text using the Speech framework. Use when implementing live microphone transcription with AVAudioEngine, recognizing pre-recorded audio files, configuring…
Transcribe audio to text with Whisper models via inference.sh CLI. Models: Fast Whisper Large V3, Whisper V3 Large.
Transcribe audio to text with Whisper models via inference.sh CLI. Models: Fast Whisper Large V3, Whisper V3 Large.
Transcribe audio to text using ElevenLabs Scribe v2. Use when converting audio/video to text, generating subtitles, transcribing meetings, or processing spoken content.
Expert skill for implementing speech-to-text with Faster Whisper. Covers audio processing, transcription optimization, privacy protection, and secure handling of voice data for…
Install and use the speechall CLI tool for speech-to-text transcription. Use when the user wants to: (1) transcribe audio or video files to text, (2) install speechall on macOS or…
Spleeter is Deezer's open-source audio source separation library with pretrained models. It can split audio into 2, 4, or 5 stems (vocals, drums, bass, piano, accompaniment) and…
Staaten- und Gebietscheck St. Kitts und Nevis: migrationsrechtlicher Workflow für Herkunfts-, Transit-, Dokumenten-, Visum-, Schutz-, Passbeschaffungs-, Rückführungs- und…
Generate Chinese / Japanese speech with StepFun's stepaudio-2.5-tts — Contextual TTS that replaces step-tts-2's `voice_label` with natural-language `instruction` (≤200 chars) plus…
Implement speech-to-text voice input in Blazor applications using Syncfusion SpeechToText component. ALWAYS use this when users need voice input, speech recognition, audio…
Telegram voice-to-voice for macOS Apple Silicon: transcribe inbound .ogg voice notes with yap (Speech.framework) and reply with Telegram voice notes via say+ffmpeg.
Generate speech from text using Telnyx and third-party TTS providers (AWS, Azure, ElevenLabs, MiniMax, Resemble, Rime, xAI). Returns base64-encoded audio or a binary stream.
Expert guidance for LocalAI, the open-source drop-in replacement for OpenAI's API that runs locally. Helps developers self-host LLMs, image generators, audio transcription, and…
The TF-differential-binding pipeline performs differential transcription factor (TF) binding analysis from ChIP-seq datasets (TF peaks) using the DiffBind package in R.