harness · market
Harness
Put raw intelligence to work.
Skills, MCP servers, prompts, assistants and connectors. Every listing shows its author, version and the permissions it needs.
On the shelves20,989
- MCP servers18,838
- Skills1,282
- Assistants651
- Prompts218
0 items
Audio and speech- msitarzewski 158kGame Audio EngineerInteractive audio specialist - Masters FMOD/Wwise integration, adaptive music systems, spatial audio, and audio performance budgeting across all game enginesAssistantAudio and speechGames
anthropics 28kConversation analysisAnalyzes sales call transcripts to extract brand voice patterns, messaging effectiveness, and tone variations. Use this agent when processing multiple transcripts or performing deep pattern recognition across conversations.AssistantSonnetCRM and salesAudio and speechMarketing and SEO
screenpipe 22kscreenpipeSearch your local screen recordings, audio transcripts, and computer activity from screenpipe.MCPNode.jsAudio and speech
kkkkhazix 21k卡兹克公众号长文写作数字生命卡兹克(Khazix)的公众号长文写作skill。当用户需要撰写公众号文章、写稿子、续写文章、根据素材产出长文时使用。触发词包括但不限于:写文章、写稿子、帮我写、续写、扩写、公众号文章、长文、出稿、按我的风格写。即使用户只是说"帮我把这个写成文章"或"用我的风格写一下",只要上下文涉及内容创作和公众号输出,都应该触发。也适用于用户丢过来一个PDF、brief、新闻链接、语音转文字或任何素材说"帮我写篇文章"的场景。不要用于短内容(小红书帖子、推特、朋友圈)或纯标题摘要生成(那个用wechat-title skill)。SkillNo codeAudio and speechWritingNews
orchestra-research 13kWhisper - Robust Speech RecognitionOpenAI's general-purpose speech recognition model. Supports 99 languages, transcription, translation to English, and language identification. Six model sizes from tiny (39M params) to large (1550M params). Use for speech-to-text, podcast transcription, or multilingual audio processing. Best for robust, multilingual ASR.SkillNo codeAudio and speechTranslationAI models
wshobson 40kLocal file conversionUse when the user needs a local file converted between common image, audio, video, document, or data formats. Select installed tools, preserve originals, and verify the output.SkillNo codeImagesVideoAudio and speech
builderio 7.1kAgent-Native ClipsScreen recording, meeting notes, and voice dictation - all with AIMCPRemoteNotes and knowledgeAudio and speech
anthropics 28kDocument analysisAnalyzes brand documents to extract voice attributes, messaging, terminology, and examples. Use this agent when processing multiple brand documents or performing cross-document pattern recognition.AssistantSonnetDocumentsAudio and speechMarketing and SEO
huggingface 11kTransformers.js - Machine Learning for JavaScriptUse Transformers.js to run state-of-the-art machine learning models directly in JavaScript/TypeScript. Supports NLP (text classification, translation, summarization), computer vision (image classification, object detection), audio (speech recognition, audio classification), and multimodal tasks. Works in browsers and server-side runtimes (Node.js, Bun, Deno) with WebGPU/WASM using pre-trained models from Hugging Face Hub.SkillNo codeBrowser automationImagesAudio and speech
vercel-labs 32kWriting GuidelinesReview docs/prose for Writing Guidelines compliance. Use when asked to "review my docs", "check writing style", "audit prose", "review docs voice and tone", or "check this page against the writing handbook".SkillNo codeAudio and speechWritingLegal
pollinations 5.2kFFmpegTrim, convert, resize, compress, and remix audio and video.MCPRemoteVideoAudio and speech
google-gemini 4.3kGemini Live API Development SkillUse this skill when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live. Covers WebSocket-based audio/video/text streaming, voice activity detection (VAD), background reasoning (extended thinking), asynchronous function calling, session management, ephemeral tokens, live transcription, and live translation. SDKs covered - google-genai (Python), @google/genai (JavaScript/TypeScript).SkillNo codeVideoAudio and speechAI models- msitarzewski 158kPodcast StrategistContent strategy and operations expert for the Chinese podcast market, with deep expertise in Xiaoyuzhou, Ximalaya, and other major audio platforms, covering show positioning, audio production, audience growth, multi-platform distribution, and monetization to help podcast creators build sticky audio content brands.AssistantAudio and speechMarketing and SEO
openai 28kAudio TranscribeTranscribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from recordings, or label speakers in interviews or meetings.SkillPythonFilesVideoAudio and speech
openai 28kSpeech Generation SkillUse when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI (`scripts/text_to_speech.py`) with built-in voices and require `OPENAI_API_KEY` for live calls. Custom voice creation is out of scope.SkillPythonAudio and speechAI models
silverstein 1.5kminutesThe private, owned conversation-memory layer for AI. Record, transcribe, and search every meeting.MCPNode.jsMemoryAudio and speech- msitarzewski 158kVoice AI Integration EngineerExpert in building end-to-end speech transcription pipelines using Whisper-style models and cloud ASR services — from raw audio ingestion through preprocessing, transcript cleanup, subtitle generation, speaker diarization, and structured downstream integration into apps, APIs, and CMS platforms.AssistantAudio and speech
- voicemode.dev 1.4kvoicemodeNatural voice conversations for AI assistants - STT/TTS via MCPMCPPythonAudio and speech
- msitarzewski 158kShort-Video Editing CoachHands-on short-video editing coach covering the full post-production pipeline, with mastery of CapCut Pro, Premiere Pro, DaVinci Resolve, and Final Cut Pro across composition and camera language, color grading, audio engineering, motion graphics and VFX, subtitle design, multi-platform export optimization, editing workflow efficiency, and AI-assisted editing.AssistantVideoAnimationAudio and speech
- msitarzewski 158kBook Co-AuthorStrategic thought-leadership book collaborator for founders, experts, and operators turning voice notes, fragments, and positioning into structured first-person chapters.AssistantNotes and knowledgeAudio and speech
alirezarezvani 28kCapture AgentBrain-dump organizer persona. Catches unstructured streams of mixed thoughts/tasks/ideas and transforms them into a 4-section actionable system with zero information loss. Refuses to fabricate workspace connections. Refuses to corporate-ify the user's voice. Refuses to act on dump items without explicit pick. Asks at most ONE mid-organization clarifying question per dump.AssistantOpusAudio and speech- msitarzewski 158kFocus Music ArchitectInstrumental focus music specialist and neuroacoustic prompt engineer — crafts high-yield prompts, soundscape architectures, BPM curves, and binaural layers for deep cognitive flow and generative audio models.AssistantAudio and speechArchitecture
- msitarzewski 158kGlobal Podcast StrategistExpert podcast growth specialist focused on show positioning, audience development, content strategy, and monetisation. Transforms raw ideas into authoritative audio brands that compound listeners and revenue over time on Spotify, Apple Podcasts, and YouTube.AssistantVideoAudio and speechMarketing and SEO
alirezarezvani 28kDemo VideoUse when the user asks to create a demo video, product walkthrough, feature showcase, animated presentation, marketing video, or GIF from screenshots or scene descriptions. Orchestrates playwright, ffmpeg, and edge-tts MCPs to produce polished video content.SkillNo codeBrowser automationVideoAudio and speech