harness · market
Harness
Put raw intelligence to work.
Skills, MCP servers, prompts, assistants and connectors. Every listing shows its author, version and the permissions it needs.
On the shelves20,989
- MCP servers18,838
- Skills1,282
- Assistants651
- Prompts218
0 items
AI models
yamadashy 29kRepomixPack local or remote codebases into AI-friendly files that LLMs and coding agents can read or searchMCPNode.jsFilesAI models
muratcankoylan 18kAdvanced EvaluationThis skill should be used for advanced LLM evaluation: LLM-as-judge systems, direct scoring, pairwise comparison, rubric calibration, evaluator bias mitigation, confidence scoring, and automated quality assessment.SkillPythonAI models
anthropics 180kMCP Server Development GuideGuide for creating high-quality MCP (Model Context Protocol) servers that enable LLMs to interact with external services through well-designed tools. Use when building MCP servers to integrate external APIs or services, whether in Python (FastMCP) or Node/TypeScript (MCP SDK).SkillPythonAI models
orchestra-research 13kWhisper - Robust Speech RecognitionOpenAI's general-purpose speech recognition model. Supports 99 languages, transcription, translation to English, and language identification. Six model sizes from tiny (39M params) to large (1550M params). Use for speech-to-text, podcast transcription, or multilingual audio processing. Best for robust, multilingual ASR.SkillNo codeAudio and speechTranslationAI models
huggingface 11kGradioBuild Gradio web UIs and demos in Python. Use when creating or editing Gradio apps, components, event listeners, layouts, or chatbots.SkillNo codeWritingAI models
droidrun 9.6kmobilerunControl real Android and iOS devices with LLM agents — tap, swipe, type, automate flows.MCPRemoteMobileAI models
orchestra-research 13kLong Context: Extending Transformer Context WindowsExtend context windows of transformer models using RoPE, YaRN, ALiBi, and position interpolation techniques. Use when processing long documents (32k-128k+ tokens), extending pre-trained models beyond original context limits, or implementing efficient positional encodings. Covers rotary embeddings, attention biases, interpolation methods, and extrapolation strategies for LLMs.SkillNo codeDocumentsAI models
google-gemini 4.3kGemini Live API Development SkillUse this skill when building real-time, bidirectional streaming applications with the Gemini Live API, or migrating legacy Live models (2.0/2.5/3.1) to Gemini 3.8 Live. Covers WebSocket-based audio/video/text streaming, voice activity detection (VAD), background reasoning (extended thinking), asynchronous function calling, session management, ephemeral tokens, live transcription, and live translation. SDKs covered - google-genai (Python), @google/genai (JavaScript/TypeScript).SkillNo codeVideoAudio and speechAI models
huggingface 11kHugging Face Dataset ViewerUse this skill for Hugging Face Dataset Viewer API workflows that fetch subset/split metadata, paginate rows, search text, apply filters, download parquet URLs, and read size or statistics.SkillNo codeAI models
huggingface 11kOverviewRun evaluations for Hugging Face Hub models using inspect-ai and lighteval on local hardware. Use for backend selection, local GPU evals, and choosing between vLLM / Transformers / accelerate. Not for HF Jobs orchestration, model-card PRs, .eval_results publication, or community-evals automation.SkillPythonAI models- msitarzewski 158kAI Data Remediation EngineerSpecialist in self-healing data pipelines — uses air-gapped local SLMs and semantic clustering to automatically detect, classify, and fix data anomalies at scale. Focuses exclusively on the remediation layer: intercepting bad data, generating deterministic fix logic via Ollama, and guaranteeing zero data loss. Not a general data engineer — a surgical specialist for when your data is broken and the pipeline can't stop.AssistantAI models
wshobson 40kLLM finetuning architectFine-tuning strategist who owns the eval gate and method/model selection. Refuses to plan training without a baselined eval harness.AssistantOpusArchitectureAI models
alirezarezvani 28kKarpathy Coder — Active Coding DisciplineUse when writing, reviewing, or committing code to enforce Karpathy's 4 coding principles — surface assumptions before coding, keep it simple, make surgical changes, define verifiable goals. Triggers on "review my diff", "check complexity", "am I overcomplicating this", "karpathy check", "before I commit", or any code quality concern where the LLM might be overcoding.SkillPythonWritingAI models
huggingface 11kOverviewPublish and manage research papers on Hugging Face Hub. Supports creating paper pages, linking papers to models/datasets, claiming authorship, and generating professional markdown-based research articles.SkillPythonDocumentsResearchAI models
orchestra-research 13kvLLM - High-Performance LLM ServingServes LLMs with high throughput using vLLM's PagedAttention and continuous batching. Use when deploying production LLM APIs, optimizing inference latency/throughput, or serving models with limited GPU memory. Supports OpenAI-compatible endpoints, quantization (GPTQ/AWQ/FP8), and tensor parallelism.SkillNo codeMemoryAI models
huangchihhungleo 2.2kclaude-real-videoLet any LLM watch a video locally — and search everything it has ever watched.MCPPythonVideoAI models
google-gemini 4.3kGemini Omni Flash SkillUse this skill for generative video editing, text-to-video, image-referenced video generation, first-frame-to-video, first-and-last-frame transitions, and video extensions using Gemini Omni 1.1 Flash (gemini-omni-1.1-flash) via the official google-genai SDK. Includes workflows for pre-processing/optimizing high-resolution or long source videos with ffmpeg, stripping audio for full sound regeneration, and handling turn-by-turn video editing and parallel execution.SkillPythonImagesVideoAI models
microsoft.com 1.9kMicrosoft Learn MCPOfficial Microsoft Learn MCP Server – real-time, trusted docs & code samples for AI and LLMs.MCPRemoteAI models
orchestra-research 13kAxolotl SkillExpert guidance for fine-tuning LLMs with Axolotl - YAML configs, 100+ models, LoRA/QLoRA, DPO/KTO/ORPO/GRPO, multimodal supportSkillNo codeAI models
openai 28kSpeech Generation SkillUse when the user asks for text-to-speech narration or voiceover, accessibility reads, audio prompts, or batch speech generation via the OpenAI Audio API; run the bundled CLI (`scripts/text_to_speech.py`) with built-in voices and require `OPENAI_API_KEY` for live calls. Custom voice creation is out of scope.SkillPythonAudio and speechAI models
alirezarezvani 28kSpinning Up in Deep RLAnswers from the knowledge base compiled from Spinning Up in Deep RL by Joshua Achiam (OpenAI). Loads the master frameworks first and reads a single chapter file on demand rather than the whole source. Refuses to answer beyond what the source covers.AssistantOpusNotes and knowledgeAI models
huggingface 11kTrackio - Experiment Tracking for ML TrainingTrack and visualize ML training experiments with Trackio. Use when logging metrics during training (Python API), firing alerts for training diagnostics, or retrieving/analyzing logged metrics (CLI). Supports real-time dashboard visualization, alerts with webhooks, HF Space syncing, and JSON output for automation.SkillNo codeMonitoringAI models
alirezarezvani 28kRAG ArchitectUse when the user asks to design a RAG pipeline, choose a chunking strategy or embedding model, pick a vector database, or evaluate retrieval quality (precision@k, recall@k, NDCG). Examples: 'design a RAG system for our docs', 'what chunk size should I use for this corpus', 'evaluate my retriever against ground truth'. NOT for general LLM cost tuning (use llm-cost-optimizer) or agent loops over retrieval (use agenthub).SkillPythonDatabasesAI models
orchestra-research 13kSentence Transformers - State-of-the-Art EmbeddingsFramework for state-of-the-art sentence, text, and image embeddings. Provides 5000+ pre-trained models for semantic similarity, clustering, and retrieval. Supports multilingual, domain-specific, and multimodal models. Use for generating embeddings for RAG, semantic search, or similarity tasks. Best for production embedding generation.SkillNo codeImagesAI models