Files
2026-08-14 19:10:57 +02:00

39 KiB
Raw Permalink Blame History

External AI agent skills — recording-studio-technician

Proven, publicly available AI agent skills mapped to this occupation. Nothing is copied from the sources: every entry is a name, a one-line summary and a link to the upstream skill package. Each section names its source repository, commit, license and retrieval date.

Tiers: core = the skill directly exercises a top market hard skill, tool or method (from gated job-ad evidence) or an essential ESCO competence of this occupation; adjacent = plausibly useful, secondary. Entries are capped at 12 per source and 80 in total per occupation (core first, strongest matches survive); everything beyond the caps is excluded and logged in the pipeline audit trail, not in this package.

Matched deterministically (ISCO group + title/competence keywords, tiered against market evidence + ESCO essentials) by pipeline/p5_enrich_ai_skills.py on 2026-07-14.

Source: anthropics/skills

  • Repository: https://github.com/anthropics/skills (commit f6656c1, retrieved 2026-07-14)
  • License: Apache-2.0; the document skills (docx/pdf/pptx/xlsx) are source-available — see the LICENSE.txt in the upstream skill folder
Skill Tier What it adds Upstream
docx adjacent Use this skill whenever the user wants to create, read, edit, or manipulate Word documents (.docx files) or Word templates (.dotx files). Triggers include: any mention of 'Word doc', 'word document', '.docx', '.dotx', or requests to … source
pdf adjacent Use this skill whenever the user wants to do anything with PDF files. This includes reading or extracting text/tables from PDFs, combining or merging multiple PDFs into one, splitting PDFs apart, rotating pages, adding watermarks, creating … source

Source: ConardLi/garden-skills

Skill Tier What it adds Upstream
web-video-presentation adjacent 把一篇文章或口播稿,做成"看起来像视频"的点击驱动 16:9 网页演示,可选合成口播音频。流程:原始文章 → 一次产出口播稿 + outline 开发计划 → 用户一次对齐 5 件事(稿子 / outline / 主题 / 素材 / 开发模式)→ 网页开发(逐章 / 顺序 / 并行)→ 可选音频合成provider-agnostic内置 MiniMax mmx-cli + OpenAI TTS可换 ElevenLabs / edge-tts / … source

Source: a5c-ai/babysitter

Skill Tier What it adds Upstream
procedural-audio core Procedural sound skill for synthesis and dynamic sound design. source
sound-design-direction core Create comprehensive audio design including music cues, sound effects, Foley, and score direction source
unity-profiler core Unity Profiler skill for performance analysis, frame debugging, memory profiling, and optimization workflows. source
multimedia-learning-design adjacent Apply Mayer's multimedia learning principles to design effective audio, video, graphics, and animations that reduce cognitive load source

Source: affaan-m/everything-claude-code

Skill Tier What it adds Upstream
google-workspace-ops core Operate across Google Drive, Docs, Sheets, and Slides as one workflow surface for plans, trackers, decks, and shared documents. Use when the user needs to find, summarize, edit, migrate, or clean up Google Workspace assets without dropping … source

Source: AgriciDaniel/claude-blog

Skill Tier What it adds Upstream
blog-audio adjacent Generate audio narration of blog posts using Google Gemini TTS. Supports summary narration, full article read-aloud, and two-speaker podcast/dialogue mode with 30 voice options. Outputs MP3 with HTML5 audio embed code. Works standalone via … source

Source: alirezarezvani/claude-skills

Skill Tier What it adds Upstream
google-workspace-cli core Google Workspace administration via the gws CLI (github.com/googleworkspace/cli). Install, authenticate, and automate Gmail, Drive, Sheets, Calendar, Docs, Chat, and Tasks. Run security audits and use local recipe templates and persona … source

Source: anbeime/skill

Skill Tier What it adds Upstream
video-frame-extractor core 视频反推工具,支持视频抽帧、视觉模型分析、提示词生成,适用于视频创作参考、内容提取、场景分析 source

Source: ayoubben18/ab-method

Skill Tier What it adds Upstream
handoff core Compact the current conversation (or a side-topic that surfaced mid-grill) into a handoff document another agent can pick up. Use when a tangent deserves its own task, or to summarize a session for a fresh agent. source

Source: bitwize-music-studio/claude-ai-music-skills

Skill Tier What it adds Upstream
mastering-engineer core Guides audio mastering for streaming platforms including loudness optimization and tonal balance. Use when the user has approved tracks and wants to master audio files. source
promo-director core Generates 15-second vertical promo videos for social media from mastered audio. Use after mastering is complete and before release, when the user wants social media content. source
sheet-music-publisher core Converts mastered audio to sheet music and creates printable songbooks. Use after mastering when the user wants sheet music or a songbook for their album. source
mix-engineer adjacent Polishes raw Suno audio by processing per-stem WAVs (vocals, backing_vocals, drums, bass, guitar, keyboard, strings, brass, woodwinds, percussion, synth, other) with targeted cleanup, EQ, and compression, then remixing into a polished … source

Source: davepoon/buildwithclaude

Skill Tier What it adds Upstream
tubeify core Remove pauses, filler words (um, uh), and dead air from raw YouTube recordings via the Tubeify API. Use when the user wants to edit a video, clean up audio, trim silences, or polish a raw recording for YouTube. source
routerbase-model-gateway core Integrate and route AI model requests through RouterBase. Use when migrating OpenAI-compatible clients, choosing model IDs, configuring fallbacks, or building chat, image, video, audio, speech, and embedding workflows. source

Source: davila7/claude-code-templates

Skill Tier What it adds Upstream
game-audio core Game audio principles. Sound design, music integration, adaptive audio systems. source
audiocraft-audio-generation adjacent PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descriptions, create sound effects, or perform melody-conditioned music generation. source

Source: Devin-AXIS/iPolloWork

Skill Tier What it adds Upstream
daytona-recording-artifacts core frame proof, HTML frames, screenshots, recording, PR proof, e2e evidence, validate visually. Daytona artifacts workflow for validated screenshots and optional videos. source

Source: EveryInc/compound-engineering-plugin

Skill Tier What it adds Upstream
ce-riffrec-feedback-analysis core Analyze Riffrec feedback captures from bundles or standalone recordings. Always load for riffrec-*.zip, session.json + events.json + recording.webm + voice.webm bundles, .mp4/.mov/.webm videos, .m4a/.mp3/.wav audio, … source

Source: foryourhealth111-pixel/Vibe-Skills

Skill Tier What it adds Upstream
hive-mind-advanced core Advanced Hive Mind collective intelligence system for queen-led multi-agent coordination with consensus mechanisms and persistent memory source

Source: googleworkspace/cli

Skill Tier What it adds Upstream
recipe-share-event-materials core Share Google Drive files with all attendees of a Google Calendar event. source

Source: gooseworks-ai/goose-skills

Skill Tier What it adds Upstream
video-polish core Takes an existing screen recording or demo video and adds professional zoom/pan effects synchronized to the narration. Uses transcript-driven zoom targeting and Remotion for rendering. Optionally replaces audio with a soundtrack. source
create-imessage-mockup core Render pixel-accurate iMessage screenshot mockups (DM or group) from a thread JSON. Supports minimal, with-keyboard, and full iPhone 15 Pro frame variants. Outputs HTML + PNG. source
render-value-prop core Render a designed 'value prop' video from a config — 3-5 noun-phrase benefit claims (<=4 words each) revealed sequentially over per-SKU product visuals, one crisp editorial frame per claim (hook sticker -> N claim beats -> brand end card). … source
review-ugc-render adjacent Mandatory pre-publish review gate for a UGC video render. Transcribes the finished render's AUDIO with Whisper and word-diffs it against the approved spoken script, then gates set_final_render — blocking a render whose generated audio … source
create-video-seedance-2-fal adjacent Generate a single 4-15s vertical video clip with ByteDance Seedance 2.0 reference-to-video via fal.ai. Multi-image reference (avatar + product + setting), native lip-synced VO + ambient audio via generate_audio: true, internal multi-cut … source

Source: infrasity-labs/dev-gtm-claude-skills

Skill Tier What it adds Upstream
blog-audio adjacent Generate audio narration of blog posts using Google Gemini TTS. Supports summary narration, full article read-aloud, and two-speaker podcast/dialogue mode with 30 voice options. Outputs MP3 with HTML5 audio embed code. Works standalone via … source

Source: jeremylongshore/claude-code-plugins-plus-skills

Skill Tier What it adds Upstream
speak-prod-checklist core Production readiness checklist for Speak language learning integrations: auth, audio pipeline, monitoring, and compliance. Use when implementing prod checklist features, or troubleshooting Speak language learning integration issues. … source
deepgram-core-workflow-a core Implement production pre-recorded speech-to-text with Deepgram. Use when building audio transcription, batch processing, or implementing diarization and intelligence features. Trigger: "deepgram transcription", "speech to text", … source
speak-sdk-patterns core Production patterns for Speak language learning API: conversation sessions, pronunciation assessment, audio preprocessing, and batch operations. Use when implementing sdk patterns features, or troubleshooting Speak language learning … source
elevenlabs-sdk-patterns core Apply production-ready ElevenLabs SDK patterns for TypeScript and Python. Use when implementing ElevenLabs integrations, refactoring SDK usage, or establishing team coding standards for audio AI applications. Trigger: "elevenlabs SDK … source
elevenlabs-reference-architecture core Implement ElevenLabs reference architecture for production TTS/voice applications. Use when designing new ElevenLabs integrations, reviewing project structure, or building a scalable audio generation service. Trigger: "elevenlabs … source
granola-install-auth core Install and configure Granola AI meeting notes with calendar and audio permissions. Use when setting up Granola for the first time, connecting Google/Outlook calendars, granting macOS Screen Recording permission, or configuring Windows … source
speak-security-basics core Security best practices for Speak API keys, audio data privacy, student data protection, and COPPA/FERPA compliance. Use when implementing security basics features, or troubleshooting Speak language learning integration issues. Trigger … source
abridge-core-workflow-a core Implement Abridge ambient clinical documentation capture-to-note pipeline. Use when building the primary encounter workflow: audio capture, real-time transcription, AI note generation, and EHR note insertion. Trigger: "abridge clinical … source
assemblyai-core-workflow-b core Execute AssemblyAI streaming transcription and LeMUR workflows. Use when implementing real-time speech-to-text, live captions, voice agents, or LLM-powered audio analysis with LeMUR. Trigger with phrases like "assemblyai streaming", … source
elevenlabs-core-workflow-a core Implement ElevenLabs text-to-speech and voice cloning workflows. Use when building TTS features, cloning voices from audio samples, or implementing the primary ElevenLabs money-path: voice generation. Trigger: "elevenlabs TTS", "text to … source
groq-core-workflow-b core Execute Groq secondary workflows: audio transcription (Whisper), vision, text-to-speech, and batch model evaluation. Trigger with phrases like "groq whisper", "groq transcription", "groq audio", "groq vision", "groq TTS", "groq speech". source
speak-common-errors core Diagnose and fix common Speak API errors: authentication failures, audio format issues, rate limits, and session management problems. Use when implementing common errors features, or troubleshooting Speak language learning integration … source

Source: JimLiu/baoyu-skills

Skill Tier What it adds Upstream
baoyu-post-to-x core Posts content and articles to X (Twitter). Supports regular posts with images/videos and X Articles (long-form Markdown). In Codex, honor explicit requests for the Codex Chrome plugin/@chrome by using the Chrome Extension workflow; … source

Source: K-Dense-AI/claude-scientific-skills

Skill Tier What it adds Upstream
open-notebook adjacent Self-hosted, open-source alternative to Google NotebookLM for AI-powered research and document analysis. Use when organizing research materials into notebooks, ingesting diverse content sources (PDFs, videos, audio, web pages, Office … source

Source: K-Dense-AI/scientific-agent-skills

Skill Tier What it adds Upstream
open-notebook adjacent Self-hosted, open-source alternative to Google NotebookLM for AI-powered research and document analysis. Use when organizing research materials into notebooks, ingesting diverse content sources (PDFs, videos, audio, web pages, Office … source

Source: microsoft/skills

Skill Tier What it adds Upstream
azure-ai-contentunderstanding-py adjacent Azure AI Content Understanding SDK for Python. Use for multimodal content extraction from documents, images, audio, and video. Triggers: "azure-ai-contentunderstanding", "ContentUnderstandingClient", "multimodal analysis", "document … source

Source: mohitagw15856/pm-claude-skills

Skill Tier What it adds Upstream
deck-autopsy core Autopsy a slide deck from photos or screenshots of its slides — the narrative arc, the numbers, and what each slide is hiding. Use when given slide images (a competitor's pitch, a conference talk, your own deck before a big meeting) and … source
youtube-script-writer adjacent Write engaging, high-retention YouTube video scripts with visual and audio cues. Use when asked to write a YouTube script, design a video outline, draft a video hook, or structure a video narrative. Produces a polished script with multiple … source

Source: nexu-io/html-anything

Skill Tier What it adds Upstream
frame-macos-notification core 拟真 macOS 通知 banner + app icon + 标题正文, 适合 video overlay / 产品发布预告 source

Source: nexu-io/open-design

Skill Tier What it adds Upstream
docx core Create, edit, and analyze Word documents with tracked changes, comments, and formatting. Useful for design briefs, copy docs, and review-ready deliverables. source
motion-frames core A single-frame motion-design composition with looping CSS animations — rotating type ring, animated globe, ticking timer, parallax labels. Renders as a hero video poster you can hand straight to HyperFrames or any keyframe-based exporter. … source
video-template-frame-product-promo core Use this plugin when the user wants a "Product Promo" HyperFrames motion video — Multi-scene product showcase with SVG assets source
video-template-frame-product-promo-30s core Use this plugin when the user wants a "Product Promo · 30s" HyperFrames motion video — Multi-scene 30-second product promo: problem-type intro, brand reveal, benefits flowchart, product surfaces, value pillars, foundation, CTA outro. … source
frame-glitch-title core Digital glitch, chromatic offset, and data-corruption title frame for video transitions or cyberpunk heroes. source
frame-liquid-bg-hero core WebGL-style fluid displacement background with a quote overlay, suited to video intros, landing heroes, or posters. source
frame-logo-outro core Segmented logo assembly, glow bloom, and tagline reveal for video outros or brand closing frames. source
frame-macos-notification core Realistic macOS notification banner with app icon, title, and body, suited to video overlays or product teasers. source
video-hyperframes core Hyperframes / Remotion-compatible continuous frame animation with autoplay support. source
video-template-frame-bold-poster core Use this plugin when the user wants a "Bold Poster Frame" HyperFrames motion video — A 1970s European editorial poster in motion — a red rule draws across, a giant tilted figure drops in, a three-line headline rises line-by-line, an italic … source
video-template-frame-bold-signal core Use this plugin when the user wants a "Bold Signal Frame" HyperFrames motion video — Bold colored card on a dark gradient — big section number, nav breadcrumb, orange card sliding in, title rising. source
video-template-frame-build-minimal core Use this plugin when the user wants a "Build Minimal Frame" HyperFrames motion video — Luxury-minimal whitespace hero — single word reveals letter by letter, warm-gold hairline, breathing indicators. source

Source: NoizAI/skills

Skill Tier What it adds Upstream
sound-fx adjacent Use this skill whenever the user wants to generate sound effects, ambient audio, or short audio clips from a text description. Triggers include: any mention of 'sound effect', 'sfx', 'generate sound', 'make a sound', 'audio effect', … source
chat-with-anyone adjacent Chat with any real person or fictional character in their own voice by automatically finding their speech online, extracting a clean reference sample, and generating audio replies. Also supports generating a matching voice from an uploaded … source

Source: Orchestra-Research/AI-research-SKILLs

Skill Tier What it adds Upstream
audiocraft-audio-generation adjacent PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descriptions, create sound effects, or perform melody-conditioned music generation. source

Source: Orkas-AI/Orkas-VideoStudio

Skill Tier What it adds Upstream
composition-design-review core Design review layer for OrkasVideoStudio COMPOSE previews and drafts. Review every immutable snapshot frame before showing a visual preview; when preview is skipped, use the draft as the fallback evidence. Return one complete, actionable … source
video-craft adjacent The craft standard that makes a video GOOD, not just rendered — hooks, story pacing, visual hierarchy, type & safe zones, motion/easing, captions, audio mix, platform conventions, shot language, generation-prompt writing, and a pre-publish … source

Source: ruvnet/claude-code-flow

Skill Tier What it adds Upstream
hive-mind-advanced core Advanced Hive Mind collective intelligence system for queen-led multi-agent coordination with consensus mechanisms and persistent memory source

Source: ruvnet/ruflo

Skill Tier What it adds Upstream
hive-mind-advanced core Advanced Hive Mind collective intelligence system for queen-led multi-agent coordination with consensus mechanisms and persistent memory source

Source: SamurAIGPT/Generative-Media-Skills

Skill Tier What it adds Upstream
muapi-seedance-2 core Expert Cinema Director skill for Seedance 2.0 (ByteDance) — high-fidelity video generation across Chinese, Global, and VIP tiers. Supports text-to-video, image-to-video, first-last-frame, omni reference, character training, omni-reference … source
muapi-ugc-video-factory adjacent Turn a person photo + a product photo + an optional script into a vertical 9:16 UGC-style video ad. Generates a lifestyle hero image (Nano-Banana Pro Edit), then animates it with native audio using Seedance 2.0 VIP image-to-video. source

Source: sanjay3290/ai-skills

Skill Tier What it adds Upstream
google-tts core Convert documents and text to audio using Google Cloud Text-to-Speech. Use this skill when the user wants to: narrate a document, read aloud text, generate audio from a file, convert text to speech, create a recording of documentation or … source

Source: silverstein/minutes

Skill Tier What it adds Upstream
minutes-mirror core Self-coaching analysis of your own behavior across meetings — talk-time ratio, filler words, hedging language, monologue length, energy patterns, and (when meetings are tagged via /minutes-tag) what your behavior in winning meetings looks … source

Source: Vincentwei1021/video-shotcraft

Skill Tier What it adds Upstream
video-shotcraft core Create cinematic product videos from shot recipe cards, a validated template, and code/audio assets (Remotion + real page screenshots + 2.5D camera moves + beat-synced cuts + sound design). Use when the user asks to turn a frontend project … source

Source: zechenzhangAGI/AI-research-SKILLs

Skill Tier What it adds Upstream
audiocraft-audio-generation adjacent PyTorch library for audio generation including text-to-music (MusicGen) and text-to-sound (AudioGen). Use when you need to generate music from text descriptions, create sound effects, or perform melody-conditioned music generation. source

Source: google/skills

Skill Tier What it adds Upstream
ima-sdk-basics adjacent Use this skill for Interactive Media Ads (IMA) SDK client-side ad insertion when you are requesting video ads client-side into websites, apps, TVs or other platforms with VAST or VMAP. Do not use for Dynamic Ad Insertion (DAI), SSAI, or … source

Source: veniceai/skills

Skill Tier What it adds Upstream
venice-audio-speech core Generate speech from text via POST /audio/speech. Covers TTS models (Kokoro, Qwen 3, xAI, Inworld, Chatterbox, Orpheus, ElevenLabs Turbo, MiniMax, Gemini Flash), voices per family, output formats (mp3/opus/aac/flac/wav/pcm), streaming, … source
venice-audio-transcription core Transcribe audio files to text via POST /audio/transcriptions. Covers supported models (Parakeet, Whisper, Wizper, Scribe, xAI STT), supported formats (wav/flac/m4a/aac/mp4/mp3/ogg/webm), response formats (json/text), timestamps, and … source
venice-chat core Call POST /chat/completions on Venice. Covers the OpenAI-compatible request shape, Venice-only venice_parameters (web search, E2EE, characters, thinking control, X search), multimodal inputs (images/audio/video), tool calls, reasoning … source
venice-audio-music adjacent Async music / audio-track generation via Venice. Covers the /audio/quote + /audio/queue + /audio/retrieve + /audio/complete lifecycle, lyrics vs instrumental, voice selection, duration, language, speed, model capability probing, and … source
venice-video adjacent Generate and transcribe videos via Venice. Covers the async /video/quote + /video/queue + /video/retrieve + /video/complete loop, text-to-video, image-to-video, video-to-video (upscale), audio input, reference images, scene and element … source