Popular Media & Files Skills
PDF, spreadsheets, slides, images, audio, and video workflows.
Media and file skills help agents work with the formats that carry day-to-day business information. They can extract content, make targeted edits, and create polished deliverables. Check local file access and external processing requirements before handling confidential material.
- Edit PDFs and documents
- Analyze spreadsheets
- Create slides and visuals
- Transform audio and video
Top Media & Files Skills
Ranked by their position in the current overall directory snapshot.
Humanizer Ru
Проверяет русскоязычный текст на следы машинной генерации и по явной просьбе пользователя переписывает его естественным языком. Отвечает на просьбы вида «очеловечь», «убери гпт-шность», «звучит как нейросеть», «проверь на ИИ», «убери штампы», «убери канцелярит», «сделай живым». Detects AI-generated Russian text and humanizes it on request. Не предназначен для текста не на русском, исходного кода, юридических документов и художественной прозы.
Apollo Company Research
Search companies, enrich profiles with firmographic data, find job postings, and track company news through Apollo. Covers company search, detailed enrichment (industry, size, funding, tech stack), hiring activity, and news monitoring for B2B sales intelligence.
Backlink Gap Analysis
Find and prioritize ethical backlink and digital-PR opportunities using SandBase DataForSEO backlink data. Use when asked to analyze backlink profiles, find referring-domain gaps, compare link competitors, or prepare a link-outreach target list.
Competitor Ad Research
Research competitor advertising strategies through Facebook Ad Library, Google SERP ads, and social media sponsored content. Covers creative analysis, messaging patterns, targeting signals, and campaign frequency for competitive marketing intelligence.
Content Ideation
Generate content ideas backed by real audience demand data from Reddit questions, YouTube search gaps, Twitter discussions, and web search trends. Identifies proven topics, underserved questions, and content formats that outperform in your niche.
Content Performance
Analyze content performance across YouTube, TikTok, Instagram, and Twitter to identify what works. Compares engagement rates, format effectiveness, posting cadence impact, and audience response patterns across platforms for content strategy optimization.
Content Translator
Translate text content between languages with high quality and natural phrasing. Covers document translation, content localization, and multi-language research support for teams working across language barriers.
Cve Lookup
Look up Common Vulnerabilities and Exposures (CVEs) with severity scores, affected software, exploitability status, and remediation guidance. Essential for security research, vulnerability management, and patch prioritization.
Document Parser
Parse and extract structured content from PDFs, Word documents, and other file formats. Converts documents into clean, machine-readable text for analysis, summarization, data extraction, and content processing workflows.
Domain Analyzer
Comprehensive domain analysis combining WHOIS ownership, DNS infrastructure, SSL security, SEO performance, tech stack, site structure, and backlink profile. One-stop domain intelligence for acquisition research, security assessment, and competitive analysis.
Github Profile Research
Research GitHub user profiles including repositories, contributions, languages, stars, and activity patterns. Useful for developer talent sourcing, open-source community research, and technical due diligence on engineering teams.
Google News Research
Search and monitor news articles across Google News through SandBase. Use when asked for news monitoring, current events research, news coverage analysis, or media tracking.
Prd
Generate high-quality Product Requirements Documents (PRDs) for software systems and AI-powered features. Includes executive summaries, user stories, technical specifications, and risk analysis.
Sales Enablement
When the user wants to create sales collateral, pitch decks, one-pagers, objection handling docs, or demo scripts. Also use when the user mentions 'sales deck,' 'pitch deck,' 'one-pager,' 'leave-behind,' 'objection handling,' 'deal-specific ROI analysis,' 'demo script,' 'talk track,' 'sales playbook,' 'proposal template,' 'buyer persona card,' 'help my sales team,' 'sales materials,' or 'what should I give my sales reps.' Use this for any document or asset that helps a sales team close deals. For competitor comparison pages and battle cards, use an available competitor-analysis capability. For marketing website copy, cold outreach, or offer design, use the corresponding capability when available.
Task Management
Simple task management using a shared TASKS.md file. Reference this when the user asks about their tasks, wants to add/complete tasks, or needs help tracking commitments.
Tweetclaw
OpenClaw guide for Twitter search, follower exports, monitoring, media, and approved X automation through Xquik. Not affiliated with X Corp.
Hand Drawn Explainer Video Nikola
制作、修改和验收中文手绘知识讲解视频,交付配音、字幕、时间轴、可编辑工程和真实 MP4。支持两条不可混淆的制作路线:同一画布持续落墨的逐笔故事动画,以及用 SVG/HTML/GSAP 编排流程卡片、关系图和知识图形的程序动画。逐笔路线可选择自然肤色 Q 版人物、小黑风格或其他手绘风格,并可采用单场景、多幕故事、左右双语义岛等画面结构。用户说“边讲边画”“一笔一笔画出来”“白板手绘”“先画左边再画右边”“知识讲解动画”或希望把文稿、SRT、人物故事做成手绘视频时使用。也支持只输出生图/图生视频提示词。不用于写实数字人或假装已经完成无法验证的成片。
Strategy Consulting Visualization
Use when turning any content into clear, professional visualizations - board slides, reports, proposals, research summaries, training materials, technical diagrams, infographics, process flows, timelines, benchmarks, waterfall charts, or data-backed visual specs for any audience.
Github Actions
Expert guidance on writing GitHub Actions workflows. Use when: (1) writing or editing any file under .github/workflows/, (2) choosing between an existing action and a shell script in CI, (3) picking action versions, (4) configuring workflow permissions or cloud authentication from CI, (5) reviewing a workflow in a PR.
Wendy Contributing
Expert guidance on contributing to WendyOS: Yocto builds, agent internals, E2E testing, and system architecture. Use when developers mention: (1) building WendyOS images, (2) meta-wendyos layers or bitbake, (3) wendy-agent development or internals, (4) containerd or nerdctl on WendyOS, (5) E2E tests for wendy-agent, (6) Yocto recipes or bbappend files, (7) mDNS/Avahi service configuration, (8) device identity or UUID generation.
Recursive Decomposition
Based on the Recursive Language Models (RLM) research by Zhang, Kraska, and Khattab (2025), this skill provides strategies for handling tasks that exceed comfortable context limits through programmatic decomposition and recursive self-invocation. Triggers on phrases like "analyze all files", "process this large document", "aggregate information from", "search across the codebase", or tasks involving 10+ files or 50k+ tokens.
Cypress Author
Creates, updates, and fixes Cypress tests (E2E/end-to-end and component tests). Use when the user asks to create tests, add tests, write tests, update tests, test this file/component, new spec, or fix a failing or flaky test. Apply even when the user does not say 'Cypress' (e.g. 'create tests for this file'). Prefer cypress-explain when the user only wants to explain or review tests without changing code.
Cypress Docs
Search and extract Cypress information from official documentation (docs.cypress.io, cypress.io); prefer LLM markdown under /llm/* and refuse unverified API or behavior claims.
Analyzing Brand
Analyzes a brand from its own website — its colours and fonts, the logo and imagery it uses, how it sounds, who it sells to, the claims it makes, and what it sells. Reads the site and its pictures with a vision model, and marks anything the site doesn't evidence as unconfirmed rather than guessing. Facts only — nothing is generated, planned or judged here. Use when a brand's guidelines, palette, typography, voice, tone or positioning need establishing, or when another skill needs the brand before it plans or generates.
Analyzing Own Ads
Audits the ads a company is running itself — what's live now, what it has already stopped, and which have run longest. Watches the video ads and reads the image ads with a vision model, then maps what the account already covers — the angles, hooks, formats and offers — what was dropped and how fast, and which long-runners have nothing newer beside them. Read from the public ad library, so it carries run lengths rather than spend or results. Research only — nothing is generated or changed here. Use when the user wants their own ads audited, wants to know what they are running, which of their ads is holding, what they already cover, or whether anything is going stale.
Cloning Video Ads
Clones an existing video ad — reads a supplied reference ad, keeps its structure, pacing, shots, camera and hook, and rebuilds it as a new video with the user's own product in place of the original's. Use when the user wants to copy, clone, recreate, or remake an ad or a competitor's video, or make an ad "like this one" with their product. Not for a product ad built from scratch with no reference video, not for a creator sharing their own take on a product and not for cloning image ads.
Identifying Competitors
Finds out who a brand competes with. Reads what the brand sells from its site, searches the web for the alternatives buyers compare it against, and proposes a shortlist with competitor names and websites. Use when the user asks who their competitors are, or to find or refresh their competitor list, or when another skill needs competitors named first.
Onboarding User
Onboards a new user — sets up their brand for the first time from their website. Analyzes the brand, finds its competitors, and saves both to the company profile that every later session reads from. Use when the user wants to onboard, get started, or set up their brand or company profile. Not for just analyzing a brand from its website — that is analyzing-brand.
Planning Campaigns
Decides what to make next — turns the brand, the product, your own running ads and competitor research into the concepts worth testing next. Each concept is a full description of one ad, the buyer it targets, the bet it tests and the evidence it rests on — complete enough for a producer to build from directly. Nothing is generated before the plan is approved — it asks which concepts to build, and only then hands each approved concept to the skill that builds it. Use when the user asks what to make, what to test next, what hooks or angles to try, for campaign concepts, or a creative plan — or wants the whole campaign run end to end, from research to finished ads. Not for a single ad with no plan behind it — that goes straight to the producer skill.
Researching Competitor Ads
Analyzes the ads competitors are running — the ones live now, the ones they have already stopped, and which have run longest. Watches the video ads and reads the image ads with a vision model, so every finding comes from the ad itself rather than its caption. Finds the patterns across the set — what is holding, what was dropped, and what nobody runs. Research only — nothing is generated or recommended here. Use when the user wants competitor ads researched, an ad teardown, a competitive or category analysis, to see what ads a named brand is running, or to find what is working and what nobody has tried yet.
Supercmo Setup
Diagnoses which capabilities are ready or blocked, sets up the keys, then sets up the brand from its website. Use when the user asks to "set me up," "setup," "help me get set up," "get started," "help me get started," "get me started," "what can you do," or to configure or finish setting up SuperCMO, or on a no_provider_configured error.
Sendmux Cli
Use the Sendmux command-line interface for terminal-driven Sendmux work. Use when the user wants install commands, profiles, key-scope preflight, --json output, colon-namespaced Sendmux commands, request body/path/query/header flags, or CLI examples for Management, Mailbox, or Sending API operations.
Compress Images
Compress images for web/SEO performance using cwebp. Use when optimizing images for faster page loads, reducing file sizes, or converting JPG/PNG to WebP format.
Transcribe Video
Generate subtitles (SRT/VTT) and plain text transcripts from video or audio files using AWS Transcribe. Use when creating captions, extracting spoken content, generating transcripts for notes, or making video content searchable.
X Post
Post to X (Twitter) from the command line. Text, images, and video.
Analyzing Products
Normalizes a product into generation-ready facts — from an e-commerce URL (a clean description plus curated product images) or from a photo alone (category, how it's used, its moving/opening parts, and key visual details). Use when a product URL or photo needs turning into inputs for image/video generation, or when another skill needs product facts before generating.
Generating Ad Videos
Makes a product ad video — a product showcase with no one on screen, or a story or lifestyle commercial where a presenter or actor plays a role in the brand's spot, with or without a voiceover. Use when the user wants a product ad, commercial, TV ad, brand film, or product-hero video. Not for a creator or customer sharing their own take on a product, like a review, unboxing, try-on or testimonial — that is generating-ugc-videos; not for a video with no product being sold — that is generating-videos.
Generating Ai Actors
Generates one photorealistic person as a single image — an AI actor, influencer, presenter or model. The image is reusable, so passing it into any later generation brings the same face, hair and build back and one person can hold across a whole set of images or videos. Use when the user wants an AI actor, or when something needs a person and no photo of one was supplied.
Generating Cartoon Videos
Generates a cartoon video for a product — a drawn, animated, anime, illustrated or painted spot that holds one art style and the same characters across every cut. A cartoon character presents, uses or sells the product, and the product itself is either kept exactly as photographed or drawn into the style. Use when the user wants a cartoon, animated, anime, illustrated or hand-drawn video, an animated ad, explainer or mascot spot, or any video whose look is drawn rather than filmed. Not for a photorealistic ad or a live actor or for animating a supplied photo.
Generating Image Ads
Turns a product photo or URL into a static image ad — one image, or a set. Reads the real product first and holds it identical across every frame and ratio. Triggers: "image ad", "make an ad", "static ad", "promo graphic", "offer graphic", "sale creative", "social ad", "banner ad", "ad creative", "before/after ad", "testimonial ad". Not for generating product photography.
Generating Storyboards
Generates the storyboard sheets for a video — one image per clip, panels in a row, each panel one action. Divides each clip into panels itself, then builds the sheets one at a time, each taking the previous one as a reference, so the person and product stay consistent except where the story deliberately changes them. Use when a video's clip lengths and story are settled and it needs sheets before anything is animated. Not for a still image on its own — that should be made using generating-images.
Generating Ugc Videos
Generates a UGC video — a creator or customer, on camera as themselves, sharing their own authentic take on a product. Use when the user wants a UGC video, creator video, influencer ad or endorsement, tiktok-style review, unboxing, try-on, OOTD, fit check, haul, tutorial, talking head, or creator testimonial. Not for a produced advertisement in the brand's voice, where a presenter or actor plays a role in the brand's spot — that is generating-ad-videos; not for a video with no person in it, or a person with no product to show.
Writing Video Prompts
Writes the motion prompt for a video clip — what moves, how the camera behaves, the physics and the audio, in the shape the chosen model wants. Works from a brief alone, from a start frame, from a source video, or from a storyboard sheet where each panel becomes a cut. Handles one clip or a whole set, returning one prompt per clip. Nothing is generated here. Use when a clip's model, duration and content are settled and it needs the prompt written.
Writing Video Scripts
Writes what is said in a video — the words a person speaks on camera, or a voiceover heard over the picture. Fits the writing to the time available, so lines are never rushed to fit or padded to fill. A video longer than a single clip gets its script split into one piece per clip, each sized to that clip's seconds. Outputs the words as text only — nothing is recorded or voiced here. Use when the user wants a script, dialogue, or a voiceover written for a video.
Nutrient Document Processing
Process documents with Nutrient DWS. Use when the user wants to generate PDFs from HTML or URLs, convert Office/images/PDFs, assemble or split packets, OCR scans, extract text/tables/key-value pairs, redact PII, watermark, sign, fill forms, optimize PDFs, or produce compliance outputs like PDF/A or PDF/UA. Triggers include convert to PDF, merge these PDFs, OCR this scan, extract tables, redact PII, sign this PDF, make this PDF/A, or linearize for web delivery.
Client Archetypes
Use this to match your tone to the buyer type. Six types. Same deal, different music.
Dark Tactics Defense
Use this to spot manipulation aimed at you. And to keep your own moves on the honest side.
Discovery Power
Use this to find out what someone wants, fears, or hides. Without interrogating them.
Related Guides
What Are Agent Skills?
A practical explanation of Skills, SKILL.md, and how they differ from MCP servers.
Read guide 8 min readHow to Install Agent Skills
Install from ClawHub, Git, or a local folder—and know what to review first.
Read guide 7 min readBest Agent Skills to Try First
A beginner-friendly path through useful, understandable skills across common workflows.
Read guideMedia Skills FAQ
What is a media agent skill?
It is a reusable instruction package that teaches an AI agent a focused media & files workflow, often including commands, checks, and supporting resources.
Which media skill should I try first?
Start with a narrow task you already understand. The current category leader is Ontology, but requirements and access scope matter more than rank alone.
Does a popular skill mean it is safe?
No. Popularity reflects adoption and interest, not a security guarantee. Read SKILL.md, review commands and dependencies, and test with minimal permissions.