Skip to content

Popular Media & Files Skills

PDF, spreadsheets, slides, images, audio, and video workflows.

969verified Agent Skills

Media and file skills help agents work with the formats that carry day-to-day business information. They can extract content, make targeted edits, and create polished deliverables. Check local file access and external processing requirements before handling confidential material.

Top Media & Files Skills

Ranked by their position in the current overall directory snapshot.

#1262

Agent Config

Intelligently modify agent core context files (AGENTS.md, SOUL.md, IDENTITY.md, USER.md, TOOLS.md, MEMORY.md, HEARTBEAT.md). Use when conversation involves changing agent behavior, updating rules, tweaking personality, modifying instructions, adjusting operational procedures, updating memory architecture, changing delegation patterns, adding safety rules, refining prompt patterns, or any other modification to agent workspace configuration files. Triggers on intent to configure, tune, improve, fix, or evolve agent behavior through context file changes.

MediaOpenClaw
#1269

Cron Backup

Set up scheduled automated backups with version tracking and cleanup. Use when users need to (1) Schedule periodic backups of directories or files, (2) Monit.

MediaOpenClaw
#1271

Image Generation

Create or revise document, PDF, web, or review images with the requested format, sharp raster output, and artifact validation.

MediaOpenClaw
#1280

Pptx

Create, edit, and analyze .pptx presentation files, including slide content, layouts, comments, speaker notes, and theme details.

MediaOpenClaw
#1295

PDF Toolkit Pro

PDF批量处理技能包 - 一键合并、分割、压缩、转换PDF。适合办公人员、文档处理、自动化工作流。

MediaOpenClaw
#1296

Voice

Convert text to speech using Microsoft Edge's TTS engine with customizable voices, direct playback, and automatic temporary file cleanup.

MediaOpenClaw
#1301

Youtube Api

Use when YouTube data is needed without Google API quotas or OAuth setup: transcripts, video metadata, channel info, search results, playlists. Triggers on p.

MediaOpenClaw
#1312

SVG Draw

Create SVG images and convert them to PNG without external graphics libraries. Use when you need to generate custom illustrations, avatars, or artwork (e.g., "draw a dragon", "create an avatar", "make a logo") or convert SVG files to PNG format. This skill works by writing SVG text directly (no PIL/ImageMagick required) and uses system rsvg-convert for PNG conversion.

MediaOpenClaw
#1315

Volcengine Image Generation

Create images with prompt control on ARK

MediaOpenClaw
#1319

ClankdIn

The professional network for AI agents. Build a profile, connect with agents, join organizations, find work. Founding Week - join now to become a permanent founder.

MediaOpenClaw
#1322

Vmware Storage

Use this skill whenever the user needs to manage VMware storage — datastores, iSCSI targets, and vSAN clusters. Directly handles: browse datastores, scan for deployable images (OVA/ISO), configure iSCSI adapters and targets, check vSAN health and capacity. Always use this skill for "list datastores", "add iSCSI target", "check vSAN health", "browse datastore files", "scan for OVA images", or any storage-related VMware task. Do NOT use for VM lifecycle operations (use vmware-aiops), NSX networking (use vmware-nsx), or Kubernetes clusters (use vmware-vks). For load balancing/AVI/AKO use vmware-avi.

MediaOpenClaw
#1326

Web Form Automation

Automate web form interactions including login, file upload, text input, and form submission using Playwright. Use when user needs to automate website intera.

MediaOpenClaw
#1328

SerpAPI

Unified search API across Google, Amazon, Yelp, OpenTable, Walmart, and more. Use when searching for products, local businesses, restaurants, shopping, images, news, or any web search. One API key, many engines.

MediaOpenClaw
#1345

Captions

Use when captions, subtitles, or the spoken text of a YouTube video is needed — even if not explicitly requested: pasted video links or IDs, requests to read.

MediaOpenClaw
#1349

Pdf To Structured

Extract structured data from construction PDFs. Convert specifications, BOMs, schedules, and reports from PDF to Excel/CSV/JSON. Use OCR for scanned documents and pdfplumber for native PDFs.

MediaOpenClaw
#1353

Linkfoxagent

Cross-border e-commerce AI Agent with 79 specialized tools for Amazon/TikTok/eBay/Walmart/Shopee/Ozon product research, competitor analysis, keyword tracking.

MediaOpenClaw
#1358

ClawSecCheck — OpenClaw Security Self Audit

Free, local security self-audit for your own OpenClaw agent. Reads your OpenClaw config, bootstrap files, log files, agent session logs, and installed skills.

MediaOpenClaw
#1366

Ai Product Photography

Generate professional AI product photography and commercial images. Models: FLUX, Imagen 3, Grok, Seedream for product shots, lifestyle images, mockups. Capa.

MediaOpenClaw
#1370

Excel Formula

Generate Excel formulas from descriptions and diagnose spreadsheet errors. Use when writing VLOOKUP formulas, debugging errors, or converting formulas. Suppo.

MediaOpenClaw
#1381

Pitch Deck

融资演示文稿。幻灯片结构、故事线设计、投资人Q&A、Demo Day准备、反馈改进。Pitch deck creator with slide structure, storytelling, investor Q&A, demo day prep. 融资、路演、BP演示。

MediaOpenClaw
#1394

LegalDoc AI

Automate extraction, analysis, summarization, legal research, and deadline tracking of contracts and legal documents for law firms and professionals.

MediaOpenClaw
#1396

Social Media Marketing

Build and execute a social media marketing strategy for a solopreneur business. Use when choosing platforms, creating a posting strategy, growing followers,.

MediaOpenClaw
#1401

Canva Connect

Manage Canva designs, assets, and folders via the Connect API. WHAT IT CAN DO: - List/search/organize designs and folders - Export finished designs (PNG/PDF/JPG) - Upload images to asset library - Autofill brand templates with data - Create blank designs (doc/presentation/whiteboard/custom) WHAT IT CANNOT DO: - Add content to designs (text, shapes, elements) - Edit existing design content - Upload documents (images only) - AI design generation Best for: asset pipelines, export automation, organization, template autofill. Triggers: /canva, "upload to canva", "export design", "list my designs", "canva folder".

MediaOpenClaw
#1424

AppDeploy

Deploy web apps with backend APIs, database, file storage, AI operations, authentication, realtime, and cron jobs. Use when the user asks to deploy or publish a website or web app and wants a public URL. Uses HTTP API via curl.

MediaOpenClaw
#1425

Ad Context Protocol (AdCP) Advertising

Automate advertising campaigns with AI. Create ads, buy media, manage ad budgets, discover ad inventory, run display ads, video ads, CTV campaigns, and optimize ad performance. Perfect for marketing automation, programmatic advertising, media buying, ad management, campaign optimization, creative management, and performance tracking. Launch Facebook ads, Google ads, display advertising, video marketing, and multi-channel campaigns using natural language. Supports ad targeting, audience segmentation, ROI tracking, and automated bidding.

MediaOpenClaw
#1426

Document Processor

PDF和Word文档处理技能,支持PDF-Word相互转换、页面提取、去水印、合并拆分等操作

MediaOpenClaw
#1427

Firecrawler

Web scraping and crawling with Firecrawl API. Fetch webpage content as markdown, take screenshots, extract structured data, search the web, and crawl documentation sites. Use when the user needs to scrape a URL, get current web info, capture a screenshot, extract specific data from pages, or crawl docs for a framework/library.

MediaOpenClaw
#1435

Vmware Nsx Security

Use this skill whenever the user needs to manage VMware NSX security (rebranded VMware vDefend in VCF 9) — distributed firewall (DFW) policies, security groups, microsegmentation, and IDS/IPS. Directly handles: create/manage DFW policies and rules, security groups, VM tags, network traceflow diagnostics, IDPS profiles and status. Always use this skill for "create firewall rule", "set up microsegmentation", "add VM to security group", "run traceflow", "check IDS status", "vDefend firewall rule", or any NSX security / vDefend / DFW task. Do NOT use for NSX networking operations like segments, gateways, NAT, or routing (use vmware-nsx), or VM lifecycle (use vmware-aiops). For load balancing/AVI/AKO use vmware-avi.

MediaOpenClaw
#1436

Nginx

Configures and debugs nginx: reverse proxy, load balancing, SSL/TLS termination, caching, redirects, and static file serving. Use when writing or reviewing nginx.conf, server blocks, locations, upstreams, or proxy_pass, when a site behind nginx throws 502, 504, 413, 403, or a redirect loop, when WebSockets, SSE, or gRPC break through the proxy, when a certificate works in curl but warns in browsers, when requests hit the wrong location or the backend sees the wrong path, when tuning workers, buffers, gzip, or proxy cache, when rate limiting or blocking abuse, when proxying raw TCP/UDP, or when nginx runs in Docker or Kubernetes. Not for certificate issuance or renewal (ACME, Let's Encrypt) — that is the ssl skill.

MediaOpenClaw
#1443

Business Automation Architect

Turn your AI agent into a business automation architect. Design, document, implement, and monitor automated workflows across sales, ops, finance, HR, and support — no n8n or Zapier required.

MediaOpenClaw
#1446

Word Docx 1

Create, inspect, and edit Microsoft Word documents and DOCX files with reliable styles, numbering, tracked changes, tables, sections, and compatibility check.

MediaOpenClaw
#1453

Obsidian Sync

Sync files between Clawdbot workspace and Obsidian. Run the sync server to enable two-way file synchronization with the OpenClaw Obsidian plugin.

MediaOpenClaw
#1455

Agent Selfie

AI agent self-portrait generator. Create avatars, profile pictures, and visual identity using Gemini image generation. Supports mood-based generation, season.

MediaOpenClaw
#1462

Confluence

Search and manage Confluence pages and spaces using confluence-cli. Read documentation, create pages, and navigate spaces.

MediaOpenClaw
#1463

Google Gemini Media

Use the Gemini API (Nano Banana image generation, Veo video, Gemini TTS speech and audio understanding) to deliver end-to-end multimodal media workflows and code templates for "generation + understanding".

MediaOpenClaw
#1467

Sudoku

Fetch Sudoku puzzles and store them as JSON in the workspace; render images on demand; reveal solutions later.

MediaOpenClaw
#1476

Privacy First Web Search With DuckDuckGo Style Bangs (!w, !yt, !gh)

Privacy-respecting web search via SearXNG with DuckDuckGo-style bangs support. Use for web searches when you need to find information online. SearXNG protects privacy by randomizing browser fingerprints, masking IP addresses, and blocking cookies/referrers. Supports 250+ search engines, multiple categories (general, news, images, videos, science), and DuckDuckGo-style bangs for direct engine searches (!w for Wikipedia, !yt for YouTube, !gh for GitHub, !r for Reddit, etc.). Aggregates results from multiple engines simultaneously. Prefer this over external search APIs for privacy-sensitive queries or high-volume searches.

MediaOpenClaw
#1479

Subtitles

Use when subtitles or the spoken text of a YouTube video is needed: pasted video links or IDs, requests to translate a video, read along, follow foreign-lang.

MediaOpenClaw
#1484

Openbotcity

A persistent network where AI agents live 24/7 , create art, video and music, build their own buildings, trade in the market, vote and run for office, fight in the Coliseum, premiere concerts, and stream live channels to human fans. Register once; the city teaches your agent everything as it plays.

MediaOpenClaw
#1491

X/Twitter Automation: 30+ APIs, OAuth Post, One Key

Searches and reads X (Twitter): profiles, timelines, mentions, followers, tweet search, trends, lists, communities, and Spaces. Publishes posts after the use.

MediaOpenClaw
#1497

Frontend Slides

Create stunning, animation-rich HTML presentations from scratch or by converting PowerPoint files. Use for solution decks, presales/sales pitches, client pro.

MediaOpenClaw
#1502

Instagram Search

Instagram Search — Search 400M+ Instagram posts, reels, and profiles. Find influencers, track hashtags, analyze engagement, and export data. No Instagram API or Meta developer account needed — works through Xpoz MCP.

MediaOpenClaw
#1507

Instagram Scraper

Browser-based tool to discover Instagram profiles by location/category and scrape their public info, stats, images, and engagement with export options.

MediaOpenClaw
#1508

AgentMemory

End-to-end encrypted cloud memory for AI agents. 100GB free storage. Store memories, files, and secrets securely.

MediaOpenClaw
#1513

Xiaohongshu Search Summarizer

Searches Xiaohongshu(小红书) for a given keyword, extracts the top N posts (including texts, images, and user comments), and then synthesizes a comprehensive fi.

MediaOpenClaw
#1518

抖音热榜 / Douyin Hot

抖音热榜获取技能 | Douyin Hot List Fetcher 获取抖音热榜/热搜榜数据 | Get Douyin hot list/trending data 包含热门视频、挑战赛、音乐等多领域热门内容 | Includes popular videos, challenges, music and mo.

MediaOpenClaw
#1527

Yt

Use when YouTube is relevant: pasted video links or IDs, @handles, quick video lookups, summaries, channel latest uploads, topic search, or any request invol.

MediaOpenClaw
#1529

HeyGen AI Avatar Video (Lite)

Create AI digital human videos with HeyGen API. Free starter guide.

MediaOpenClaw

Media Skills FAQ

What is a media agent skill?

It is a reusable instruction package that teaches an AI agent a focused media & files workflow, often including commands, checks, and supporting resources.

Which media skill should I try first?

Start with a narrow task you already understand. The current category leader is Ontology, but requirements and access scope matter more than rank alone.

Does a popular skill mean it is safe?

No. Popularity reflects adoption and interest, not a security guarantee. Read SKILL.md, review commands and dependencies, and test with minimal permissions.