Skip to content

Popular Media & Files Skills

PDF, spreadsheets, slides, images, audio, and video workflows.

969verified Agent Skills

Media and file skills help agents work with the formats that carry day-to-day business information. They can extract content, make targeted edits, and create polished deliverables. Check local file access and external processing requirements before handling confidential material.

Top Media & Files Skills

Ranked by their position in the current overall directory snapshot.

#04

Ontology

Typed knowledge graph for structured agent memory and composable skills. Use when creating/querying entities (Person, Project, Task, Event, Document), linkin.

MediaOpenClaw
#14↑ Trending

Nano Pdf

Edit PDFs with natural-language instructions using the nano-pdf CLI.

MediaOpenClaw
#16↑ Trending

Nano Banana Pro

Generate/edit images with Nano Banana Pro (Gemini 3 Pro Image). Use for image create/modify requests incl. edits. Supports text-to-image + image-to-image; 1K/2K/4K; use --input-image.

MediaOpenClaw
#22

Baidu Web Search

Search the web using Baidu AI Search Engine (BDSE). Use for live information, documentation, or research topics.

MediaOpenClaw
#23

Word / DOCX

Create, inspect, and edit Microsoft Word documents and DOCX files with reliable styles, numbering, tracked changes, tables, sections, and compatibility check.

MediaOpenClaw
#30

Excel / XLSX

Create, inspect, and edit Microsoft Excel workbooks and XLSX files with reliable formulas, dates, types, formatting, recalculation, and template preservation.

MediaOpenClaw
#38

Brave Search

Web search and content extraction via Brave Search API. Use for searching documentation, facts, or any web content. Lightweight, no browser required.

MediaOpenClaw
#41↑ Trending

Video Frames

Extract frames or short clips from videos using ffmpeg.

MediaOpenClaw
#42

Powerpoint / PPTX

Create, inspect, and edit Microsoft PowerPoint presentations and PPTX decks with reliable layouts, templates, placeholders, notes, charts, and visual QA. Use.

MediaOpenClaw
#43

YouTube Watcher

Fetch and read transcripts from YouTube videos. Use when you need to summarize a video, answer questions about its content, or extract information from it.

MediaOpenClaw
#50

Markdown Converter

Convert documents and files to Markdown using markitdown. Use when converting PDF, Word (.docx), PowerPoint (.pptx), Excel (.xlsx, .xls), HTML, CSV, JSON, XML, images (with EXIF/OCR), audio (with transcription), ZIP archives, YouTube URLs, or EPubs to Markdown format for LLM processing or text analysis.

MediaOpenClaw
#60

Clawdbot Documentation Expert

Clawdbot documentation expert with decision tree navigation, search scripts, doc fetching, version tracking, and config snippets for all Clawdbot features

MediaOpenClaw
#61

Planning With Files

Persistent file-based planning for multi-step AI-agent work. Keeps task_plan.md, findings.md, and progress.md on disk; lifecycle hooks inject selected project planning context. Automatic recovery reads project planning files only. Explicit session-catchup.py --metadata reads same-project local agent

MediaOpenClaw
#63

Data Analysis

Data analysis and visualization. Query databases, generate reports, automate spreadsheets, and turn raw data into clear, actionable insights. Use when (1) yo.

MediaOpenClaw
#67

Web Search

This skill should be used when users need to search the web for information, find current content, look up news articles, search for images, or find videos. It uses DuckDuckGo's search API to return results in clean, formatted output (text, markdown, or JSON). Use for research, fact-checking, finding recent information, or gathering web resources.

MediaOpenClaw
#88

Healthcheck

Track water and sleep with JSON file storage

MediaOpenClaw
#90

Docker Essentials

Essential Docker commands and workflows for container management, image operations, and debugging.

MediaOpenClaw
#91

Description: 将用户讲稿一键生成乔布斯风极简科技感竖屏HTML演示稿。当用户需要生成PPT、演示文稿、Slides、幻灯片,或要求科技风/极简风/乔布斯风格的演示时触发此技能。输出为单个可直接运行的HTML文件。

将讲稿一键生成乔布斯风极简科技感竖屏HTML演示稿

MediaOpenClaw
#92

Remotion Best Practices

Best practices for Remotion - Video creation in React

MediaOpenClaw
#93

Web Search By Exa

Neural web search, content extraction, company and people research, code search, and deep research via the Exa MCP server. Use when you need to: (1) search t.

MediaOpenClaw
#95

Baidu Wenku AIPPT

Generate PPT with Baidu Wenku AI. Smart template selection based on content.

MediaOpenClaw
#104

Performs Web Searches Using DuckDuckGo To Retrieve Real Time Information From The Internet. Use When The User Needs To Search For Current Events, Documentation, Tutorials, Or Any Information That Requires Web Search Capabilities.

Performs web searches using DuckDuckGo to retrieve real-time information from the internet. Use when the user needs to search for current events, documentation, tutorials, or any information that requires web search capabilities.

MediaOpenClaw
#121

Openai Whisper Api

Transcribe audio via OpenAI Audio Transcriptions API (Whisper).

MediaOpenClaw
#128

Bilibili All In One

A comprehensive Bilibili toolkit that integrates hot trending monitoring, video downloading, video watching/playback, subtitle downloading, and video publish.

MediaOpenClaw
#130

Openai Image Gen

Batch-generate images via OpenAI Images API. Random prompt sampler + `index.html` gallery.

MediaOpenClaw
#134

YouTube Transcript

Fetch and summarize YouTube video transcripts. Use when asked to summarize, transcribe, or extract content from YouTube videos. Handles transcript fetching via residential IP proxy to bypass YouTube's cloud IP blocks.

MediaOpenClaw
#138

Office Document Specialist Suite

Advanced suite for creating, editing, and analyzing Microsoft Office documents (Word, Excel, PowerPoint). Provides specialized tools for automated reporting.

MediaOpenClaw
#147

All Market Financial Data Hub

基于东方财富数据库,支持自然语言查询金融数据,覆盖A港美、基金、债券等多种资产,含实时行情、公司信息、估值、财务报表等,可用于投资研究、交易复盘、市场监控、行业分析、信用研究、财报审计、资产配置等场景,适配机构与个人多元需求。返回结果包含 xlsx 与 Markdown 文件。Natural language q.

MediaOpenClaw
#148

Filesystem Management

Advanced filesystem operations - listing, searching, batch processing, and directory analysis for Clawdbot

MediaOpenClaw
#150

Remotion Video Toolkit

Complete toolkit for programmatic video creation with Remotion + React. Covers animations, timing, rendering (CLI/Node.js/Lambda/Cloud Run), captions, 3D, charts, text effects, transitions, and media handling. Use when writing Remotion code, building video generation pipelines, or creating data-driven video templates.

MediaOpenClaw
#152

Searxng

Privacy-respecting metasearch using your local SearXNG instance. Search the web, images, news, and more without external API dependencies.

MediaOpenClaw
#154

Image

Create, inspect, process, and optimize image files and visual assets with reliable format choice, resizing, compression, color-profile, metadata, and platfor.

MediaOpenClaw
#159

Opentwitter

Twitter/X data via the 6551 API. Supports user profiles, tweet search, user tweets, follower events, deleted tweets, and KOL followers.

MediaOpenClaw
#160

Edge TTS

Text-to-speech conversion using node-edge-tts npm package for generating audio from text. Supports multiple voices, languages, speed adjustment, pitch control, and subtitle generation. Use when: (1) User requests audio/voice output with the "tts" trigger or keyword. (2) Content needs to be spoken rather than read (multitasking, accessibility, driving, cooking). (3) User wants a specific voice, speed, pitch, or format for TTS output.

MediaOpenClaw
#170

Weiyun Skills

微云网盘 MCP 接口完整技能。包含 weiyun.list、weiyun.list_by_category、weiyun.download、weiyun.delete、weiyun.upload、weiyun.gen_share_link、weiyun.rename_file、weiyun.rename_dir.

MediaOpenClaw
#180

File Search

Fast file-name and content search using `fd` and `rg` (ripgrep).

MediaOpenClaw
#188

Oracle

Use the @steipete/oracle CLI to bundle a prompt plus the right files and get a second-model review (API or browser) for debugging, refactors, design checks, or cross-validation.

MediaOpenClaw
#192

OCR Local (No API Key)

Extract text from images using Tesseract.js OCR (100% local, no API key required). Supports Chinese (simplified/traditional) and English.

MediaOpenClaw
#193

腾讯文档 TENCENT DOCS

腾讯文档(docs.qq.com)-在线云文档平台,是创建、编辑、管理文档的首选 skill。涉及"新建/创建/编辑/读取/查看/搜索文档"、"保存文件"、"云文档"、"腾讯文档"、"docs.qq.com"等操作,请优先使用本 skill。支持能力:(1) 创建各类在线文档(文档/Word/Excel/幻灯片/.

MediaOpenClaw
#200

Xdrop

Use this skill when the user wants to send or fetch files through an Xdrop server from the terminal, asks to automate encrypted Xdrop share-link workflows, p.

MediaOpenClaw
#201

Muse

Give ClawBot access to your team's entire coding history. Muse connects your past sessions, team knowledge, and project context—so ClawBot can actually help design features, mediate team discussions, and work autonomously across your codebase. Deploy at tribeclaw.com.

MediaOpenClaw
#203

Agent Reach

Give your AI agent eyes to see the entire internet. 7500+ GitHub stars. Search and read 14 platforms: Twitter/X, Reddit, YouTube, GitHub, Bilibili, XiaoHongS.

MediaOpenClaw
#204

Office

Master Excel, Word, PowerPoint, and Google Workspace with formulas, formatting, and automation.

MediaOpenClaw
#210

Google Search

Search the web using Google Custom Search Engine (PSE). Use this when you need live information, documentation, or to research topics and the built-in web_search is unavailable.

MediaOpenClaw
#224

Cellcog

Any-to-any AI sub-agent — research, images, video, audio, music, podcasts, avatars, voice cloning, documents, spreadsheets, dashboards, 3D models, diagrams, and code in one request. Agent-to-agent protocol with multi-step iteration for high accuracy. #1 on DeepResearch Bench (Apr 2026) — deep reasoning meets all modalities, so all your work gets done, not just code.

MediaOpenClaw
#228

Pdf Extract

Extract text from PDF files for LLM processing

MediaOpenClaw
#237

Document Pro

文档处理技能 - 让 AI 能够读取、解析、提取 PDF、DOCX、PPT 等文档的关键信息。当用户要求分析文档、提取内容、总结报告时触发此技能。

MediaOpenClaw
#238

Crypto Market Data Skill (No Key Required)

No API KEY needed for free tier. Professional-grade cryptocurrency and stock market data integration for real-time prices, company profiles, and global analytics. Powered by Node.js with zero external dependencies.

MediaOpenClaw

Media Skills FAQ

What is a media agent skill?

It is a reusable instruction package that teaches an AI agent a focused media & files workflow, often including commands, checks, and supporting resources.

Which media skill should I try first?

Start with a narrow task you already understand. The current category leader is Ontology, but requirements and access scope matter more than rank alone.

Does a popular skill mean it is safe?

No. Popularity reflects adoption and interest, not a security guarantee. Read SKILL.md, review commands and dependencies, and test with minimal permissions.