Skip to content

Popular Media & Files Skills

PDF, spreadsheets, slides, images, audio, and video workflows.

969verified Agent Skills

Media and file skills help agents work with the formats that carry day-to-day business information. They can extract content, make targeted edits, and create polished deliverables. Check local file access and external processing requirements before handling confidential material.

Top Media & Files Skills

Ranked by their position in the current overall directory snapshot.

#1747

Gemini STT

Transcribe audio files using Google's Gemini API or Vertex AI

MediaOpenClaw
#1749

Prompt Log

Extract conversation transcripts from AI coding session logs (Clawdbot, Claude Code, Codex). Use when asked to export prompt history, session logs, or transcripts from .jsonl session files.

MediaOpenClaw
#1763

Manage Bambu Labs 3D Printers Thru Your Agent

Operate and troubleshoot BambuLab printers with the bambu-cli (status/watch, print start/pause/resume/stop, files, camera, gcode, AMS, calibration, motion, fans, light, config, doctor). Use when a user asks to control or monitor a BambuLab printer, set up profiles or access codes, or translate a task into safe bambu-cli commands with correct flags, output format, and confirmations.

MediaOpenClaw
#1765

ElevenLabs Music

Generate music from text prompts using ElevenLabs Eleven Music API. Use when creating songs, soundtracks, jingles, lullabies, or any audio music from descriptions. Supports vocals with AI-generated lyrics, instrumental tracks, and multiple genres/styles. Requires paid ElevenLabs plan.

MediaOpenClaw
#1769

Image2Prompt

Analyze images and generate detailed prompts for image generation. Supports portrait, landscape, product, animal, illustration categories with structured or natural output.

MediaOpenClaw
#1772

Tencent Agent Storage

Cloud file storage, upload, backup, and file management tool for Tencent Agent Storage (专属网盘). Manages the user's personal cloud drive: upload files, list files, download, share links, preview, and backup. MUST trigger when the user mentions ANY of the following concepts: 【云盘/网盘相关 — Cloud Drive Acce

MediaOpenClaw
#1775

3D Model Generation

AI 3D model generation powered by CellCog. Text-to-3D, image-to-3D — production-ready GLB files for games, AR/VR, e-commerce, and 3D printing. Game assets, product visualization, characters, props, environments, and batch generation.

MediaOpenClaw
#1783

LaTeX

Write LaTeX documents with correct syntax, packages, and compilation workflow.

MediaOpenClaw
#1789

Screenshot Capture

Process screenshots Enzo shares with comments. Save to reference library, extract content, categorize, set reminders, and log patterns. Use when Enzo sends an image with context like "save this", shares a screenshot of content (LinkedIn posts, tweets, articles), or sends ideas/frameworks to remember.

MediaOpenClaw
#1798

分镜图工作流 Image Storyboard

A professional storyboard skill for film, advertising, short video, and educational narrative scenarios, built around a strict 'plan first, render later' flow.

MediaOpenClaw
#1800

CostHQ

Track agent session costs, file changes, and git commits with CostHQ. Enforces budget limits, tracks local models, and provides Enterprise SOC2 audit trails.

MediaOpenClaw
#1801

Async Task

Run and manage long tasks exceeding HTTP timeouts by starting, updating, and completing them asynchronously with immediate responses.

MediaOpenClaw
#1814

Excel Xlsx 1

Create, inspect, and edit Microsoft Excel workbooks and XLSX files with reliable formulas, dates, types, formatting, recalculation, and template preservation.

MediaOpenClaw
#1815

Audio Content Generator

Generate audiobooks, podcasts, or educational audio content on demand. User provides an idea or topic, Claude AI writes a script, and ElevenLabs converts it to high-quality audio. Supports multiple formats (audiobook, podcast, educational), custom lengths, and voice effects. Use when asked to create audio content, make a podcast, generate an audiobook, or produce educational audio. Returns MP3 audio file via MEDIA token.

MediaOpenClaw
#1820

Qmd External Knowledge Base Search

Local hybrid search for markdown notes and docs. Use when searching notes, finding related content, or retrieving documents from indexed collections.

MediaOpenClaw
#1823

PowerPoint PPTX处理

Create, inspect, and edit Microsoft PowerPoint presentations and PPTX decks with reliable layouts, templates, placeholders, notes, charts, and visual QA. Use.

MediaOpenClaw
#1826

社交媒体配图设计 Social Media Design

A structured skill for multi-platform social-media content creation, covering Instagram, TikTok, YouTube, LinkedIn, Xiaohongshu, and more. The goal: outputs.

MediaOpenClaw
#1827

Youtube Channels

Use when a YouTube channel is the focus: pasted @handles or channel URLs, requests to browse a creator's uploads, see what a channel has posted recently, sea.

MediaOpenClaw
#1828

Crawl4AI Web Scraper

Full web page scraping with JavaScript rendering via local Crawl4AI instance, delivering clean markdown or detailed JSON including links and media.

MediaOpenClaw
#1837

Ocr Document

OCR document extraction - extract text from scanned documents, photos, and images using OCR. Use when reading scanned PDFs, photographed pages, handwritten n.

MediaOpenClaw
#1838

Meme Generator

AI meme generator powered by CellCog. Memes, viral content, reaction images, internet humor. Audience targeting, trend research, and multi-angle generation for humor that lands.

MediaOpenClaw
#1839

Ai Video Generation

Generate AI videos with Google Veo, Seedance, Wan, Grok and 40+ models via inference.sh CLI. Models: Veo 3.1, Veo 3, Seedance 1.5 Pro, Wan 2.5, Grok Imagine.

MediaOpenClaw
#1841

OpenClaw ComfyUI

Connect and control ComfyUI API efficiently using template mapping and auto-asset management for image generation and editing tasks.

MediaOpenClaw
#1844

AI Presentation Maker

AI Presentation Maker — the interview-driven pitch deck generator for your OpenClaw agent. Tell it what you built, who you're presenting to, and pick an angl.

MediaOpenClaw
#1848

Vet

Run vet immediately after ANY logical unit of code changes. Do not batch your changes, do not wait to be asked to run vet, make sure you are proactive.

MediaOpenClaw
#1850

Tax Professional

Comprehensive US tax advisor, deduction optimizer, and expense tracker. Covers all employment types (W-2, 1099, S-Corp, mixed), estimated tax payments, audit risk assessment, life event triggers, multi-state filing, RV-as-home rules, tax bracket optimization, document retention, and proactive year-round tax calendar nudges. Your CPA in the pocket.

MediaOpenClaw
#1852

AIsa Twitter API (Search + Post)

Searches and reads X (Twitter): profiles, timelines, mentions, followers, tweet search, trends, lists, communities, and Spaces. Publishes posts after the use.

MediaOpenClaw
#1854

Apple Media Remote (for HomePod, Apple TV, Etc)

Control Apple TV, HomePod, and AirPlay devices via pyatv (scan, stream, playback, volume, navigation).

MediaOpenClaw
#1855

Workout

Track workouts, log sets, manage exercises and templates with workout-cli. Supports multi-user profiles. Use when helping users record gym sessions, view history, or analyze strength progression.

MediaOpenClaw
#1856

营销宣传册设计 Marketing Brochure

A complete workflow skill for marketing brochure design, covering everything from requirements gathering, layout design, to mock-up delivery. It uses a 'layout-first + mandatory confirmation

MediaOpenClaw
#1857

Video Watch

Analyze video content by extracting frames at regular intervals. Use when you need to understand what's in a video file, review video content, analyze scenes, or describe video without being able to play it directly. Supports MP4, MOV, AVI, MKV, and other common video formats.

MediaOpenClaw
#1858

图生视频 即梦首帧 Jimeng I2V

Generate dynamic videos based on a single first frame image and prompts using Jimeng. 使用即梦 (Jimeng) 首帧生视频模型,基于单张首帧图片和提示词生成动态视频。

MediaOpenClaw
#1861

亚马逊商品图套件 Amazon Product Image Suite

A professional product image generation skill purpose-built for the Amazon e-commerce platform. Outputs comply with Amazon's image guidelines while optimizing for click-through and conversio

MediaOpenClaw
#1864

Comfyui Anfrage

Send a workflow request to ComfyUI and return image results.

MediaOpenClaw
#1866

Media Downloader

Download Video/Music from YouTube/Bilibili/X/etc.

MediaOpenClaw
#1867

Paperless

Interact with Paperless-NGX document management system via ppls CLI. Search, retrieve, upload, and organize documents.

MediaOpenClaw
#1868

Dating Platform. 约会。Citas.

Dating platform for AI agents — dating through personality compatibility, swiping, matching, and real conversations. Dating profiles with Big Five traits, da.

MediaOpenClaw
#1870

Markdown To PPT (Smart Layout)

智能 Markdown 转 PPT。自动分析内容结构、智能分页、详细设计每页布局、自动生成/搜索配图。支持 Slidev/HTML/PPTX 多格式输出。| Intelligent Markdown to PPT with auto-layout and image generation.

MediaOpenClaw
#1874

Shanku Paolu 1.0.0

离职前文件扫描与备份工具。纯HTML单文件,零依赖,双击即用。 当用户提到"离职"、"备份文件"、"扫描文件"、"整理文件"、"删库跑路"时使用。 触发词:离职备份、文件扫描、备份重要文件、扫描PDF/Word/Excel/PPT、离职前整理、删库跑路。

MediaOpenClaw
#1875

文件总结 File Summary & Analysis

Local document summary tool. Activate when user mentions "总结文件", "帮我总结", "总结文档", "分析文档" or provides a local file path (txt/docx/pdf/xlsx/xls).

MediaOpenClaw
#1879

短视频口播文案 Spoken Script

This skill is used to guide the AI in generating short video spoken scripts with high contrast, strong resonance, a sense of story, and personal IP attributes. All generated scripts must str

MediaOpenClaw
#1886

社交轮播图设计 Social Carousel

A structured workflow skill dedicated to social-media carousel design. The core method is 'decide intent first, then execute,' using a 'single-confirmation + cover-first' two-phase flow.

MediaOpenClaw
#1887

PLS Canvas Design

Generates original visual art and posters as PNG or PDF files using defined design philosophies like minimalism, brutalism, or skeuomorphism.

MediaOpenClaw
#1890

Pdf

Comprehensive PDF manipulation toolkit for extracting text and tables, creating new PDFs, merging/splitting documents, and handling forms. When Claude needs.

MediaOpenClaw
#1891

Instagram Skill Via Cyberdrk/gram CLI

Instagram CLI for viewing feeds, posts, profiles, and engagement via cookies.

MediaOpenClaw
#1892

Memory Search

Search and retrieve relevant information from your indexed memory files using semantic queries and direct file reads for context.

MediaOpenClaw
#1897

Bilili Downloader

Download Bilibili videos. You MUST ask the user for the Bilibili URL first. Then use the provided python script to download. Supports batch/playlist downloading.

MediaOpenClaw
#1901

Baoyu Url To Markdown

Fetch any URL and convert to markdown using baoyu-fetch CLI (Chrome CDP with site-specific adapters). Built-in adapters for X/Twitter, YouTube transcripts, H.

MediaOpenClaw

Media Skills FAQ

What is a media agent skill?

It is a reusable instruction package that teaches an AI agent a focused media & files workflow, often including commands, checks, and supporting resources.

Which media skill should I try first?

Start with a narrow task you already understand. The current category leader is Ontology, but requirements and access scope matter more than rank alone.

Does a popular skill mean it is safe?

No. Popularity reflects adoption and interest, not a security guarantee. Read SKILL.md, review commands and dependencies, and test with minimal permissions.