Media AI Skills

708 open-source Media AI skills that teach any AI model a new workflow.

Search and filter AI skills

Showing 511-561 of 708 AI skills

AI skills directory results

gooseworks-ai logo

Render Multiworld

gooseworks-ai
OrganizationPopular

Assemble a silent, music-led 3-world product-tour ad — trim and hard-cut-concat the per-world WIDE-arrival + top-down-macro clips, composite the HTML/Playwright brand end card ("FIND YOUR DAILY." +…

208
1.2K
4 files
View
PlayableIntelligence logo

Worldlabs

PlayableIntelligence
Organization

Generate photorealistic 3D worlds and environments with the World Labs Marble API — Gaussian Splat scenes from text prompts or reference images. Use when the user says "generate a 3D world", "create…

41
331
Instructions
View
hoodini logo

Video To Landing Page

hoodini
Community

Turn any video into a cinematic scroll-driven landing page — Apple-style hero where scrolling progresses the visible frame through the video. Use when the user provides a video file and asks for "a…

62
280
4 files
View
TerminalSkills logo

Ai Video Generator

TerminalSkills
Organization

Generate short-form videos with AI — script writing, text-to-speech narration, stock footage selection, subtitle generation, and video assembly. Use when: creating TikTok/YouTube Shorts/Reels content,…

21
155
1 files
View
zoom logo

Translator

zoom
Organization

Zoom AI Services Translator for synchronous text translation and asynchronous batch file translation. Use for plain-text translation, one-target-language jobs, S3 text archives, Build-platform JWT…

16
78
8 files
View
spinabot logo

Openai Whisper Api

spinabot
OrganizationPopular

Transcribe audio via OpenAI Audio Transcriptions API (Whisper).

50
4.4K
1 files
View
SamurAIGPT logo

Muapi Youtube Shorts

SamurAIGPT
OrganizationPopular

Auto-generate viral 9:16 YouTube Shorts (or TikTok / Reels clips) from a long-form video. Thin platform-aware wrapper around the AI Clipping skill — picks sensible defaults for short-form social…

492
4.3K
1 files
View
Pluviobyte logo

Ra Audio To Subtitles

Pluviobyte
CommunityPopular

Generate production subtitle artifacts from the final narration audio or final merged video using Volcengine Doubao ASR word timestamps. Use for local IndexTTS2 videos, Xiaohei page videos,…

181
1.6K
3 files
View
first-fluke logo

Oma Video

first-fluke
OrganizationPopular

Create short, explainer, or recorded-demo videos through the OMA video CLI. Use for scripts, narration, assets, composition, and video delivery.

149
1.3K
13 files
View
vellum-ai logo

Deepgram Voice

vellum-ai
OrganizationPopular

Select and tune a Deepgram TTS voice - curated voice list, full Aura voice catalog via API key, and tuning parameters

186
1.3K
1 files
View
gooseworks-ai logo

Render Myth Vs Fact

gooseworks-ai
OrganizationPopular

Assemble a myth-vs-fact kinetic-typography explainer video ad (≈29.5s, 9:16) from N myth/fact pairs + hook / turn / punch copy + palette + a brand end-card PNG + a VO track — a hook, 3 red-strike MYTH…

208
1.2K
14 files
View
glebis logo

Elevenlabs Tts

glebis
Community

This skill converts text to high-quality audio files using ElevenLabs API. Use this skill when users request text-to-speech generation, audio narration, or voice synthesis with customizable voice…

56
379
3 files
View
guia-matthieu logo

Video Processing

guia-matthieu
Community

Process video files with ffmpeg automation. Use when: compressing videos for upload; extracting audio from video; resizing for social formats; clipping segments; merging multiple videos; generating…

27
150
2 files
View
cometchat logo

Cometchat Flutter V5 Sdk

cometchat
Organization

Add voice & video calling to any Flutter app FROM SCRATCH with the headless CometChat Calls SDK v5 (`cometchat_calls_sdk`, pub.dev) — no UI Kit. init→login→generateCallToken→joinSession, which hands…

2
109
4 files
View
spinabot logo

Openai Whisper

spinabot
OrganizationPopular

Local speech-to-text with the Whisper CLI (no API key).

50
4.4K
Instructions
View
rlaope logo

Omh External Connector Readiness

rlaope
CommunityPopular

[omh] External connector readiness - assess whether a named plugin, connector, API, data provider, or multimodal route is safe, affordable, fresh, and observable; use executor-runtime-readiness for…

194
2.7K
1 files
View
questflowai logo

Libei Macro Hedge

questflowai
OrganizationPopular

Use when evaluating markets through Li Bei's macro hedge lens: regime identification, trend + contrarian timing, position management by conviction, and risk control for drawdown.

138
1.8K
1 files
View
benchflow-ai logo

Voice Activity Detection (VAD)

benchflow-ai
OrganizationPopular

Detect speech segments in audio using VAD tools like Silero VAD, SpeechBrain VAD, or WebRTC VAD. Use when preprocessing audio for speaker diarization, filtering silence, or segmenting audio into…

367
1.8K
Instructions
View
first-fluke logo

Oma Voice

first-fluke
OrganizationPopular

Generate speech or transcribe audio locally with Voicebox. Use for narration, voice assets, dictation, and meeting transcription.

149
1.3K
5 files
View
gooseworks-ai logo

Render Narrated Ugc Wardrobe Stitch

gooseworks-ai
OrganizationPopular

Assemble a narrated-UGC "stitch reply" ad from a config — a single spoken VO carries a verbatim testimonial while ~30 per-cut i2v clips (one creator across ~5 wardrobes in ~3 worlds, plus product…

208
1.2K
4 files
View
github logo

Automate This

github
OrganizationPopular

Analyze a screen recording of a manual process and produce targeted, working automation scripts. Extracts frames and audio narration from video files, reconstructs the step-by-step workflow, and…

5K
39.1K
Instructions
View
TokenRhythm logo

Short Drama Review Normalizer

TokenRhythm
OrganizationPopular

Internal deterministic consent gate for meta-short-drama. Normalizes draft review and post-revision confirmation replies, and fails closed before external image/video generation.

566
7K
1 files
View
FreedomIntelligence logo

Bio Chipseq Peak Calling

FreedomIntelligence
OrganizationPopular

ChIP-seq peak calling using MACS3 (or MACS2). Call narrow peaks for transcription factors or broad peaks for histone modifications. Supports input control, fragment size modeling, and various output…

410
3K
2 files
View
Pluviobyte logo

Ra Local Talking Head Cut

Pluviobyte
CommunityPopular

Produce a polished local talking-head or narrated screen-recording rough cut without a cloud editor. Use when Codex must clean Chinese or mixed Chinese-English speech, correct product terminology…

181
1.6K
10 files
View
gooseworks-ai logo

Render Offer Ad

gooseworks-ai
OrganizationPopular

Render a punchy ~12s vertical (9:16) music-only direct-response OFFER ad as a 4-beat kinetic-typography film — HEADLINE slam → real PRODUCT drop → CLAIM/proof → CTA pill — from one config of copy…

208
1.2K
17 files
View
glebis logo

Fathom

glebis
Community

Fetch meetings, transcripts, summaries, and action items from Fathom API. Use when user asks to get Fathom recordings, sync meeting transcripts, or fetch recent calls.

56
379
3 files
View
guia-matthieu logo

Whisper Transcription

guia-matthieu
Community

Transcribe audio and video files to text using OpenAI Whisper. Use when: converting podcasts to blog posts; creating video subtitles; extracting quotes from interviews; repurposing video content to…

27
150
2 files
View
letta-ai logo

Ai News

letta-ai
Organization

Fetch and summarize recent AI news from curated RSS feeds (Hugging Face, VentureBeat, The Verge, OpenAI, Anthropic, DeepMind, etc.) and YouTube channels (Yannic Kilcher, Two Minute Papers, AI…

25
144
4 files
View
cometchat logo

Cometchat Flutter V6 Components

cometchat
Organization

The Flutter v6 UI Kit widget catalog — which drop-in widgets exist, which barrel each comes from, and how to compose or swap them (conversations, messages, users, groups, threads, calling, bubbles).…

2
109
Instructions
View
jonathimer logo

Youtube Devrel

jonathimer
Community

When the user wants to create developer YouTube content, technical screencasts, or video tutorials. Trigger phrases include "YouTube," "developer video," "screencast," "video tutorial," "live coding,"…

6
87
Instructions
View
zhaoxuya520 logo

Competition Stego Media

zhaoxuya520
CommunityPopular

Internal downstream skill for ctf-sandbox-orchestrator. CTF-sandbox workflow for image, audio, video, document, and container steganography. Use when the user asks to inspect metadata, alpha or…

5K
36.3K
2 files
View
davidondrej logo

Deepapi

davidondrej
CommunityPopular

Use DeepAPI for all web search, deep research, and web scraping (websites, LinkedIn, GitHub, X/Twitter, YouTube, Instagram) instead of built-in search, research, fetch, or browser tools. Prefer…

599
4.1K
9 files
View
benchflow-ai logo

Ffmpeg Video Editing

benchflow-ai
OrganizationPopular

Video editing with ffmpeg including cutting, trimming, concatenating segments, and re-encoding. Use when working with video files (.mp4, .mkv, .avi) for: removing segments, joining clips, extracting…

367
1.8K
Instructions
View
Pluviobyte logo

Ra Video Download

Pluviobyte
CommunityPopular

Download source video or audio from Douyin, YouTube, Bilibili, Twitter/X, Xiaohongshu, and other yt-dlp-supported URLs into the content-creation workspace. Use when the user says 下载视频, 下载音频, 保存这个链接,…

181
1.6K
2 files
View
alchaincyf logo

Huashu Speech Coach

alchaincyf
CommunityPopular

演讲与分享教练。基于Patrick Winston(MIT AI教授)的How to Speak方法论,帮助准备线下培训、技术分享、B站教程视频等演讲场景。当用户提到"演讲"、"分享"、"培训"、"讲课"、"PPT演讲"、"开场"、"结尾"、"如何讲"、"演讲结构"时使用此技能。

244
1.6K
1 files
View
gooseworks-ai logo

Render Photo Grid Card

gooseworks-ai
OrganizationPopular

Render a 'photo-grid promo card' video from a config — a real-DOM card with a FIXED header/footer and a CONTINUOUSLY SCROLLING 2-row grid of MIXED tiles (video clips, product/lifestyle stills,…

208
1.2K
5 files
View
BankrBot logo

Bankr Communities

BankrBot
OrganizationPopular

Bankr Space ↔ bankr.bot/agents two-way sync (BANKR-PROJECT-SYNC.md Paths B+C). Original tweets from GET /agent-profiles/:id/tweets shown on Spaces. Holder votes: yes/no or multiple-choice polls…

622
1.2K
40 files
View
hoodini logo

Yuv Design System

hoodini
Community

Yuval Avidani's YUV.AI brand and design system. Apply ONLY when YUV.AI-branded output is requested — presentations, decks, keynotes, portfolio, brand site, profile, speaker bio, brand assets, or any…

62
280
29 files
View
guia-matthieu logo

Youtube Downloader

guia-matthieu
Community

Download and process YouTube content for research. Use when: downloading competitor videos for analysis; extracting audio for podcasts; getting transcripts for content repurposing; archiving webinars;…

27
150
2 files
View
dykyi-roman logo

Check Command Injection

dykyi-roman
Community

Analyzes PHP code for command injection vulnerabilities. Detects shell_exec, exec, system, passthru with user input, missing escapeshellarg/escapeshellcmd.

25
98
Instructions
View
TokenRhythm logo

Srt From Script

TokenRhythm
OrganizationPopular

Build an SRT subtitle file from a 3-shot short-drama script (ai-video-script OUTPUT FORMAT). Reads each SHOT_N block's DURATION_S + VOICEOVER, emits cumulative-timestamped SRT cues. Pure…

566
7K
1 files
View
dotnet logo

Maui Safe Area

dotnet
OrganizationPopular

.NET MAUI safe area and edge-to-edge layout guidance for .NET 10+. Covers the new SafeAreaEdges property, SafeAreaRegions enum, per-edge control, keyboard avoidance, Blazor Hybrid CSS safe areas,…

416
5.4K
1 files
View
xu-xiang logo

Content Engine

xu-xiang
CommunityPopular

为X、LinkedIn、TikTok、YouTube、新闻通讯和跨平台重新利用的多平台活动创建平台原生内容系统。适用于当用户需要社交媒体帖子、帖子串、脚本、内容日历,或一个源资产在多个平台上清晰适配时。

318
1.9K
Instructions
View
benchflow-ai logo

Filler Word Processing

benchflow-ai
OrganizationPopular

Process filler word annotations to generate video edit lists. Use when working with timestamp annotations for removing speech disfluencies (um, uh, like, you know) from audio/video content.

367
1.8K
Instructions
View
Pluviobyte logo

Ra Video Production Director

Pluviobyte
CommunityPopular

End-to-end video production orchestration for the content-creation workspace. Use when the user asks to make, recreate, package, render, QC, or archive a video; when the user says 制作待制作队列, 按交接稿制作,…

181
1.6K
18 files
View
vellum-ai logo

Elevenlabs Voice

vellum-ai
OrganizationPopular

Select and tune an ElevenLabs TTS voice - curated voice list, custom/cloned voices via API key, and tuning parameters

186
1.3K
1 files
View
gooseworks-ai logo

Render Podcast Skit

gooseworks-ai
OrganizationPopular

Assemble a two-host fake-podcast skit ad from a config — per-line lipsync clips hard-concatenated in script order, scaled/padded to 1080×1920, WHITE bottom-center captions (up to 5 words per cue,…

208
1.2K
4 files
View
hoodini logo

Yuv Reel Covers

hoodini
Community

Generate unified, on-brand Instagram Reel covers for Yuval (YUV.AI Neon Phoenix system) — the signature look is a giant Hebrew headline BEHIND the subject cutout + a punch line IN FRONT (depth…

62
280
12 files
View
anbeime logo

Infinitetalk

anbeime
CommunityPopular

自媒体创作者与内容创作者在制作数字人播报或视频配音时,只需输入单张人像与音频,即可自动生成唇形、表情、动作完美同步的无限时长说话视频。一键打造高质量虚拟主播内容,告别繁琐拍摄,让音视频创作更高效!

645
6.9K
8 files
View
davidondrej logo

Fireflies Transcript

davidondrej
CommunityPopular

Fetch raw Fireflies.ai meeting transcripts. Use only when the user explicitly invokes /fireflies-transcript; for YouTube, use youtube-transcript.

599
4.1K
1 files
View
benchflow-ai logo

Whisper Transcription

benchflow-ai
OrganizationPopular

Transcribe audio/video to text with word-level timestamps using OpenAI Whisper. Use when you need speech-to-text with accurate timing information for each word.

367
1.8K
Instructions
View

Set up your own AI workspace now

Get notified about new features and future giveaways by subscribing to our newsletter 👇