Video, Audio & Voice Agent Skills
Video generation and editing, transcription, audio processing, speech, voice and podcast workflows.
Category assignments are rule-generated in Phase 1 unless a record says human confirmed.商品短视频广告 Product Video Ad
商品短视频广告。卖点 → 分镜脚本 → 分段生成 → 拼接加字幕成片。当用户说「做条广告」「短视频」「投流素材」「分镜脚本」「带货视频」时使用。
- Publisher
- dlazy
- Fixed release
- 1.0.13
视频生成 通义万相 Wan 2.7
Tongyi Wanxiang 2.7 video model 鈥?one model covers text-to-video, first/last-frame-to-video, and reference-to-video.
- Publisher
- dlazy
- Fixed release
- 1.3.23
网站转视频 Website to Video
website to video, url to video, landing page to video, link to ad, web promo video — capture the site, derive brand, storyboard, voiceover, build, validate on a Remotion template. Use when the user gives a URL (or just pastes a link) and wants a promo, social ad, or product demo.
- Publisher
- dlazy
- Fixed release
- 1.3.24
视频抠像 Video Segmentation
Video human segmentation tool: invokes Aliyun's async SegmentVideoBody and returns a same-length black/white mask video, suitable for downstream compositing or matting. 视频人像分割工具:调用阿里云 SegmentVideoBody 异步任务,返回与原视频同时长的 mask 视频(黑白蒙版),可用于后续合成或抠像处理。
- Publisher
- dlazy
- Fixed release
- 1.3.25
声音克隆 Vidu Audio Clone
Clone voice and generate new text reading audio with one click using Vidu Audio Clone. 使用 Vidu 声音克隆技术,通过参考音频一键复制音色并生成新文本的朗读音频。
- Publisher
- dlazy
- Fixed release
- 1.3.23
视频翻译配音 Video Translate & Dub
video translation, video dubbing, subtitle translation, translate video to Chinese, add subtitles to video, AI dubbing, srt translation, 视频翻译, 视频配音, 字幕翻译 — transcribes a video with word-level timings, translates the subtitles, then burns them in and optionally lays down a fitted dub track. Composes the dlazy fun-asr, LLM and TTS tools with ffmpeg locally; delivers a finished mp4 plus srt files, not a script.
- Publisher
- dlazy
- Fixed release
- 1.0.16
视频对口型 Video Retalk
Tongyi VideoRetalk lip sync / lip-sync (mouth sync, dubbing) video model — takes a talking-person video plus a voice audio track and regenerates the video so the speaker's mouth/lips match the new audio. Use this for lip syncing a person video to new speech. Optionally provide a reference face image to pick the target person when the video contains multiple faces. 通义声动人像 VideoRetalk 口型同步(对口型、lip sync / lip-sync、配音对嘴)视频模型,输入一段人物讲话视频与一段人声音频,生成讲话口型与音频匹配的新视频;适用于让人物视频的口型对上新的语音。当视频中存在多张人脸时,可额外提供人脸参考图来指定要替换口型的目标人物。
- Publisher
- dlazy
- Fixed release
- 1.3.24
分镜视频生成 Storyboard Video
1. Get the storyboard info
- Publisher
- dlazy
- Fixed release
- 1.2.22
视频仿制 Video Replicate
Video replicate tool: extracts the first frame and audio from the source video, runs video understanding for a prompt, and returns a Seedance 2.0 replicate bundle (first frame + audio + video). 视频复刻工具,从源视频中提取首帧与音频,并通过视频理解生成描述提示,输出 Seedance 2.0 复刻方案(首帧 + 音频 + 视频)三件套。
- Publisher
- dlazy
- Fixed release
- 1.3.24
快速视频生成 Veo 3.1 Fast
Fast response and generation of short videos with Google Veo 3.1 Fast. 使用 Google Veo 3.1 Fast 极速版模型,快速响应并生成短时长的文生视频或图生视频。
- Publisher
- dlazy
- Fixed release
- 1.3.23
视频生成 Veo 3.1
Generate high-quality cinematic effects videos with Google Veo 3.1. 使用 Google Veo 3.1 模型,生成高质量的电影级特效视频,支持文生视频与图生视频。
- Publisher
- dlazy
- Fixed release
- 1.3.23
网址转视频 URL to Video
url to video, link to video, webpage to video, landing page to video — paste a URL and turn the page into a promo / ad / demo video: capture, derive brand, storyboard, voiceover, build. Use when the user gives a URL or link and wants a video.
- Publisher
- dlazy
- Fixed release
- 1.0.20
分镜脚本 Storyboard
storyboard to video, character animation, animated short, AI anime, multi-shot video — script, characters and shot prompts, ref-sheets and first/last frames, i2v shot videos, voice/TTS plus music/SFX/subtitles, then Remotion assembles and renders. Use for a multi-shot animated short with consistent characters.
- Publisher
- dlazy
- Fixed release
- 1.3.24
短视频生成 Short Video
short video maker, tiktok video, youtube shorts, instagram reels, douyin video, vertical video, 9:16 video — hook-first single-thread storyboard, per-shot first frames and i2v clips, TTS voiceover, then Remotion assembles with burned-in subtitles. Delivers a finished 15-25s vertical mp4, not a script. Use for social shorts; for conversion-focused product ads use product-to-ecommerce-video instead.
- Publisher
- dlazy
- Fixed release
- 1.2.25
视频生成 Seedance 2.0
ByteDance's latest video generation model. Supports multi-modal reference (images, video, audio) to generate videos, as well as first/last frame and text-to-video modes. 字节跳动最新视频生成模型 Seedance 2.0,支持多模态参考(图片 + 视频 + 音频)生视频、首尾帧及文生视频,适合高质量多样化视频创作。
- Publisher
- dlazy
- Fixed release
- 1.3.22
视频生成 Seedance 2.5
ByteDance's next-generation video model: up to 30 seconds per clip with native 4K, substantially better instruction following and long-form narrative. Supports multi-modal references (image + video + audio) and first/last frame control.
- Publisher
- dlazy
- Fixed release
- 1.2.20
脚本转视频 Script to Video
script to video, screenplay to video, shot list to video — turn a script into a storyboarded, shot-by-shot video: break down scenes, generate shots, assemble, validate. Use when the user gives a script or scene breakdown and wants a video.
- Publisher
- dlazy
- Fixed release
- 1.0.21
语音合成 通义 Qwen TTS
Alibaba Bailian qwen3-tts text-to-speech. Choose from curated system voices (including dialects) or design a custom voice from a natural-language description. 阿里云百炼 qwen3-tts 文本转语音,支持系统音色(含方言)或通过自然语言描述自定义新音色(声音设计)。
- Publisher
- dlazy
- Fixed release
- 1.3.24
声音克隆 通义 Qwen Audio Clone
Alibaba Bailian qwen3-tts voice cloning. Upload a clean voice sample to clone a custom voice usable in subsequent TTS calls. 阿里云百炼 qwen3-tts 声音复刻,上传一段干净人声样本即可复刻自定义音色,可在后续 TTS 中使用。
- Publisher
- dlazy
- Fixed release
- 1.3.23
产品视频生成 Product Video
product video, product demo video, product ad video, product showcase — turn a product's photos or link into a polished demo / ad video. Use when the user wants a product demo, showcase, or ad video.
- Publisher
- dlazy
- Fixed release
- 1.0.21
动态图形视频 Motion Graphics
motion graphics, kinetic typography, animated text video, animated infographic, explainer animation — authored as Remotion code (text, shapes, data, logos, transitions), then polished and exported. Use for code-driven animated graphics rather than AI-generated footage.
- Publisher
- dlazy
- Fixed release
- 1.3.24
视频生成 MiniMax H3
MiniMax Hailuo omni-modal video model with native stereo audio, producing 5-15 second clips at up to 2K. Supports text-to-video, first/last frame transitions and multi-asset references for character and scene consistency.
- Publisher
- dlazy
- Fixed release
- 1.2.21
数字人视频 即梦 OmniHuman 1.5
Generate realistic digital human broadcast videos from portrait images and audio/text using Jimeng OmniHuman 1.5. 使用即梦 (Jimeng) OmniHuman 1.5 模型,通过人像图片和音频/文本生成逼真的数字人播报视频。
- Publisher
- dlazy
- Fixed release
- 1.3.23
首尾帧视频 即梦 Jimeng First-Tail
Generate coherent transition videos using Jimeng's first and tail frame models. 使用即梦 (Jimeng) 首尾帧生视频模型,通过提供的第一帧和最后一帧图片生成连贯的过渡视频。
- Publisher
- dlazy
- Fixed release
- 1.3.24
创意转视频 Idea to Video
Turn a user's idea into the full pipeline: **story → characters → 3-view portraits → scenes → shots → keyframes → shot videos → concat**. First emit a **plan...
- Publisher
- dlazy
- Fixed release
- 1.3.27
视频生成 快马 Happyhorse 1.0
Happy Horse 1.0 video model — one model covers text-to-video (t2v), first-frame-to-video (i2v), reference-to-video (r2v), and video editing (edit). The selected mode is automatically routed to the matching sub-model. Happy Horse 1.0 视频模型,一站式覆盖文生视频(t2v)、首帧生视频(i2v)、参考图生视频(r2v)与视频编辑(edit):根据所选模式自动路由到对应子模型。
- Publisher
- dlazy
- Fixed release
- 1.3.24
全能生成 Generate
A comprehensive generation skill. Can generate images, videos, and audio by automatically selecting the appropriate dlazy CLI model. 综合生成技能。能够根据用户意图自动选择合适的 dlazy CLI 模型来生成图片、视频或音频。
- Publisher
- dlazy
- Fixed release
- 1.3.25
文章转视频 Article to Video
article to video, text to video, news to video, essay to video — turn a written article into a narrated explainer video: outline, storyboard, voiceover, build, validate. Use when the user pastes or gives an article and wants a video.
- Publisher
- dlazy
- Fixed release
- 1.0.27
语音合成 Gemini 2.5 TTS
Generate multilingual, highly natural audio using Gemini 2.5 text-to-speech. 使用 Gemini 2.5 强大的文本转语音能力,生成多语言、高自然度的音频。
- Publisher
- dlazy
- Fixed release
- 1.3.23
录音转写 Fun ASR
Alibaba Bailian Fun-ASR recording transcription. Supports Chinese, English and other languages, with auto language detection and speaker diarization. Suitable for subtitles, transcription, and meeting notes. 阿里云百炼 Fun-ASR 录音文件识别,支持中英文及多语种,自动语种识别与说话人分离。适合字幕、转录与会议记录。
- Publisher
- dlazy
- Fixed release
- 1.3.24
解说视频生成 Explainer Video
explainer video, explainer video generator, animated explainer, training video — turn a document, topic, or brief into a narrated explainer video: outline, storyboard, voiceover, build, validate. Use when the user wants an explainer, courseware, or training video.
- Publisher
- dlazy
- Fixed release
- 1.0.21
语音转文字 ElevenLabs STT
ElevenLabs scribe_v1 speech-to-text with auto language detection and optional speaker diarization. Suitable for subtitles, transcription, and meeting notes. ElevenLabs scribe_v1 语音转文字,支持自动语种识别与说话人分离,适合字幕、转录与会议记录。
- Publisher
- dlazy
- Fixed release
- 1.3.24
声音克隆 ElevenLabs Voice Clone
ElevenLabs Instant Voice Cloning (IVC). Upload a clean voice sample to clone a custom voice usable with ElevenLabs TTS. ElevenLabs 即时音色克隆(IVC),上传一段干净人声样本即可复刻自定义音色,可用于 ElevenLabs TTS 配音。
- Publisher
- dlazy
- Fixed release
- 1.3.23
语音合成 ElevenLabs TTS
ElevenLabs eleven_v3 text-to-speech with 12 curated multilingual voices and stability/similarity/style controls. Great for dubbing, audiobooks, and character dialog. Before picking a voice, you can search for the right one via elevenlabs-search. ElevenLabs eleven_v3 文本转语音,提供 12 种精选英文/多语种音色,支持稳定性、相似度、风格控制。适合配音、有声内容与角色对话。选择音色前,可以从 elevenlabs-search 检索合适的音色。
- Publisher
- dlazy
- Fixed release
- 1.3.24
多人对话配音 ElevenLabs Dialogue
ElevenLabs eleven_v3 multi-voice dialogue: assign a different voice per line (up to 10) and render the whole conversation in one shot. Supports audio tags like [giggling], [whispers] — great for character dialogue, podcasts, and short skits. Before picking a voice, you can search for the right one via elevenlabs-search. ElevenLabs eleven_v3 多人对白合成:为每行台词指定不同音色(最多 10 个),一次性生成完整对话音频。支持 [giggling]、[whispers] 等情绪标签,适合角色对白、播客与短剧。选择音色前,可以从 elevenlabs-search 检索合适的音色。
- Publisher
- dlazy
- Fixed release
- 1.3.23
语音合成 豆包 Doubao TTS
Synthesize text into natural and fluent speech using Doubao TTS. 使用豆包 (Doubao) TTS 文本转语音模型,将文字合成为自然流畅的语音播报。
- Publisher
- dlazy
- Fixed release
- 1.3.23
文档转视频 Doc to Video
doc to video, word to video, markdown to video, document to video — parse the document, outline, storyboard, voiceover, build, validate. Use when the user gives a Doc / Word / Markdown file and wants an explainer, report broadcast, or training video.
- Publisher
- dlazy
- Fixed release
- 1.0.21
博客转视频 Blog to Video
blog to video, blog post to video, article to video — convert a blog post into a narrated video with storyboard, voiceover, and build. Use when the user gives a blog post (text or link) and wants a video version for social or YouTube.
- Publisher
- dlazy
- Fixed release
- 1.0.23
Pet Training Command Execution Recognition | 宠物训练指令执行识别(坐/卧/等)
Triggers when a user provides a training-area video of a pet for analysis; supports local uploads or network URLs to call server-side APIs for command-execution recognition, detecting whether the pet's body posture matches the issued commands (Sit / Down / Stay), comparing posture timing against command timestamps, and judging execution success. When the command is not executed, the result can trigger an external voice repeat-prompt signal (not a medical / behavior-therapy advice). Application scenarios: smart dog…
- Publisher
- smyx-skills
- Fixed release
- 1.0.13
Foleyix
Create and download audio with your Foleyix account
- Publisher
- yinchao.lv
- Fixed release
- 1.1.0
Pet Eating Speed Slow Feed Analysis | 宠物进食速度检测与慢食干预
Triggers when a user provides a video of the pet food-bowl area for analysis; supports local uploads or network URLs to call server-side APIs for eating-speed detection, recording start/end timestamps of feeding, estimating eating speed (g/s and seconds-per-bowl), and when the speed falls below the safety threshold (e.g. < 30 sec/bowl) emitting an intervention signal (slow-feed baffle pop-up or voice prompt) to prevent choking and vomiting (without diagnosing diseases). Application scenarios: smart slow-feeder bow…
- Publisher
- smyx-skills
- Fixed release
- 1.0.13
varg-ai
Generate AI videos, images, speech, and music using varg. Use when creating videos, animations, talking characters, slideshows, product showcases, social content, or single-asset generation. Supports zero-install cloud rendering (just API key + curl) and full local rendering (bun + ffmpeg). Triggers: "create a video", "generate video", "make a slideshow", "talking head", "product video", "generate image", "text to speech", "varg", "vargai", "render video", "lip sync", "captions".
- Publisher
- Alex
- Fixed release
- 2.0.10
Podcast Mention Memory
Indexes and transcribes podcast and radio audio to track who mentioned a brand, with show, episode, timestamp, speaker, and sentiment details.
- Publisher
- bluecolumnconsulting-lgtm
- Fixed release
- 1.0.0
Environmental Anomaly Trigger | 畜禽舍环境异常联动
Combines livestock behavior in continuous barn videos with environmental sensor data (temperature, humidity, ammonia, etc.) to identify group stress responses caused by abnormal in-barn conditions. | 结合畜禽行为与环境传感器,识别温湿度异常时的群体应激反应。
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
Podwise
Search, summarize and learn from podcasts.
- Publisher
- Xin Wu
- Fixed release
- 1.0.0
Fish Respiratory Rate (Gill Opening / Closing) Monitor | 鱼类呼吸频率(鳃盖开合)监测
Through fixed cameras on aquariums, the system analyzes fish gill-cover opening / closing motion video, detects periodic gill opening and closing, and calculates respiratory rate (breaths per minute). | 通过鱼缸固定摄像头,分析鱼类的鳃盖开合运动视频,检测鳃盖的周期性开启和闭合,计算呼吸频率(次/分钟)。当呼吸频率超过正常阈值(例如 > 80 次/分钟,具体依品种和水温而定)时,输出'缺氧预警',提示用户检查水质(溶氧量)、水温或鱼的健康状态。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.10
550W Watermark & Text Eraser
Erase image text, video subtitles and watermarks; resolve links
- Publisher
- SunshineHu
- Fixed release
- 3.1.3
Dietary Behavior Health Analyzer | 饮食行为健康分析工具
Analyzes videos to evaluate human eating behaviors, habits, and dietary patterns. It identifies tendencies towards unhealthy eating and provides structured analysis reports along with nutritional improvement recommendations. | 饮食行为健康分析工具,针对人的饮食行为、进食习惯、饮食结构进行视频分析,识别不良饮食行为倾向,提供结构化分析报告和营养改善建议
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.15
Living-Alone Sleep Rhythm Anomaly Analysis | 独居者作息规律异常分析
Using a fixed camera in the living room or bedroom of a person living alone, the system continuously analyzes night video (typically 22:00-06:00) to detect lights-off time (when light sources turn off) and early-morning activity (human movement or body motion between 0-6 AM). It builds a personal historical baseline (e.g., average lights-off time and early-morning activity frequency over the past 7-14 days). | 通过家庭客厅或卧室固定摄像头,夜间(通常指22:00-6:00)连续分析视频,检测熄灯时间(光源关闭的时刻)、凌晨活动(0-6点期间的人体移动或肢体动作)。建立个人历史基线(如过去7-14天的平均熄灯时间和凌晨…
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
Sound Event Memory
Indexes and recalls non-speech ambient audio events (glass break, alarms, dog barks) across devices with timestamps and confidence scores for incident review.
- Publisher
- bluecolumnconsulting-lgtm
- Fixed release
- 1.0.0
Child Window/Balcony Climbing Detection | 儿童攀爬窗户/阳台识别
Using fixed cameras in living rooms or child-activity areas (aimed at windows or balconies), AI pose estimation and object detection analyze the video in real time to recognize whether a child is climbing windows, leaning out, or gripping window-sill edges. | 通过家庭客厅或儿童活动区域的固定摄像头(需对准窗户或阳台),利用AI姿态估计和目标检测技术实时分析视频,识别儿童是否发生攀爬窗户、身体探出窗外、抓握窗台边缘等危险行为,当检测到危险动作时立即输出预警,联动手机APP推送警报或触发声光报警器。该技能可有效预防儿童坠楼事故。
- Publisher
- smyx-skills
- Fixed release
- 1.0.13
simplepractice
Read a SimplePractice Client Portal through the simplepractice-mcp server — upcoming appointments, invoices/statements/superbills/receipts, balance and saved cards, paperwork waiting to be signed, and practice announcements. Use when the user asks about their therapy or healthcare appointments, what they owe a practice, a superbill for insurance, or forms their provider has sent them.
- Publisher
- chrischall
- Fixed release
- 1.2.2
Bravado
Turn a project link into a choreographed motion film
- Publisher
- DKTR N9NE
- Fixed release
- 1.0.1
Feed Intake Estimation | 畜禽采食量估算
Estimates daily feed intake per livestock individual from continuous feeder videos by tracking the change of feed remaining in the trough, and outputs intake trend with anomaly alerts. | 通过食槽视频估算每日采食量变化,异常时预警。
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
smyx_infant_cry_cause_classification_analysis | 婴幼儿哭声原因分类
Using the built-in microphone of a baby monitor or smart camera to capture infant cry audio, AI acoustic analysis extracts cry features such as frequency, pitch, rhythm, and duration, and classifies the possible causes behind the cry (hunger, sleepiness, pain/discomfort, boredom/need for comfort, fear, etc.), outputting the most likely cause and its confidence. | 通过婴儿监护器或智能摄像头的内置麦克风采集婴儿哭声音频,利用AI声学分析技术提取哭声的频率、音调、节奏、持续时间等特征,分类识别婴儿哭声背后的可能原因(饥饿、困倦、疼痛/不适、无聊/需要安抚、恐惧等),输出最可能的原因类别及置信度。系统实时监测哭声,当检测到哭声时自动分析并在父母手机APP上推送结果(如'…
- Publisher
- smyx-skills
- Fixed release
- 1.0.13
used-car-walkaround
Turn one used-car condition sheet into a listing hero still and a speakable walkaround script, then turn that still into one used car walkaround clip. This used car walkaround studio makes a used-car listing still and a listing walkaround video from the facts you already have. Use it for dealer lot walkaround video, used-car listing video, and inventory walkaround clips.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
YouTube Thumbnail Maker
Create YouTube thumbnails from a video topic, title, script, key frame, portrait, product photo, or channel reference. This AI thumbnail maker compares three directions in text first, then renders the one you pick as a 16:9 image with clear visual hooks, readable hierarchy, and a headline-safe composition for explainers, reviews, tutorials, vlogs, games, podcasts, and long-form creative videos. It pairs the image with title-matching advice and refines an accepted direction into a consistent channel look. Rendering…
- Publisher
- beatra-ai
- Fixed release
- 0.2.6
Wrong Item Talking Clips
Turn a user-supplied wrong-item script and authorized stills into one wrong item talking clip per still. This mistake explanation talking video studio voices a line from the script for each photo, then animates that photo into a 2 to 15s wrong-item explanation talking clip, ready to share as a homework error talking clip. Use it for error explanation talking pack and wrong-item script talking clip work that stays one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
Wealth Product Talking Clips
Turn a user-supplied product factsheet and authorized stills into one wealth product talking clip per still. This product factsheet talking video studio writes a speakable product highlights talking clip for each photo, then animates a 2 to 15s product factsheet talking clip. Use it for wealth product talking pack and factsheet talking video pack work that stays one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
Voiceover & Narration Studio
Use Voiceover & Narration Studio as an AI voice generator, text-to-speech workspace, and AI voiceover generator. Choose from the current voice library, turn scripts into ready-to-edit AI narration and voiceover, or create and reuse a custom brand voice through voice cloning. It supports short-video voiceover, script-to-voiceover, course narration, ordered audiobook narration, supplied multilingual text to speech, Cantonese text to speech, and recurring brand audio, with current price estimates, clear output planni…
- Publisher
- beatra-ai
- Fixed release
- 0.2.2
AI Voice Cloning Studio
Create a reusable personal or brand voice from a clean audio sample with this AI voice cloning studio and voice cloning software. Clone my voice, build a custom AI voice, or create an AI voice clone from a short single-speaker sample; give the custom voice a memorable name and reuse it for narration, courses, product stories, customer updates, series, and brand content. Compare sample quality, review the current estimate, and hear the reusable voice in a short test reading before expanding it into longer spoken pr…
- Publisher
- beatra-ai
- Fixed release
- 0.2.4
Visitor Desk Voice Pack
Turn a written visitor reception script into one visitor reception voice clip per labeled cue. This front desk voice studio records each cue the desk already wrote, from reception script audio to visitor greeting voice lines, and delivers 8 to 20 clips as one visitor desk voice pack. Use it for visitor reception voice packs that keep one cue on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
video-generation-studio
Plan and create short AI videos from a written shot, a supplied image, exact first and last frames, multimodal references, or existing footage. Video Generation Studio supports text-to-video, image-to-video, reference-guided generation, AI video editing, and AI video extension for product videos, ad creative, social clips, b-roll, transitions, reveals, and cinematic concepts. Review each delivered clip for action, subject stability, camera, continuity, audio when requested or returned, and destination fit, then ch…
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
viral-video-teardown-remake
Turn a short video that already worked into your own version. Paste the link and this viral video teardown and short-video remake workflow reads the reference itself on TikTok, Douyin, Xiaohongshu, Instagram, YouTube, or X — caption, author, visible metrics, comments, and on YouTube the full transcript — or work from a file, screenshots, or your own description instead. It breaks the clip into its hook, body beats, and call to action, names the script pattern behind it, scores what carried the performance, then re…
- Publisher
- beatra-ai
- Fixed release
- 0.3.4
AI Video Realism Retoucher
Polish an existing AI-generated short video with a focused realism retouch. This focused AI video retouch refines artificial lighting, synthetic materials, oversaturated color, and repeated or distracting detail in one selected pass, while carrying forward the shot's subject, camera framing, timing, and intended mood. Use it for AI video cleanup, video retouching, natural-looking video polish, product clips, ad creative, social video, and short-form footage that needs a cleaner, more believable finish.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
TikTok Comment Reply Voice
Turn public TikTok comments into one spoken reply clip per written line. This TikTok comment reply voice studio reads the comments on a finished post, then records each TikTok comment reply from the reply lines you already wrote. Operators use it when they need TikTok comment replies as a comment reply voice pack, or spoken comment replies they can import one file at a time.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
talking-pet-video
Make a pet talk by animating one clear pet photo with a short message or prepared voice clip. This talking pet video and talking dog and cat generator workflow turns a single pet photo into a shareable talking-pet clip from one pet image and a short spoken line, and reviews breed and face identity, mouth motion, speech clarity, and synchronization. Use it for pet greetings, funny pet dialogue, pet reactions, pet stories, festive pet messages, and pet-creator content.
- Publisher
- beatra-ai
- Fixed release
- 0.2.1
Talking Avatar & AI Presenter Video
Create a talking avatar from one portrait and a short script or speech track. This AI presenter and digital human video workflow can prepare narration with a selected voice or use a supplied recording, then direct a stable talking-head clip with restrained expression, natural movement, clear delivery, and focused lip-sync review. Use it for AI spokesperson videos, product explainers, training, course lessons, announcements, onboarding, social talking-head content, and photo-to-talking-video messages, with narratio…
- Publisher
- beatra-ai
- Fixed release
- 0.2.4
short-drama-voice-pack
Turn a vertical short-drama episode script into a labeled short drama voiceover pack with one consistent voice per role. This short drama dialogue studio and vertical short drama voiceover shop casts the episode, records every spoken line as labeled clips, and keeps those character voices through the episode so editors can place AI short drama voice and multi-character voiceover without recasting. Use it for short drama voice acting, episode dialogue audio, vertical drama TTS, micro drama voiceover, and serialized…
- Publisher
- beatra-ai
- Fixed release
- 0.1.6
Shift Handoff Voice Pack
Turn a written shift handoff checklist into one shift handoff voice clip per labeled cue. This shift handover voice studio records each handover checklist audio and shift change voice from the list the desk already wrote, then delivers 8 to 20 shift handoff clip files. Use it for shift handoff voice packs that keep one cue on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
photo-singing-video
Make a photo sing by animating one clear portrait with a short singing audio excerpt. This photo singing video and AI singing portrait workflow turns a single face into an expressive singing clip from one portrait and a chosen song excerpt, and reviews identity, mouth and facial movement, performance energy, audio presence, and synchronization. Use it for old photo singing, birthday greetings, character art, playful posts, song promos, and memorable messages, with one portrait and one singing audio excerpt per run…
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
Spoken Seeding Video Maker
Make a spoken recommendation video from nothing but a topic. This talking-style seeding video maker and short video script generator picks the script pattern that fits your product or subject, writes the hook, the body beats, and the closing ask with the on-screen action and the spoken line written separately, then produces ready-to-edit still beat frames, a narration track in a voice you choose, an optional music bed, and one vertical clip animated from the opening frame with the full narration. Use it for produc…
- Publisher
- beatra-ai
- Fixed release
- 0.2.2
novel-promo-video-maker
Turn a novel chapter, web-novel excerpt, or story script into narrated vertical short video scenes with illustrated shots that keep every character looking the same from beat to beat. This AI story video maker pulls the hook and the beats out of your text, draws one shot for each, records the narration in a voice you choose, and sets each clip to its own narration, so a chapter arrives as an ordered set of scenes for story channels, book trailers, web-novel promotion, chapter recaps, and faceless storytelling acco…
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
Market Inspection Talking Clips
Turn user-supplied merchant inspection notices and authorized stills into one market inspection talking clip per still. For each photo it writes the spoken line from your notice and animates a 2 to 15s merchant notice talking clip. The result is a market inspection talking pack, a merchant notice talking video pack with one clip per photo.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
IVR Voice Pack
Build a labeled IVR voice pack for a phone tree: welcome, menu, hold, transfer, after-hours, and error prompts in one consistent brand voice. This phone tree voice studio and hotline voiceover shop records about twelve default prompts as a voice menu you can drop into the switch. Use it for IVR prompts, call center voice, auto attendant audio, and customer-service phone menus.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
Insurance Clause Talking Clips
Turn user-supplied insurance policy clauses and authorized stills into one insurance clause talking clip per still. This insurance clause talking video studio voices a line from the supplied clauses for each photo, then animates that photo into a 2 to 15s clause explanation talking clip. Use it for insurance clause talking packs that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
Pet Race Foul Detection (False Start / Lane Crossing) | 宠物赛跑/竞赛作弊识别(起跑/窜道)
Triggers when a user provides a pet racing track start/finish video URL or file for analysis; uses HD cameras at the starting line and finish line to analyze race video in real time, detecting each pet's (greyhounds, racehorses, etc.) start time, finish order, and lane assignment, automatically determining false starts (start before the signal) or lane crossing (deviating from own lane into an adjacent lane) fouls and outputting judgment results. Assists referee decisions and improves race fairness. Application: p…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.11
Geography Place Talking Clips
Turn a user-supplied geography place table and authorized stills into one geography place talking clip per still. This geography place talking video studio writes a speakable place-fact talking clip for each photo, then animates a 2 to 15s place highlight talking clip. Use it for geography place talking packs that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
game-script-voice-pack
Turn a murder-mystery or indie-game script into a labeled multi-character voice pack. This AI game voice pack and character dialogue studio casts one consistent voice per role, records every line in script order, and delivers numbered clips with character IDs ready for an engine or tabletop production. Use it for game voice acting, indie game dialogue, murder mystery voiceover, multi character TTS, NPC voice acting, interactive story audio, tabletop voice pack work, and scripted game audio.
- Publisher
- beatra-ai
- Fixed release
- 0.1.6
Fund Quarterly Report Talking Clips
Turn a user-supplied fund quarterly report highlight sheet and authorized stills into one fund quarterly report talking clip per still. This fund quarterly report talking video studio voices a line from the printed report facts and manager commentary for each photo, then animates that photo into a 2 to 15s quarterly report highlight talking clip. Use it for fund quarterly report talking packs that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
Fund Dividend Talking Clips
Turn a user-supplied fund dividend announcement and authorized stills into one fund dividend talking clip per still. This fund dividend talking video studio writes a speakable account-arrival talking clip for each photo, then animates a 2 to 15s dividend announcement talking clip. Use it for fund dividend talking packs that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
Earnings Script Reads
Turn an official earnings script into one spoken earnings script read per labeled section. This earnings call voice studio turns the prepared remarks the company already wrote into earnings report narration and quarterly results audio, recorded as 8 to 20 earnings script voice clips. Use it for investor update voice packs that stay one section, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
Douyin Comment FAQ Stills
Turn Douyin video comments into one comment FAQ still per picked question. This Douyin comment reply studio writes a listing FAQ card from each seller-answered comment, then lays out a 4 to 8 still comment FAQ set. Use it for listing FAQ graphics and a Douyin comment FAQ card that stay one question, one still.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
Douyin UGC Ad Creator
Create a Douyin shopping video, Douyin UGC ad, or AI creator product pitch from a product photo, product details, and an on-camera direction. This AI UGC ad creator composes a vertical presenter frame with the product, shapes a conversational hook and spoken recommendation, and delivers a short Douyin product video for launches, creator-style reviews, demonstrations, unboxings, and paid social creative.
- Publisher
- beatra-ai
- Fixed release
- 0.2.2
Douyin Comment Demo Clips
Turn Douyin comment objections into one comment talking clip per still. This douyin comment demo studio writes a speakable demo talking clip from each seller-picked objection, then animates a 2 to 15s objection demo clip. Use it for comment demo video and a douyin objection video that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
0bull-tiktok-post
Publish a video to TikTok from a real rented iPhone through the 0bull MCP server, then follow the submission to a final status.
- Publisher
- 0bull - iOS Phone Farm Remote Control
- Fixed release
- 1.0.0
Booking Confirmation Voice Pack
Turn confirmed booking facts into one booking confirmation voice clip per labeled notice. This appointment confirm audio studio writes speakable reschedule and cancel booking voice lines for a shop, then records 8 to 20 reservation confirm voice reads. Use it for shop booking voice packs and reschedule voice clips that stay one line, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
Creator Drop Talking Clips
Turn a confirmed new-drop script and authorized stills into one drop talking clip per still. This product drop video studio writes a speakable launch talking clip and new-drop announcement for each photo, then animates a 2 to 15s talking teaser. Use it for creator drop videos and talking teasers that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
course-video-studio
Turn finished lecture scripts into presenter-led course videos with lecture narration. This course video studio prepares the spoken narration for each lesson, then records a digital-teacher delivery, so enablement teams get ready-to-publish lecture videos and training presenter clips.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
course-narration-studio
Turn lesson materials into polished course narration audio. This course narration studio organizes lessons into speakable scripts, records the teacher-voice narration for each section, and delivers slide-ready audio for courses, training, and knowledge-product audio.
- Publisher
- beatra-ai
- Fixed release
- 0.1.7
Club Activity Talking Clips
Turn a user-supplied club-activity script and authorized stills into one club activity talking clip per still. This club notice talking video studio voices a line from the script for each photo, then animates that photo into a 2 to 15s event script talking clip. Use it for club activity talking pack, club signup talking clip, and activity notice talking pack work that stays one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
Beatra Universal
Create image, video, music, speech, and look up public social data with Beatra
- Publisher
- beatra-ai
- Fixed release
- 2.9.1
Beatra AI Voice Studio
Use Beatra AI Voice Studio as an AI voice generator, text-to-speech workspace, and AI voiceover generator. Choose from the current voice library, turn scripts into ready-to-edit AI narration and voiceover, or create and reuse a custom brand voice through voice cloning. It supports short-video voiceover, script-to-voiceover, course narration, ordered audiobook narration, supplied multilingual text to speech, Cantonese text to speech, and recurring brand audio, with current price estimates, clear output planning, an…
- Publisher
- beatra-ai
- Fixed release
- 0.3.0
AdsTurbo 广告视频复刻 · AdsTurbo Ad Video Clone
看中一条广告就照着做一条:自动拉片得出分镜与提示词,再生成同款结构的新视频。也可只做视频分析不生成。
- Publisher
- AdsTurbo
- Fixed release
- 1.2.1
AdsTurbo 视频精修 · AdsTurbo Video Cleanup
画面内容不变,只让成片更干净:去水印、去 logo、去硬字幕、提升分辨率到 2K/4K、添加或翻译字幕。
- Publisher
- AdsTurbo
- Fixed release
- 1.2.1
AdsTurbo 虚拟人口播 · AdsTurbo Digital Human Voiceover
让虚拟人把文案念出来:可选平台数字人形象、用照片+声音克隆专属形象,或直接给一张人像加一段音频做对口型。含 TTS 配音与形象管理。
- Publisher
- AdsTurbo
- Fixed release
- 1.2.1
Beatra AI Video Studio
Plan and create short AI videos from a written shot, a supplied image, exact first and last frames, multimodal references, or existing footage. Beatra AI Video Studio supports text-to-video, image-to-video, reference-guided generation, AI video editing, and AI video extension for product videos, ad creative, social clips, b-roll, transitions, reveals, and cinematic concepts. Review each delivered clip for action, subject stability, camera, continuity, audio when requested or returned, and destination fit, then cho…
- Publisher
- beatra-ai
- Fixed release
- 1.3.0
Homestay Welcome Talking Clips
Turn authorized host or room stills and confirmed stay facts into one homestay welcome talking clip per still. This homestay welcome talking-clip studio writes a speakable check-in and amenity script, then animates each still into a 2–15s welcome or facility clip in stay order. Use it for homestay welcome videos, Airbnb check-in greetings, amenity explainers, and guest arrival talking clips that stay one photo, one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
AI Video Continuation
Continue one short source clip naturally before or after its existing action from a continuity state and the next visual beat. This AI video continuation and generative extend workflow adds matching footage to the opening or ending of a clip — holding a moment longer, completing a reveal, adding a lead-in, or creating a natural ending — and reviews the seam, motion direction, subject identity, camera and lighting continuity, audio behavior, and final duration. Use it for stronger endings, longer holds, action cont…
- Publisher
- beatra-ai
- Fixed release
- 0.2.1
AI Music Video Clip Maker
Create a short visual clip guided by a song's mood, rhythm, and visual concept. This AI music video clip maker and song-to-video generator turns a music excerpt and visual direction into a cinematic music promo clip, animates approved cover art or a portrait in time with the music, interpolates motion between opening and ending art, or uses audio as a loose mood reference for a new visual concept. Use it for new-song teasers, album promo clips, cover art animation, mood visuals, virtual performer scenes, and socia…
- Publisher
- beatra-ai
- Fixed release
- 0.1.9
AI Comic Drama Shot Maker
Turn a comic panel, character sheet, webtoon frame, or frozen story beat into one dynamic comic-drama shot. This AI comic drama generator and motion comic maker animates an approved comic panel or manga frame into a short live shot, interpolates motion between a first and last comic panel, combines loose character, style, and scene references into a new comic-drama shot, or creates an original comic first frame and brings it to life. Use it for web-novel comic shots, motion comics, webtoon to video, manga panel an…
- Publisher
- beatra-ai
- Fixed release
- 0.2.0
AI Audiobook Narration
Turn final manuscript or course text into an audiobook with one consistent narrator. This AI audiobook generator and AI audiobook narration workflow organizes chapters, settles names and specialist pronunciations, shapes long-form pacing, helps choose a suitable audiobook narrator, and creates a representative sample before the remaining book. Use it for manuscript-to-audiobook production, chapter-by-chapter audiobooks, course narration, long-text-to-speech, and sample chapters, with up-to-date price estimates, or…
- Publisher
- beatra-ai
- Fixed release
- 0.2.5
Fish Abnormal Swimming Posture (Side-swim / Upside-down) Detection | 鱼类游动姿态异常(侧游/倒立)识别
Through fixed cameras on aquariums, the system analyzes fish swimming videos and computes the angle between the fish body axis and the horizontal plane (normal fish bodies stay nearly horizontal). | 通过鱼缸固定摄像头,分析鱼类的游动视频,检测鱼体轴线与水平面的夹角(正常鱼体基本保持水平),当鱼体倾斜角度超过阈值(默认 > 30°)或出现倒立(头部向下 > 45°)、旋转(绕自身纵轴连续翻转)等异常游姿时,标记为异常,并记录异常时长占观察总时长的比例。该技能有助于早期发现鱼鳔失调、神经系统疾病或水质中毒等健康问题,提醒养鱼爱好者及时干预。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.11
guaikei语音转文字
将视频转为文字与结构化文案的技能。当用户提出"视频转文字 / 视频提取文案 / 视频转稿 / 字幕提取 / 视频总结 / 视频内容分析 / 会议纪要 / 课程拆解 / 直播复盘 / 采访整理 / 短视频二创脚本 / 口播稿 / 小红书文案 / 抖音文案 / 公众号文案"等需求时使用。支持本地视频文件与抖音、小红书等平台视频链接,调用千问大模型自动转写并剔除语气词、口误与重复内容,并可通过自定义 Prompt 生成总结、改写、金句提取、分镜头、中英翻译等风格化文案。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
Sick Poultry/Swine Behavior Detection | 病鸡/病猪行为识别
Detects morbid behavioral cues in poultry and pigs from continuous barn videos — such as difficulty standing, ruffled feathers/piloerection, isolation, drowsiness and appetite loss — and outputs behavior type with risk level up to 2-3 days ahead of visible clinical signs. | 识别站立困难、羽毛蓬松、离群、嗜睡等病态行为,比人工观察提前2-3天。
- Publisher
- smyx-skills
- Fixed release
- 1.0.8
Fish Flashing & Scraping Detection (Ectoparasite Warning) | 爬宠体温调节行为识别(晒点/躲避)
Through fixed enclosure cameras, the system analyzes behavior videos of reptiles (lizards, snakes, turtles) and detects movement frequency and dwell duration between the basking zone (heated area under the basking lamp) and the hiding zone (cave/cool side). | 通过爬宠箱固定摄像头,分析爬行动物(如蜥蜴、蛇、龟)的行为视频,检测宠物在晒点(加热灯下方高温区域)与躲避区(洞穴、冷区)之间的移动频次、停留时长以及活动节律。系统连续监测,生成每日温区利用报告,异常时推送提醒。
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
Autism Stereotyped Behavior Detection (Spinning / Hand-Flapping) | 自闭症儿童刻板行为识别(转圈/摆手)
Using a fixed camera in rehabilitation centers or homes, the system analyzes children's behavior videos with pose estimation and temporal action detection to recognize repetitive stereotyped behaviors, including spinning (body rotation ≥ 360°), hand flapping (non-functional repetitive arm movement), body rocking (rhythmic forward-backward or side-to-side trunk motion), etc. | 通过康复机构或家庭固定摄像头,分析儿童行为视频,利用姿态估计和时序动作检测技术识别重复性刻板动作,包括转圈(身体旋转360°以上)、摆手(手臂非功能性重复摆动)、摇晃(躯干前后或左右有节律摆动)等。该技能可辅助康复师和家长客观记录行为变化,评估干预效果。
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
music-video
Use when someone wants a full music video — original song or vocals, performance clips, B-roll, and lyric-synced edits.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
narrated-multi-scene
Use when someone wants a multi-part story with voiceover — episodic B-roll, chaptered promo, or several linked video scenes without on-camera dialogue.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
interactive-explainer
Use when someone wants an educational explainer with a host and characters — history or science shorts with dialogue, not voiceover-only B-roll.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
image-to-video
Use when someone wants one short film beat from images — a narrated scene, story moment, or cinematic B-roll with optional voiceover.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
whisperx
Use when someone needs word-level timestamps from audio — lyric alignment, cut-safe line boundaries, or caption source timing before burn-in with video-editing.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
stable-audio-2.5
Use when someone wants light instrumental background music — an ambient bed under dialogue or underscore for reels and explainers.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
gemini-3.1-flash-tts
Use when someone needs spoken narration or voiceover — explainer tracks, documentary lines, or voice to pair with generated video.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
music-2.5
Use when someone wants an original AI song with vocals — sung lyrics, a style prompt track, or source audio for a music video.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
p-video-edit
Use when someone wants to edit an existing video with a text instruction — recolor, restyle, remove or add objects, change environment or lighting, update on-screen text, or apply optional reference-guided product and accessory edits. Not for a new clip from scratch or ffmpeg assembly.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
p-video-replace
Use when someone wants to swap a person, outfit, or product inside existing footage while keeping the camera move and audio.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
p-video-avatar
Use when someone wants a person on camera speaking a script — lip-synced host, spokesperson, or narrated avatar from a portrait photo.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
video-editing
Use when assembling or polishing already-rendered clips with ffmpeg — concat, crossfades, burned captions and subtitles, text/logo overlays, before/after sliders, background music beds, platform export — or when composing a multi-layer HTML combination video with Hyperframes. Not for AI video generation, prompt craft, or model-based video edits.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
p-video
Use when someone wants a simple short clip from text or images — quick B-roll, drafts, or start/end frame animation. Not when the brief needs cinematic generation, highest quality, tight lip-sync, or imported audio at 1080p.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
p-video-2
Use when someone wants a polished short clip from text, images, or imported audio — 1080p B-roll, start/end frame animation, or a motion shot with a mixed track. Not for cinematic generated-audio clips or talking-head-only hosts.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
p-video-2-pro
Use when someone wants a cinematic clip from text or start/end frames — product ads, documentary shots, or dialogue with generated audio. Not for 1080p, imported audio tracks, or talking-head-only hosts.
- Publisher
- Pruna AI
- Fixed release
- 1.0.14
WorkRally
WorkRally CLI (workrally) — 面向 AI Agent 的 AIGC 漫剧视频创作全流程工具集。支持 AI 生图、AI 生视频、视频提示词优化、画布生音频/音乐、混元 3D 模型生成、AI 生音频、项目/剧集/场次/分镜的完整 CRUD、资产库、媒资管理、无限画布、文件上传下载等。Use when user asks to generate images, generate videos, generate audio, generate music, generate 3d, optimize video prompts, manage projects, series, shots, upload files, download assets, manage materials, or interact with WorkRally platform via command line.
- Publisher
- 腾讯开源
- Fixed release
- 2.10.1
Outdoor Sports Event Risk Analysis Tool | 户外体育赛事风险分析工具
Conducts video safety risk analysis for participants in outdoor sports competitions, long-distance running, marathons, etc.; identifies sports injuries and sudden health risks, outputs professional analysis reports, and provides timely warnings to ensure sports safety. | 户外体育赛事风险分析工具,针对户外体育比赛、长跑马拉松等运动项目的参赛人员进行视频安全风险分析,识别运动损伤和突发健康风险,输出专业分析报告,及时预警保障运动安全
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.15
ElevenLabs
ElevenLabs API integration with managed authentication. AI-powered text-to-speech, voice cloning, sound effects, and audio processing. Use this skill when users want to generate speech from text, clone voices, create sound effects, or process audio. For other third party apps, use the api-gateway skill (https://clawhub.ai/byungkyu/api-gateway). Calls run through the `maton` CLI with OAuth login, or over raw HTTP with a Maton API key where the CLI cannot be installed. Every call is authenticated as the user's conne…
- Publisher
- byungkyu
- Fixed release
- 1.2.6
Child Separation Anxiety Detection (Pre-School Crying) | 儿童分离焦虑识别(上学前哭闹)
Using a fixed camera at the home entrance or kindergarten gate, the system analyzes pre-school videos and detects crying facial expressions (frowning, open-mouth crying, tearing), physical clinging actions (grabbing parent's clothes, hugging parent's leg, pulling door frame), and resistance behaviors (stepping back, lying on the ground), then comprehensively evaluates the separation-anxiety level (mild / moderate / severe). | 通过家庭或幼儿园门口固定摄像头,分析儿童上学前的视频,检测哭闹面部表情(皱眉、张嘴哭泣、流泪)、肢体抓拽动作(抓住家长衣服、抱住家长腿、拉扯门框)以及抗拒行为(后退、躺地)等,综…
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
Child Restless Sleep / Nightmare Detection | 儿童睡眠中频繁翻身/噩梦识别
Using a fixed camera in the child's bedroom (infrared night vision), the system continuously captures video and audio at night to analyze the child's sleep behavior. It detects rollover frequency (rollovers per minute), cries (recognizing specific cry-sound features), and sleep talk (speech during sleep), and generates a sleep-quality report. When rollovers occur too often (e.g., > 3 per hour), strong crying is detected, or sleep talk is observed, the system pushes 'possible nightmare' or 'restless sleep' alerts t…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.11
media-use
Agent Media OS, the single skill for every media need in a HyperFrames project. Resolve BGM, SFX, image, icon, brand logo, voice, color grade, or LUT into a frozen local file or paste-ready block + ledger record (one verb, `resolve`); generate via TTS / music / image models when the catalog misses; produce voiceover, transcription, captions, and background removal through one shared audio engine; operate on media (cut / reframe / transform); and reuse assets across projects. Also use for vague feedback that real f…
- Publisher
- HeyGen
- Fixed release
- 1.0.63
Pregnancy Emotion Soothing | 孕妇情绪波动舒缓
Through fixed cameras (and optional microphones) at the pregnant woman's home or prenatal exam waiting room, the system analyzes facial expressions (sudden crying, frowning, anxiety), prolonged silent sitting (≥ 30 consecutive minutes without social interaction or activity), and tone of conversations with family members (rapid, impatient). | 通过孕妇家中或产检候诊室的固定摄像头(及可选麦克风),分析孕妇的面部表情(突然哭泣、皱眉、焦虑)、长时间静坐不语(连续超过30分钟无社交互动或活动)、以及与家人对话的语气(急促、不耐烦)。当检测到显著情绪波动时,自动触发安抚动作:通过智能音箱播放孕期舒缓音乐或正念引导音频,或向丈夫手机APP推送提醒('妻子情绪波动,请打电话关心')。
- Publisher
- smyx-skills
- Fixed release
- 1.0.13
Wiki Capture (Dexio)
Capture what an agent session settled into an LLM wiki before it is lost: at the end of a task or session (or over saved transcripts), pick out decisions, verified facts, corrections and working procedures, drop the chatter, and file each item on the page that owns it. Use when a session or task ends, before context is compacted, when a person says to remember something, or on a schedule over past agent sessions (Claude Code, Codex, Hermes, OpenClaw).
- Publisher
- Dexio
- Fixed release
- 1.1.0
video-watermark-anti-theft
视频水印与防盗用处理技能,用于回答「视频怎么防被搬运」「被盗用了怎么证明是我的」这类问题
- Publisher
- zhaoxinghua09-cell
- Fixed release
- 1.0.4
invoice-risk-scan
发票要素风险扫描 — 报销/入账前扫发票要素:税号格式、金额、日期合理性,异常即拦(零依赖)
- Publisher
- zhaoxinghua09-cell
- Fixed release
- 1.0.3
Child Bedtime Soothing (Fear of Dark / Post-Nightmare) | 儿童睡前情绪安抚(怕黑/噩梦后)
Through a fixed camera (with infrared night vision) and microphone in the child's bedroom, the system analyzes pre-sleep and night-time video and audio to detect pre-sleep crying (continuous crying, calling 'Mama'), fear-of-the-dark expressions (curling up, looking around), and nightmare awakenings (sudden sitting up, trembling, screaming). | 通过儿童卧室的固定摄像头(红外夜视)及麦克风,分析儿童睡前及夜间视频,检测睡前哭闹(持续性哭声、呼喊'妈妈')、怕黑表现(身体蜷缩、四处张望)、噩梦惊醒(突然坐起、颤抖、尖叫)等行为。当检测到上述情绪不安时,自动触发安抚动作:开启小夜灯(柔光)、播放预先录制的妈妈讲故事音频或轻柔摇篮曲。
- Publisher
- smyx-skills
- Fixed release
- 1.0.13
hyperframes
Create video compositions, animations, title cards, overlays, captions, voiceovers, audio-reactive visuals, and scene transitions in HyperFrames HTML. Use when asked to build any HTML-based video content, add captions or subtitles synced to audio, generate text-to-speech narration, create audio-reactive animation (beat sync, glow, pulse driven by music), add animated text highlighting (marker sweeps, hand-drawn circles, burst lines, scribble, sketchout), or add transitions between scenes (crossfades, wipes, reveal…
- Publisher
- OpenAI
- Fixed release
- git-5fd93af4cd0c
narrator
Activate when the user explicitly names the narrator skill or requests production-ready narration in either case: (1) numbered takes fitted to fixed video windows, or (2) one continuous long-form story read in a locked voice that must be measured and retried for timbre or internal pauses. Collect missing text, voice_id, or voice_type after activation; their absence is not a reason to skip this workflow. Own pacing, speech-duration gates, retries, and ready audio delivery; never time-stretch. Do not use for ordinar…
- Publisher
- OpenAI
- Fixed release
- git-5fd93af4cd0c
subtitles
Burn timed subtitles onto a finished video or configure Whisper-timed caption burning during faceless-video assembly. Takes video/audio generation jobs or a local finished video, optionally with authored narration text, and returns a captioned video or the exact backend subtitle configuration. Timings always come from Whisper on the video's own audio — never estimated. THREE looks: `paper` (torn cream paper label, handwritten ink), `bold` (UGC ALL-CAPS, white with black stroke, platform safe zones), `clean` (slim…
- Publisher
- OpenAI
- Fixed release
- git-5fd93af4cd0c
Reptile Feeding Refusal / Vomiting Detection | 爬宠进食拒绝/呕吐识别
Through fixed enclosure cameras, the system analyzes feeding-time and post-feeding videos of reptiles (snakes, lizards, turtles) to detect prey-attack behavior, successful swallowing, and regurgitation (vomiting). | 通过爬宠箱固定摄像头,分析喂食时及喂食后一段时间的视频,检测爬行动物(如蛇、蜥蜴、龟)的进食行为:是否主动攻击猎物(如鼠、昆虫)、是否成功吞食、以及是否在进食后短时间内将食物吐出(反吐)。当宠物对猎物无视、逃避(拒食)或将已吞入的食物吐出时,记录异常事件并输出提示。
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
Video to Text (ZH)
提取视频语音并转成中文文字稿。当用户给出抖音、B站(bilibili/b23.tv)、YouTube、快手、视频号等常见视频网站的链接,想要"提取视频里说话的内容""视频转文字""语音转文字""视频字幕提取"时,必须使用本技能。即使用户只发了一个视频链接没说要做什么,也应主动用本技能识别并转写其中的说话内容。
- Publisher
- Luo4507
- Fixed release
- 1.0.0
Vibbit Skills
Create e-commerce content with Vibbit, from product research and creative planning to finished videos. Generate images, voiceovers, music and sound effects. Produce digital-human videos and short dramas, adapt references, or replace people, products and scenes. Translate and dub videos with subtitles and optional lip sync. Reuse assets, review results and make targeted revisions, with task tracking and recovery for interrupted work.
- Publisher
- zGissing
- Fixed release
- 2.14.2
HiFi Review
Objective, source-traceable evaluation of headphones/IEMs & DAC/amp gear
- Publisher
- VincentJiang06
- Fixed release
- 1.1.1
logic-pacer
Rewrite EXISTING admired Chinese/English expository prose so its reasoning is easier to follow — shrink the inferential STEP SIZE and re-anchor each step on ground the reader already holds (given-new), while KEEPING the voice, the vocabulary (never 对齐词汇), the facts/claims/stance, and staying lean (net length <= ~1.3x). Method: detect >=2-move leaps, unfold each into its minimal chain, subtract ornament. Use for "这段逻辑跳太快,放慢但别动文风/词汇", "reduce the inferential step size", "$logic-pacer". ABSTAIN if the prose is alread…
- Publisher
- VincentJiang06
- Fixed release
- 1.1.0
album-review
Deep, source-traceable long-form Chinese album review (乐评). Use when the user names a music credit (artist/composer/band) + an album and wants one comprehensive critique. Triggers: "写一篇深度乐评", "全面评测这张专辑", "$album-review". NOT for audio-gear evaluation (→ hifi-review).
- Publisher
- VincentJiang06
- Fixed release
- 0.3.0
embedded-captions
Add captions or subtitles to an existing single-subject talking-head video without editing the footage. Use for plain verbatim captions, cinematic captions embedded behind the subject, VFX captions, “炸/特效/酷炫字幕,” or a named identity from the 35-style catalog. Route by visual identity, not by backend engine. The quiet `anchor` rail is the default; embed every word only when the user explicitly wants a fully cinematic treatment. The workflow runs locally end to end, including transcription and subject matting; split…
- Publisher
- HeyGen
- Fixed release
- 1.0.20
AI KEY·口播剪辑
AI KEY·口播剪辑 ——口播成片剪辑技能。把拍好的素材剪成可发布的成片:**按文案规则做内容层重组**(删废镜头/去重复/重排顺序)→ 语音剪辑(去口癖/停顿)→ 加速 → 字幕 → B-roll → 交付。 🔴 **只删和重排,不加词** —— 说话人没说过的话一个字都不加。 中文口播的剪辑工作流(判据按中文口播写;用户用别的语言时照常用他的语言交流,并说明规则是按中文口播定的)。 默认参数(抖音 · 保持原长 · 1.15× · HarmonyOS Sans 粗体字幕 · 黄色 #FFE20A 高亮)属于澳洲AI教父账号:**确认是这个账号**才直接用;其他账号先读 `config`,读不到就问,不套用。 检测不到 ChatCut 会引导安装,并给出**本机转写旁路** —— 内容层的活全部不需要 ChatCut。 用户要 B-roll / 特效 / 转场而手上没有真素材时,可按文案逐句判哪几句能配生成画面(比喻、过程、氛围、转场可以;「我做过」「真实发生」这类证言绝不),写成 Seedance 提示词交给 chatcut-video-gen 出片,提交前先确认花费,发布时提醒打开 AI 生成声明。 触发方式:/…
- Publisher
- 詹明明
- Fixed release
- 0.4.5
talking-head-recut
Package an existing talking-head / interview / podcast video with timed, designed GRAPHIC OVERLAY cards — kinetic titles, lower-thirds, data callouts, quotes, side panels, picture-in-picture — synced to the transcript, on a 16:9 / 9:16 / 4:5 canvas of your choice; the clip plays untouched underneath. Trigger on "graphic overlays", "on-screen graphics", "package / dress up my video". Not plain subtitles (/embedded-captions). Unclear → /hyperframes.
- Publisher
- HeyGen
- Fixed release
- 1.0.12
music-to-video
Turn a music track (an audio file, a video to pull audio from, or a track generated from a mood brief) into a beat-synced video — lyric video, slideshow, or kinetic promo. The music drives all pacing; any user-supplied images/videos are cut onto the same beat grid, and a complete video needs zero assets. Narrated pieces → the input-matched workflow (see /hyperframes). Unclear → /hyperframes.
- Publisher
- HeyGen
- Fixed release
- 1.0.19
HyperFrames Keyframes
Author seek-safe video keyframes
- Publisher
- HeyGen
- Fixed release
- 1.0.7
hyperframes-audio
Use when audio already placed in a HyperFrames composition needs to be mixed: fade-in/fade-out, crossfade, track gain or volume, volume automation, ducking, a music bed that fights a voiceover (voiceover carve), effects on a track (EQ, compressor, limiter, gate, saturation, delay, reverb, chorus, phaser, bitcrush), automation envelopes drawn on a track's volume or any effect parameter, or one submix bus carrying a chain, a fader and an automation clock for several tracks at once (`<hf-audio-group>`). Don't use for…
- Publisher
- HeyGen
- Fixed release
- 1.0.12
sogni-creative-agent-skill
CLI and agent skill for personal LoRAs, image, video, and music generation using Sogni AI's decentralized GPU network. Supports Pixal3D single- and multi-view image-to-GLB, BiRefNet background removal, Qwen3-TTS speech/voice cloning/design, MiniMax Music 3, SAM 3 object selection, promptless RTX VSR image upscaling through 16K, promptless FlashVSR video upscaling to 1080p/1440p, one-click image-folder loop reels, personas (named people with saved reference photos and voice clips), persistent memories, custom perso…
- Publisher
- Sogni AI
- Fixed release
- 3.54.3
Elderly Gait Instability / Shuffling Step Detection | 老年人步态不稳/小碎步识别
Using a fixed camera in a hallway or living room to record video of an elderly person walking in a straight line, AI pose estimation and gait analysis extract parameters such as step length (cm), gait speed (m/s), trunk sway angle (left-right tilt), and cadence to evaluate gait stability. | 通过走廊或客厅的固定摄像头拍摄老年人直线行走的视频,利用AI姿态估计和步态分析技术检测步幅长度(cm)、步速(m/s)、躯干摇摆角度(左右倾斜度)以及步频等参数,评估步态稳定性。当步幅过小(小碎步)、步速过慢、躯干摇摆幅度过大时,输出跌倒风险等级(低/中/高)。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.11
Fish Feeding Behavior Activity Analysis | 鱼类摄食行为活跃度分析
Through built-in cameras of smart feeders or fixed cameras on aquariums, the system captures fish feeding videos after feeding. Using AI object detection and motion analysis, it identifies the number of fish gathering for food, feeding intensity (fish swimming speed, feeding action frequency), and remaining feed amount, and computes a comprehensive feeding activity score (0-100). | 通过智能喂食器内置摄像头或鱼缸固定摄像头,在投喂后拍摄鱼群摄食视频,利用 AI 目标检测和运动分析技术,识别鱼群聚集抢食的数量、摄食强度(鱼只游动速度、摄食动作频率)以及剩余饲料量,综合计算摄食活跃度评分(0-100 分)。当活跃度评分低于阈值时,输出'食欲下降'提示…
- Publisher
- smyx-skills
- Fixed release
- 1.0.16
Body Size/Weight Estimation | 畜禽体长/体重估测
Estimates livestock body length and body weight from side-view videos or frames, tracking fattening progress in a contactless manner. | 通过视频视觉估测体长、体重,追踪育肥进度。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.12
loudkit-voice
Set up free, open-source Loudkit voice notes for an existing agent and messenger. Use to check compatibility, run the companion on an Apple Silicon Mac, and configure Hermes, OpenClaw or another agent without replacing its session. The companion is a developer preview.
- Publisher
- pepinu
- Fixed release
- 0.1.1
sogni-creative-agent-skill
CLI and agent skill for personal LoRAs, image, video, and music generation using Sogni AI's decentralized GPU network. Supports Pixal3D single- and multi-view image-to-GLB, BiRefNet background removal, Qwen3-TTS speech/voice cloning/design, MiniMax Music 3, SAM 3 object selection, promptless RTX VSR image upscaling through 16K, promptless FlashVSR video upscaling to 1080p/1440p, one-click image-folder loop reels, personas (named people with saved reference photos and voice clips), persistent memories, custom perso…
- Publisher
- Mark Ledford
- Fixed release
- 3.54.0
aigate
Self-hosted AI platform — one `docker-compose up`, one OpenAI-compatible endpoint at http://localhost:4000. Bundles inference (Groq/Cerebras/OpenRouter/HuggingFace/Mistral/Cohere/Ollama/vLLM/llama.cpp/claudebox/pibox-zai/Anthropic/OpenAI), MCP tool use, a stealth browser cluster, image generation (FLUX/DALL-E/SD), speech synthesis (Kokoro/Qwen3-TTS/Chatterbox/OpenAI TTS), transcription (Whisper/Parakeet), S3-compatible object storage, agentic code execution (Claude Code + pi-coding-agent + sandboxed piston), web s…
- Publisher
- Ciprian Mandache
- Fixed release
- 6.0.0
heirloom-provenance
Use when someone dies or downsizes and the family objects need their stories captured before the people who know them are gone — structured provenance registry for heirlooms (origin, owner chain, story, appraisal, photo refs, audio-interview transcripts), provenance question scripts for interviewing elderly relatives before it is too late, estate-ready provenance reports for executors and appraisers, printed heirloom labels/tags with QR codes linking to the full record, and fairness tools for distribution (wish-li…
- Publisher
- voronindenis5
- Fixed release
- 1.0.0
Phosor AI
Generate AI videos, images and speech (text-to-video, image-to-video, reference-to-video, speech-to-video, animate, text-to-image, image-to-image, image edit, text-to-speech), bring your own LoRA models, and generate AI product/model photography for e-commerce (Image Studio) via the Phosor AI platform. Use when the user wants to create videos or images from text prompts, animate images, generate lip-synced video from audio, synthesize speech from text, generate images with a custom LoRA, generate product photograp…
- Publisher
- Phosor AI
- Fixed release
- 1.3.2
Aggressive Behavior Detection | 畜禽争斗行为识别
Detects aggressive interactions in livestock and poultry from continuous barn videos — including fighting, biting, chasing and butting — and outputs behavior type, intensity level and alert level. | 识别打斗、撕咬等攻击行为,及时预警。
- Publisher
- smyx-skills
- Fixed release
- 1.0.11
Driver Blink-Rate & Eye-Closure Fatigue Detection | 驾驶员眨眼频率与闭眼时长检测
Using an in-cabin DMS camera, the system analyzes the driver's facial video in real time, detects eye open/closed state, calculates blink rate per minute (normal range 15-20 blinks/min), and identifies single-blink closure duration. | 通过车载DMS摄像头实时分析驾驶员面部视频,检测眼部开闭状态,计算每分钟眨眼频率(正常约为15-20次/分钟),并识别闭眼持续时间。当眨眼频率异常降低(如<10次/分钟)或出现单次闭眼超过2秒(微睡眠前兆)时,输出疲劳驾驶预警,联动车内语音提醒或震动座椅,预防因疲劳导致的事故。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.10
Elderly Hand Resting-Tremor Detection | 老年人手部震颤(静止性)识别
Using a fixed home camera to record video of an elderly person's hand at rest (placed on a table or armrest with no voluntary movement), AI video-motion analysis detects periodic shaking, extracts tremor frequency (Hz) and amplitude (pixel displacement), and identifies the presence of resting tremor (commonly associated with Parkinson's disease and other neurological conditions). | 通过家庭固定摄像头拍摄老年人手部(置于桌面或自然静止)的视频,利用AI视频分析技术检测手部在静止状态下的周期性抖动频率(Hz)和幅度(像素位移),识别是否存在静止性震颤(常见于帕金森病等神经系统疾病)。该技能可作为早期筛查工具,提示家属或护理人员关注老年人神经系统健康…
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
farsight-intro-sequence
Generate 100% original Farsight commercial video intros by analyzing structural elements of inspiration videos while strictly avoiding copyright infringement.
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
Arrhythmia Early Warning Analysis Tool | 心律失常早期预警分析工具
Based on facial video, identifies abnormal rhythms such as premature beats, atrial fibrillation, tachycardia/bradycardia, assists in early detection of heart health risks. | 心律失常早期预警技能,基于面部视频识别早搏、房颤、心动过速/心动过缓等异常节律,辅助心脏健康风险早发现
- Publisher
- smyx-skills
- Fixed release
- 1.0.18
gemini-mcp
Generate and edit images, video, and music with Google Gemini models via MCP. Use when the user asks to generate, create, or edit images (Gemini / Nano Banana), produce a consistent set of images, compose/blend multiple images, generate a short video (text→video or image→video, via the omni model), or generate music/audio clips (via Lyria). Triggers on phrases like "generate an image of", "edit this image with Gemini", "create a set of consistent images", "make a video of", "generate a video", "generate music", "m…
- Publisher
- chrischall
- Fixed release
- 2.3.0
Uydi Voice Design & Clone
Uydi Voice enables an AI agent to design custom voices, clone a user's authorized voice sample, synthesize narration, create expressive sound scenes, and produce multi-voice Voice Canvas projects with the Uydi voice platform (https://uydi.com). Use it when a user asks to create or describe a voice, clone their own voice from a recording, convert text to speech, generate narration audio, create dialogue, sound effects, ambience or music; arrange multiple speakers or languages into one audio file; list or manage Uyd…
- Publisher
- yinchao.lv
- Fixed release
- 1.2.0
三剪客 · a7w 全量协议桥
一行 base_url,让上百个开源 AI 项目瞬间用上算力集市 api.a7w.cn 的全部能力——75 个大模型 + 21 个生成应用(TTS / 音色克隆 / 语音转文字 / 出图 / 数字人 / 视频生成 / 超分 / 换装 / 音乐)。自动补齐 OpenAI 协议的语音与图像端点,把 ASR 的字级时间戳直接输出成 SRT / VTT / verbose_json 字幕;Dify、n8n、MoneyPrinterTurbo、pyvideotrans、VideoCaptioner、ComfyUI 等无需改一行代码即可接入。本地零依赖运行,一行命令启动,实测 12/12 通过。遇到问题可加技术微信 9872659。
- Publisher
- 0人公司agent
- Fixed release
- 1.1.1
Pet Litter Box Usage Monitor (Frequency & Duration) | 宠物猫砂盆使用频次与时长
Triggers when a user provides a litter-box area video URL or file for analysis; uses object detection and tracking to identify each cat's entry/exit times at the litter box, records daily usage frequency and per-visit duration (entry → exit), and compares against the historical baseline. When frequency rises/falls significantly, or per-visit duration is abnormally long/short, outputs a urinary-disease (cystitis, urinary obstruction, kidney disease) early-warning alert. Especially useful for individualized health m…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.10
Child Social Interaction Frequency & Duration Analysis | 儿童社交互动频次与时长分析
Using fixed cameras in kindergartens or early-education centers, the system analyzes multi-person video to detect social-interaction behaviors among children, including approach (distance < 1 m), conversation (face-to-face with mouth movement), and cooperative play (collaborative play, chasing, etc.). | 通过幼儿园或早教中心的固定摄像头,分析多人视频,检测儿童之间的社交互动行为,包括接近(距离<1米)、对话(面对面且嘴部运动)、共同游戏(合作玩耍、追逐等)。系统定期生成社交互动热力图,为教师提供参考。
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
RMBase
Query RMBase v3.0 RNA modification data with provenance
- Publisher
- leo-itlger
- Fixed release
- 0.1.0
Fish Flashing & Scraping Detection (Ectoparasite Warning) | 鱼类擦缸/蹭底行为识别(外寄)
Through fixed aquarium cameras, the system analyzes fish behavior videos and detects abnormal frictional actions between fish bodies and tank walls, substrate, or rockwork — 'flashing' (fish flipping sideways and brushing tank walls rapidly) and 'scraping' (fish belly/flank rubbing on substrate). The system counts abnormal contact frequency per minute. | 通过鱼缸固定摄像头,分析鱼类的行为视频,检测鱼体与缸壁、底砂、造景石等物体的异常摩擦动作(擦缸:鱼体侧身快速蹭过缸壁;蹭底:鱼体腹部或侧面贴底砂摩擦)。统计每分钟的异常接触频次,当频次超过阈值(默认 5 次/分钟)且持续时间超过 10 秒时,输出'外寄风险提示',提醒用户检查是否有寄生虫(如小瓜虫、车轮虫、三代虫)感染或皮…
- Publisher
- smyx-skills
- Fixed release
- 1.0.12
yaps
Use Yaps for voice, audio, video, images, translation, and memory when no focused skill fits. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-video-clipping
Shorten a talking-head video with Yaps Auto Cut. Review the pauses, adjust the cut, and export. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-video-to-audio
Save the sound from a video as MP3, WAV, or M4A with Yaps. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-transcription
Turn an audio or video file into text with Yaps. Good for interviews, podcasts, and voice memos. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-text-to-speech
Turn text into spoken audio with Yaps. Choose a supported voice and language, then save the file. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-srt-generator
Make a timed SRT subtitle file from audio or video with Yaps. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-auto-captions
Add captions to a video with Yaps. Review the words and timing, choose a style, then export. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-dictation
Type with your voice using Yaps on desktop. Get set up, fix a problem, or recover a recent dictation. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
yaps-audio-cleaner
Reduce noise, hiss, and static in a speech recording with Yaps. Save a separate clean file. New users: install Yaps and sign in.
- Publisher
- Yaps AI
- Fixed release
- 1.0.2
Safe Driving Behavior Analyzer | 安全驾驶行为分析工具
Analyzes videos of vehicle drivers to identify unsafe driving behaviors. It generates professional analysis reports to help enhance road safety awareness. | 安全驾驶行为分析工具,针对机动车驾驶人员的驾驶行为进行视频分析,识别不安全驾驶行为,输出专业分析报告,提升道路交通安全意识
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.13
reel-watch
Watch a video for the user: Instagram reels, TikToks, YouTube, X or local files. Gemini watches it with audio (or local frames + Whisper without a key), then you report the exact tools, links, repos and commands it shows. Use when a message has a video link or file.
- Publisher
- Fernane Mohamed Hanafi
- Fixed release
- 1.0.0
Text to Audio GPT
GPT-only: озвучивает текст и документы выразительным голосом и собирает проверенный MP3; подходит для чтения вслух, аудиокниг и озвучки. Требует GPT browser control и FFmpeg/FFprobe, не нативен для OpenClaw; Android ChatGPT требует отдельного импорта ZIP.
- Publisher
- ciklopentan
- Fixed release
- 1.0.2
Child Drowsiness / Fatigue Detection | 儿童打瞌睡/疲劳检测
Using a fixed camera in the classroom or above the home desk, the system analyzes the child's (student's) facial video in real time, detecting eye closure ratio (PERCLOS — the proportion of time eyes are closed more than 80% within a unit time), head-nodding frequency (rapid downward nod followed by lift), and changes in eye-region glossiness, and computes a comprehensive fatigue index (0-100). | 通过教室或家庭书桌的固定摄像头,实时分析儿童(学生)的面部视频,检测眼部闭合比例(PERCLOS,单位时间内眼睛闭合超过80%的时间占比)、点头动作频率(头部快速下点后抬起)以及眼部区域的光泽度变化,综合计算疲劳指数(0-100)。该技能…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.10
deyo
Use only when the current user explicitly asks to use Deyo to transcribe explicitly provided URLs or exact local audio/video file paths, or explicitly asks for Deyo install, status, or troubleshooting. Do not trigger from a mere Deyo mention, ambient context, an implicit attachment, directory browsing, a glob, stdin, or inferred permission to log in, install software, or read files.
- Publisher
- Casa Taloyum
- Fixed release
- 1.0.13
Elderly Sleep Nightmare / Startle Detection | 老年人睡眠中间惊醒/梦魇行为识别
Using a fixed bedroom camera (infrared night vision + microphone), the system analyzes elderly nighttime sleep video and detects abnormal events such as sudden sitting-up (quick lying-to-sitting transition), screams (high-pitched short cries), and arm-thrashing (purposeless rapid arm movements), and records the occurrence time, frequency and duration of each event. | 通过卧室固定摄像头(红外夜视),分析老年人夜间睡眠视频,检测突然坐起(快速从躺卧变为坐立)、惊叫声音(高频短促叫声)以及挥舞手臂(无目的性的快速手臂动作)等行为,记录发生时间、频次及持续时间。该技能可帮助家属或护理人员了解老人夜间睡眠质量,识别可能的梦魇、快速眼动期睡眠行为障碍等异常现象,为医疗评…
- Publisher
- smyx-skills
- Fixed release
- 1.0.11
XReplyAI - Social Post Manager
Generate, schedule, and publish posts across 10 platforms — X, LinkedIn, Instagram, Threads, Facebook, YouTube, TikTok, Pinterest, Bluesky, and Mastodon — in your voice using AI. Manage preferences and track billing.
- Publisher
- John Moon
- Fixed release
- 0.3.26
Elderly Loneliness / Depression-Tendency Behavior Analysis | 老年人孤独/抑郁倾向行为分析
Using fixed cameras at home (living room, bedroom) of elderly people living alone, the system analyzes daily videos and detects negative behavior indicators during solo time: dazing (long-duration motionless gazing without purposeful action), sighing (rapid chest rise-and-fall with audible expiration), and self-talking (mouth movement without any conversation partner). | 通过独居老人在家中的固定摄像头(如客厅、卧室),分析日常视频,检测独处期间的消极行为指标:发呆(长时间静止注视,缺乏目的性动作)、叹气(胸部快速起伏伴呼气声)、自言自语(口部活动但无对话对象)等。该技能可辅助家属或社区工作者了解老人心理状态,及时进行情感关怀或心理干预。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.9
550W AI Subtitle & Watermark Remover
Remove hardcoded video subtitles, watermarks from short-video share links, and image watermarks or unwanted text with the 550W Open API; check processing tasks and credits. Use for user-selected media or public share links.
- Publisher
- SunshineHu
- Fixed release
- 2.0.0
AI KEY·今天拍什么
AI KEY·今天拍什么 ——短视频选题技能。先探寻现象、检验假设、比较解释,把问题转成观众在意的选择。交互式选题:供需判定 + 对标信号 + 本人真实数据佐证 + 配比检查,不一键生成选题清单。 触发方式:/aikey-topic、/拍什么、/选题、/aikey-选题、「今天拍什么」「帮我出选题」「这个题能不能做」「有个想法你帮我判断下」「这题只有一头」「这题没劲」「帮我把它讲出两头来」 Topic selection for short videos: supply-demand test, benchmark signals, real data evidence. Trigger: /aikey-topic, "what topic should I shoot", "is this topic worth making" —— AI KEY · 不给公式,给判据。每条规则都标了实测代价。
- Publisher
- 詹明明
- Fixed release
- 0.5.3
AI KEY·找对标
AI KEY·找对标 ——找对标。给一个方向就去抖音/小红书/视频号搜人,给一个名字就去认人;三筛过滤(赚钱 / 看懂 / 能仿)挑出真正值得抄的那个,然后把他的**三批内容**全扒下来——最早 10 条(他怎么起的号)、数据最好 10 条(什么能爆)、最新 10 条(他现在在哪)——封面、大字、逐字稿、互动数据一条不落,最后出完整拆解和抄袭路线图。 触发方式:/aikey-benchmark、/找对标、/对标、「我该学谁」「帮我找个对标」「这个号值不值得学」「把这个博主拆一下」「他是怎么起号的」「扒一下这个账号」 Find and dissect a benchmark creator: search by direction or by name, filter on money/understandable/copyable, then pull the earliest 10, best-performing 10, and latest 10 posts with covers, cover text, transcripts and engagement data, and produce a full tea…
- Publisher
- 詹明明
- Fixed release
- 0.2.3
Child Dangerous Object Contact Detection | 儿童接触危险物品识别
Using fixed cameras in the living room, child's room, kitchen, or other home zones, AI object detection and pose estimation analyze the video in real time to recognize a child's hand actions and the objects in hand, identifying whether the child grabs scissors, knives, medicine bottles, lighters, or other preset dangerous items, or inserts fingers into electrical socket holes. | 通过家庭客厅、儿童房或厨房等区域的固定摄像头,利用AI目标检测和姿态估计技术实时分析儿童手部动作及手中持有的物品,识别儿童是否抓握剪刀、刀具、药品瓶、打火机等预设危险品,或是否将手指插入电源插座孔。一旦检测到危险行为,立即输出预警,联动手机APP或智能音箱发出警报,提醒家长…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.6
YouTube Data
YouTube channels, videos, transcripts, and comments for agents — content research and channel intelligence.
- Publisher
- superagnt_
- Fixed release
- 2.0.6
invoice-organizer
Automatically organizes invoices and receipts for tax preparation by reading messy files, extracting key information, renaming them consistently, and sorting them into logical folders. Turns hours of manual bookkeeping into minutes of automated organization.
- Publisher
- javotillo
- Fixed release
- 1.0.0
Office Prolonged Sitting & Posture Warning | 成人久坐/姿态预警(办公室)
Using a fixed camera in the office (aimed at the workstation), the system analyzes office workers' sitting-posture video in real time, detecting continuous sitting duration, neck forward angle (head offset relative to shoulders), and back curvature (hunchback degree). | 通过办公区固定摄像头(对准工位)实时分析办公人员的坐姿视频,检测连续坐姿时间、颈部前倾角度(头部相对于肩部的偏移)、背部弯曲度(驼背程度)。当久坐时间超过预设阈值(默认1小时)且未起身活动时,输出'久坐提醒';当颈部前倾角>20°或背部弯曲超过阈值时,输出'姿态异常提醒'。
- Publisher
- smyx-skills
- Fixed release
- 1.0.11
Avatar-Outfit-Motion Video
头像服装动作组合搭配视频——头像、服装、动作三必备主体要素 + 配音口播、背景、物品三可选次要要素,六要素皆可随机生成或指定(配音口播=音频驱动对口型、背景=场景生成/替换、物品=手持广告产品),走「判定→生成→组合→守护」五域管线:生成可独立交付、可自由组合的六要素资产,组合分固定组合(三视图锁身份、动作迁移/配音口播成片)与自由组合(任意要素→静态产物),提示词/产物双方式,输出 AI 换装/舞蹈/走秀/数字人口播/姿势短视频或静态资产。触发词:换装视频、AI 模特穿搭、虚拟试衣、动作迁移、音频对口型、配音口播、数字人、角色三视图、深度视频提取、背景替换、产品植入、手持道具、静态换装图、产
- Publisher
- 波动几何
- Fixed release
- 1.0.9
Rehab Patient Frustration / Giving-up Tendency Motivation | 康复患者沮丧/放弃倾向激励
Through fixed cameras in rehabilitation centers or home rehab areas, the system analyzes video of patients during rehabilitation training to detect frustration / giving-up tendency behaviors: sighing (rapid chest-abdomen rise-fall with exhalation), training interruption (actively stopping before reaching preset reps or duration), head-down silence (head lowered, avoiding eye contact, long-term silence), sluggish or. | 通过康复中心或家庭康复区的固定摄像头,分析患者在进行康复训练时的视频,检测沮丧/放弃倾向行为:叹气(胸腹快速起伏伴呼气声)、中断训练(在未达到预设次数或时间前主动停止动作)、低头不语(头部低垂,…
- Publisher
- smyx-skills
- Fixed release
- 1.0.11
uttera
Audio for the agent. Transcribes audio, summarises long recordings, translates, turns text into speech, generates sound effects and music, scores pronunciation, and verifies signed reports using Uttera. Use it when the user sends an audio file, asks for something to be read aloud, asks what was said
- Publisher
- Uttera Labs
- Fixed release
- 1.3.0
Autism Spectrum Disorder Behavior Analysis Tool | 孤独症谱系障碍行为分析工具
Performs special video analysis on behavioral characteristics of children with autism, identifies core symptom features, provides structured analysis reports and intervention recommendations. | 孤独症谱系障碍行为分析工具,针对儿童孤独症行为特征进行专项视频分析,识别核心症状特征,提供结构化分析报告和干预建议
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.13
gait test
Triggers when a user provides a pet side-view walking video URL or file for analysis; uses AI pose estimation to track limb joint trajectories, analyzes stri...
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.5
Pet Vocal Emotion Deep Classification | 宠物叫声情绪深度分类
Triggers when a user provides a pet vocalization audio/video URL or file for analysis; extracts acoustic features such as frequency, duration, interval, and harmonic structure via AI audio analysis, and classifies the vocalization into 6+ emotion categories (howling, growling, excitement, loneliness, fear, whining/coaxing) with confidence scores. Helps owners understand pet emotional states, improve human-pet interaction, and detect potential stress or health issues early. Application: daily companionship (smart c…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.8
Anxiety-Related Behavior Recognition (Hand-rubbing / Nail-biting / Pacing) | 焦虑症相关行为(搓手、咬指甲、来回踱步)识别
Using a fixed camera at home or in the office, the system analyzes daily videos of an individual (e.g., adult, adolescent) and detects anxiety-related behaviors: hand rubbing (repeated rubbing of both hands), nail biting (hand approaching mouth with biting motion), and pacing (repeated back-and-forth walking in a small area). | 通过家庭或办公室的固定摄像头,分析个体(如成人、青少年)的日常行为视频,检测手部搓揉(双手反复摩擦)、指甲啃咬(手部靠近嘴部并有啃咬动作)、来回踱步(在狭小区域内反复折返行走)等焦虑相关行为。系统实时监测,当焦虑行为指数超过阈值时推送提醒(如'您今日焦虑行为较多,建议进行放松练习')。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.4
Family / Couple Conflict Intensity Detection | 夫妻/家庭争吵强度识别
Using a fixed camera with microphone in the living room, the system analyzes audio and video in real time, detecting sound intensity (dB) and the intensity of body movements (e.g., rapid hand waving, finger pointing, pushing, throwing objects). It comprehensively evaluates the family conflict intensity level (low / medium / high). | 通过客厅固定摄像头(含麦克风),实时分析音频和视频,检测声音强度(分贝)和肢体动作激烈程度(如快速挥手、戳指、推搡、摔物等)。综合评估家庭争吵的冲突强度等级(低/中/高),当强度达到中或高时,通过手机APP推送提醒(如'检测到高强度冲突,建议冷静沟通或暂时分开')。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.8
Infant Suffocation Risk Detection | 婴幼儿趴睡窒息风险识别
Using a baby monitor (smart camera) fixed above the crib, the system analyzes infant sleep video in real time to detect sleep posture (supine, side, prone) and whether the mouth/nose area is occluded by a blanket, pillow, plush toy or other object. | 通过婴儿监护器(智能摄像头)固定于婴儿床上方,实时分析婴儿睡眠视频,检测婴儿的睡姿(仰卧、侧卧、俯卧)以及口鼻区域是否被被子、枕头、玩偶等异物遮挡。当检测到俯卧或口鼻被遮挡时,输出风险等级(中风险/高风险),并立即向父母手机APP推送警报,预防婴儿猝死综合征(SIDS)和窒息意外。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.5
Kitchen Stove Left-On Detection | 老年人厨房忘关火识别
Using a fixed kitchen camera (must be able to capture the stove area), the system analyzes video in real time to detect whether there is human activity in the kitchen area, and at the same time identifies stove flames or heat sources (e.g., thermal/infrared features) to determine whether the gas stove is on. | 通过厨房固定摄像头(需能拍摄到灶台区域)实时分析视频,检测厨房区域内是否有人体活动,同时识别灶台火焰或热源(如红外特征)以判断燃气灶是否处于开启状态。当检测到厨房无人连续超过预设时间(默认10分钟)且灶火仍处于开启状态时,输出'忘关火'预警,可联动智能燃气阀自动关闭阀门,并推送提醒至家属或护理人员手机,预防火灾和燃气泄漏事故。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.9
youtube-data
Use when structured YouTube data is needed: pasted video/channel/playlist links, transcripts for analysis, video metadata, channel upload history, search results, or playlist contents — without Google API quotas or OAuth. Triggers on YouTube URLs, creator names, topic research, or any request needing YouTube content, even if not mentioned explicitly. Not for uploads, account management, or written-source-only research.
- Publisher
- Rohit Das
- Fixed release
- 1.6.2
youtube-api
Use when YouTube data is needed without Google API quotas or OAuth setup: transcripts, video metadata, channel info, search results, playlists. Triggers on pasted YouTube links, creator names, @handles, topic research, video summaries, channel browsing, or any request where YouTube content would help — even if not mentioned explicitly. Not for uploads, account management, or written-source-only research.
- Publisher
- Rohit Das
- Fixed release
- 1.6.2
MarkItDown
Convert documents AND web pages to Markdown with Microsoft's MarkItDown CLI (`markitdown`). Covers PDF, Word, PowerPoint, Excel, images (EXIF/LLM description), audio/video transcription, HTML, YouTube and direct URLs. Use when the user asks to read / analyze / summarize / extract / translate / Q&A about a rich-format file or a public web page, or to deposit such content into a knowledge base; converting to plain Markdown first also cuts token cost. Also ships optional local-only token-cost estimators (`scripts/tok…
- Publisher
- Wing
- Fixed release
- 1.8.0
youtube-playlist
Use when a YouTube playlist is involved: pasted playlist links or IDs, requests to list playlist videos, browse playlist contents, or work through a playlist for transcripts or research. Also use when the user wants all videos from a series, course, or collection. Not for creating playlists or account management.
- Publisher
- Rohit Das
- Fixed release
- 1.6.0
video-transcript
Use when video content needs to be extracted as text: pasted YouTube links or IDs, requests to transcribe, summarize, quote, translate, convert video to text, or extract information from video content. Also use when a user shares a video URL without explanation and wants to know what it says. Not for uploads or account management.
- Publisher
- Rohit Das
- Fixed release
- 1.5.1
transcript
Use when the spoken content of a YouTube video is needed — even if not explicitly requested: pasted video links or IDs, requests to summarize, quote, transcribe, translate, fact-check, or extract anything from a video. Also use for research or learning when a video is the source. Not for uploads or account management.
- Publisher
- Rohit Das
- Fixed release
- 1.5.1
subtitles
Use when subtitles or the spoken text of a YouTube video is needed: pasted video links or IDs, requests to translate a video, read along, follow foreign-language content, or extract what was said. Also use for language learning or accessibility. Fetches timestamped subtitles from any YouTube video. Not for uploading subtitles or account management.
- Publisher
- Rohit Das
- Fixed release
- 1.5.1
captions
Use when captions, subtitles, or the spoken text of a YouTube video is needed — even if not explicitly requested: pasted video links or IDs, requests to read, quote, or translate a video, accessibility needs, deaf/HoH use cases, content review, or language learning. Fetches timestamped caption data from any YouTube video. Not for uploading subtitles or account management.
- Publisher
- Rohit Das
- Fixed release
- 1.5.1
yt
Use when YouTube is relevant: pasted video links or IDs, @handles, quick video lookups, summaries, channel latest uploads, topic search, or any request involving YouTube content — even if YouTube is not mentioned explicitly. Covers transcripts, search, and channel latest. Not for uploads or account management.
- Publisher
- Rohit Das
- Fixed release
- 1.5.2
youtube-research
Pulls structured YouTube data — video and channel details, transcripts/captions, comments, playlists, and search — via the Crawlora API as clean JSON, with no yt-dlp or HTML scraping. Use when the user provides a YouTube URL or asks for a transcript, comments, channel/video metadata, or video search results.
- Publisher
- Crawlora
- Fixed release
- 1.0.18
guaikei视频字幕提取转文字
将已完成的点播视频转写为文字与结构化文案。当用户提供本地视频文件或抖音、小红书等公网视频链接,要求视频转文字、视频转稿、字幕提取、视频总结、会议纪要、课程拆解、二创改写时使用。不支持实时直播流、需登录的加密视频与纯音乐无人声视频;转写后可按自定义 Prompt 生成总结、金句、分镜头、翻译等
- Publisher
- engheng-art
- Fixed release
- 1.0.0
gift-occasion-planner
Use when birthdays, anniversaries, and holidays keep sneaking up on you every single year — maintains a registry of people and their gift occasions (birthday, anniversary, Mother's/Father's Day, etc.) with per-person gift profile (interests, sizes, favorites, taboo list, past gifts), generates a rolling 60-day lookahead of upcoming occasions, recommends what to give based on profile + past-gift history (never repeat, always evolve), budgets per person/occasion with running yearly totals, produces an actionable buy…
- Publisher
- voronindenis5
- Fixed release
- 1.0.0
auditor-ganchos-cuenta-formula100k
Auditoría end-to-end del PATRÓN DE GANCHO de una cuenta IG/TikTok. Activar cuando pidan: "audita los últimos N reels de @cuenta", "top 10 vs 10 peores", "analiza esta cuenta y su patrón ganador", "compara top vs bottom", "qué hace funcionar los reels de X", "reverse-engineering de @cuenta", "reporte visual de los ganchos de [cuenta]". Pipeline: Apify instagram-reel-scraper (métricas, portadas y transcripts) → integra estadísticas nativas de IG si las pasan (Composio o pegadas) → análisis Top-N vs Bottom-N + compar…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.2
YouTube Data
YouTube channels, videos, transcripts, and comments for agents — content research and channel intelligence.
- Publisher
- Jaen
- Fixed release
- 2.0.5
Videosays · Video to Text
Transcribe video links and extract spoken scripts or subtitles from Douyin, Xiaohongshu, Bilibili, YouTube, TikTok, and other supported platforms. Use for single or batch transcription, timestamped transcripts, SRT/VTT export, credit balance, and transcription history.
- Publisher
- wegofuture
- Fixed release
- 1.3.1
video-extend-edit
Requires OFOX_API_KEY — create one at https://app.ofox.ai, plus a video you already have. Makes an existing clip longer, or replaces its ending. Use when a user wants more of footage they already have, e.g. "extend this 5-second clip to 15", "keep going from where this one ends", "re-shoot the ending from 4 seconds on", or "add another shot onto this". A frame is pulled out of the clip at zero cost and becomes the first frame of a newly generated segment, which is then joined onto the original. Do not use to chang…
- Publisher
- OFOX AI
- Fixed release
- 2.0.0
talking-head
Requires OFOX_API_KEY — create one at https://app.ofox.ai. Turn a portrait plus a short script into a clip of one person speaking those words to camera. You supply the text and the model generates the voice — audio cannot be uploaded, measured. Defaults to alibaba/wan-3.0-prime rather than this repo's usual seedance-2.5, because seedance-2.5 refuses a real person's photo at submission. Use when a user has a face and some words and wants the face to say them, e.g. "make this headshot read my intro", "a spokesperson…
- Publisher
- OFOX AI
- Fixed release
- 2.0.0
shorts-reels
Requires OFOX_API_KEY — create one at https://app.ofox.ai. Generate cheap vertical 9:16 drafts in one priced batch, pick a winner off the contact sheet, then re-render only that one. The cheap tier is the point — several times cheaper per second, so a whole set can cost less than one flagship clip. Use when a user wants Shorts/Reels/TikTok raw material rather than one finished video, e.g. "give me 5 vertical clips to choose from", "a few Reels drafts for this product", "some cheap options before we commit", or "ba…
- Publisher
- OFOX AI
- Fixed release
- 2.0.0
seedance-short-drama
Requires OFOX_API_KEY — create one at https://app.ofox.ai. Generate a realistic-human, dialogue-driven short-drama clip — one shot, or a few hard-cut shots inside one job — from a script or scene description using the Ofox video API (Seedance 2.5). Runs a short creative brief when the input leaves beat, aspect ratio, emotional arc or camera register open (one held take, a travelling take, or a multi-shot cut list) ("Let the AI decide" is offered on the taste questions, never on a must-ask one, and never as the def…
- Publisher
- OFOX AI
- Fixed release
- 2.0.0
music-video
Requires OFOX_API_KEY — create one at https://app.ofox.ai, plus the music file the finished video must carry. Your audio never reaches the API (measured) — the visuals are written to the track's tempo, mood and sections, then your own file is laid on locally at zero cost. One job caps at 30 seconds, so a three-minute song is six jobs minimum and the cost table says that before anything is spent. Use when a user has a specific piece of music and wants visuals for it, e.g. "make a music video for this track", "visua…
- Publisher
- OFOX AI
- Fixed release
- 2.0.0
explainer
Requires OFOX_API_KEY — create one at https://app.ofox.ai. Turn an article, doc or release note into a short explainer clip — one person to camera, or a voiceover over illustrative footage. The user supplies the source text and the model generates the speech; audio cannot be uploaded, measured. A 30-second clip holds about eighty spoken words in eight sentences — measured, and well under a tenth of a 1,200-word post — so this skill does not summarise an article, it picks the single idea worth saying and helps choo…
- Publisher
- OFOX AI
- Fixed release
- 2.0.0
aivideo-remix
拆解并复刻爆款短视频:粘贴抖音/快手等短视频链接,AI 自动分镜拆解,逐镜头生成图片与视频,合并成片,支持断点续作。当用户想"拆解这条爆款视频 / 复刻对标视频 / 爆款二创 / 分镜拆解再创作"时使用。平台 www.aidaoyan66.online,注册即送1000体验积分,Seedance 系列官方价6~8折。
- Publisher
- chenzhigao61
- Fixed release
- 1.0.0
aivideo-pptvideo
把PPT或图文素材变成15~60秒竖屏讲解短视频:AI写解说脚本、选音色配音、逐页生成画面、合成成片并烧录字幕。当用户想"把这份PPT做成短视频 / PPT转视频 / 图文做成解说视频 / 知识口播成片"时使用。平台 www.aidaoyan66.online,注册即送1000体验积分,Seedance 系列官方价6~8折。
- Publisher
- chenzhigao61
- Fixed release
- 1.0.0
aivideo-koubo
输入口播文案生成数字人播报视频:AI 优化脚本、选数字人形象与音色(支持声音克隆与男女真人音色库)、按生成时长计费、支持预览下载,另有双人剧情对话与带货收尾模式。当用户想"做一条口播视频 / 数字人播报 / 不出镜带货口播 / 剧情对话视频"时使用。平台 www.aidaoyan66.online,注册即送1000体验积分。
- Publisher
- chenzhigao61
- Fixed release
- 1.0.0
Baoyu YouTube Transcript
Downloads YouTube video transcripts/subtitles and cover images by URL or video ID. Supports multiple languages, translation, chapters, and speaker identification. Caches raw data for fast re-formatting. Use when user asks to "get YouTube transcript", "download subtitles", "get captions", "YouTube字幕", "YouTube封面", "视频封面", "video thumbnail", "video cover image", or provides a YouTube URL and wants the transcript/subtitle text or cover image extracted.
- Publisher
- agentforge-cyber
- Fixed release
- 1.0.0
subtitle-generator
当用户需要生成字幕、制作字幕、字幕对齐、ASR识别、语音转文字时使用此技能。触发词:生成字幕, 字幕生成, 制作字幕, 视频字幕, 语音识别, ASR字幕, Whisper字幕, 字幕制作, 视频转字幕, 字幕对齐, 音频转字幕, 语音转文字, 自动字幕, ASR识别, faster-whisper, subtitle, subtitles, caption, transcription, speech to text
- Publisher
- agentforge-cyber
- Fixed release
- 1.0.0
local-tts
文本转语音(TTS)。三条通道:edge-tts(微软神经网络语音,免费无需 key,15 个中文音色可用)、Azure Speech 官方 API(75 个中文音色全解锁,含晓秋/晓辰/HD/MAI,50 万字符/月免费)、pyttsx3(Windows SAPI 离线兜底)。当用户/agent 需要把文本转成语音文件(mp3/wav)、生成语音播报、配音、解锁被 edge-tts 封锁的音色时使用。
- Publisher
- Azrael Noah
- Fixed release
- 1.1.0
subtitle-download
Batch-download missing subtitles for a media library of organized movie folders.
- Publisher
- Mina Atef
- Fixed release
- 1.0.0
ym-gift-pick-helper
按收礼人关系、预算、已知偏好与禁忌推荐礼物方案,给出理由与备选,避免泛泛清单,不确定时给出可问清偏好的问题。当用户说「送什么礼物」「选个礼物」「礼物推荐」时使用。 也适用于「选礼物」「送礼建议」「gift ideas」这类说法。
- Publisher
- liuyuming0823
- Fixed release
- 0.1.0
Youtube Playlist Transcripts
Fetch every transcript of a YouTube playlist as one background bulk job via the BulkTranscripts API, or list the playlist in order and fetch videos one by one. Use when the user shares a playlist link, wants a lecture series or course turned into study notes, or needs playlist transcripts as text. Requires a free BulkTranscripts API key in BULKTRANSCRIPTS_API_KEY, created at https://bulktranscripts.co/app?tab=mcp (Google sign-in, 30 free credits, no card).
- Publisher
- pratie
- Fixed release
- 1.0.1
Youtube Transcripts
Fetch YouTube video transcripts, search YouTube, list channel or playlist videos, run a background bulk job for a whole channel or playlist, and track new uploads via the BulkTranscripts API. Use when the user shares a YouTube link, asks to summarize/analyze/quote a video, wants transcripts for a whole channel or playlist, needs YouTube research, or asks what a channel posted recently. Requires a free BulkTranscripts API key in BULKTRANSCRIPTS_API_KEY, which the user creates at https://bulktranscripts.co/app?tab=m…
- Publisher
- pratie
- Fixed release
- 1.0.2
Plaud Meeting Archive / Plaud会议归档助手
Archive selected Plaud audio, transcripts, or summaries in JishuDB; preserve provenance and verify imports before cited follow-up analysis. / 将选定的Plaud音频、转写或纪要归档到JishuDB,保留来源关联,验证导入后进行有引用的分析。
- Publisher
- bowen-aijishu
- Fixed release
- 0.1.1
shelbys-speech
把「谢尔比的记忆」档案库(含完整记忆系统:记忆库+重要md+日志+项目档案)按主题识别,一篇篇生成博士水平、SCI 格式的技术论文(中英文双语、自动检查修复 Word、问答式修改、确认后输出 PDF)。可识别多个主题→多篇论文,每篇先询问用户。选题明确后开放完整记忆系统访问权限(只读),素材更全。作者默认<用户名> & 谢尔比,可随时按用户要求修改。用户说写论文/技术论文/博士论文/SCI/投稿/演讲/分享/总结成文时使用。关键词:论文、SCI、博士论文、投稿、英文论文、PDF、演讲、分享、提炼总结、多主题、shelbys speech、完整记忆。
- Publisher
- hechunxian
- Fixed release
- 1.0.0
transcripcion-youtube-formula100k
Convierte videos de YouTube en batches de guiones virales con FÓRMULA 100K. Usar SIEMPRE que alguien comparta una URL de YouTube y quiera extraer ideas, guiones o contenido de ese video. También con: "transcribe este video", "sácame ideas de este YouTube", "convierte este video en guiones", "quiero hacer un batch de scripts de este video", "repurposea este video", "dame guiones basados en este video", o cualquier variación que combine una URL de YouTube con intención de crear contenido nuevo. Obtiene la transcripc…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.2
radar-contenido-f100k
Radar de contenido de FÓRMULA 100K: monta y corre un sistema que vigila Instagram y te avisa qué FORMATOS están despegando. Cosecha con Apify (búsquedas cortas convertidas a hashtags) y con los Reels guardados que la usuaria reenvía por Telegram; filtra cuentas pequeñas cuyas vistas multiplican sus seguidores; mira cada reel (capturas o análisis de video de Higgsfield) antes de anotarlo; describe la mecánica (no el tema) con anatomía de 4-6 pasos; y la anota en un registro de recurrencia. Una mecánica cruza el umb…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.1
analiza-tu-contenido-f100k
Hace el módulo «Analiza tu Contenido» de FÓRMULA 100K sobre la cuenta de la alumna y lo entrega en La Hoja del Mes (infografía imprimible en HTML). Trae sus últimos 30 Reels con sus estadísticas básicas (Composio con su Instagram profesional, o Apify + capturas) y sus transcripts, pide las avanzadas por captura de la app (retención y actividad en el perfil), lee los 5 filtros y el filtro tapado, corona moldes y ángulos, saca 3 conclusiones fijas (formato ganador, ángulo que más viraliza y posible ángulo de venta)…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.1
adaptador-formatos-f100k
Toma un video/post viral de CUALQUIER nicho (incluso uno random y ajeno) — por enlace de Instagram, TikTok o YouTube — DECODIFICA su formato/mecánica viral (no su tema) y lo ADAPTA al RUBRO DE QUIEN LO PIDE (cualquiera: se lee del Segundo Cerebro o se pregunta), proponiendo 2-3 ángulos listos para guionizar. Activar SIEMPRE que alguien diga "adapta este formato a lo nuestro", "roba el formato de este reel", "este video random me gustó, sirve para mi nicho?", "convierte este viral en algo mío", "decodifica este for…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.1
YouTube Search
Search YouTube for videos, channels or playlists, or search inside one channel, via the BulkTranscripts API, then fetch transcripts of the results. Use when the user asks to find videos about a topic, research what YouTube says about something, or locate a creator's videos on a subject. Requires a free BulkTranscripts API key in BULKTRANSCRIPTS_API_KEY, created at https://bulktranscripts.co/app?tab=mcp (Google sign-in, 30 free credits, no card).
- Publisher
- pratie
- Fixed release
- 1.0.0
YouTube Upload Monitor
Check a YouTube channel's newest uploads for free via the BulkTranscripts API and fetch transcripts only for videos that are new. Use when the user asks what a channel posted recently, wants alerts or digests for new uploads, or wants to track competitors' channels. Requires a free BulkTranscripts API key in BULKTRANSCRIPTS_API_KEY, created at https://bulktranscripts.co/app?tab=mcp (Google sign-in, 30 free credits, no card).
- Publisher
- pratie
- Fixed release
- 1.0.0
YouTube Video Transcript
Get the transcript of a single YouTube video (or Short, or TikTok video) as clean text, timestamped segments, or SRT/VTT/Markdown via the BulkTranscripts API. Use when the user shares a video link and wants it summarized, quoted, translated, fact-checked or turned into notes. Requires a free BulkTranscripts API key in BULKTRANSCRIPTS_API_KEY, created at https://bulktranscripts.co/app?tab=mcp (Google sign-in, 30 free credits, no card).
- Publisher
- pratie
- Fixed release
- 1.0.0
portada-video-carrusel-formula100k
Portada animada para un carrusel de Instagram: genera (o toma) la foto gancho con la cara de la usuaria en Higgsfield, la anima como cinemagraph de cámara fija (5 s, sin audio) y renderiza el título como PNG transparente con Chromium. En este servidor entrega las piezas (clip animado + título transparente + portada estática en PNG) listas para unir en CapCut o Edits. Usar cuando pida una portada animada o "tipo video" para un carrusel, "anima la foto gancho", "cinemagraph de portada", "la portada que se mueve", "c…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.0
motion-reels-f100k
Diseña motion graphics para un reel 9:16 talking-head estilo softgirlnocode: con una captura del video detecta dónde está la persona y calcula zonas seguras; con el .srt auto-clasifica el guion para elegir el overlay correcto (antes/ahora → comparison card, listas → chapter card, herramientas → UI card, identidad → pill labels, frases → kinetic, citas → quote pull) y cumple la densidad de 1 gráfico cada 5 segundos. Entrega las tarjetas como PNG transparentes 1080×1920 renderizadas desde HTML + un motion board + un…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.0
motion-takeover-f100k
Planea TAKEOVERS faceless a pantalla completa para un video de YouTube 16:9 grabado con tu cara, estilo 'Claude Routines': en los momentos clave la pantalla se llena de mockups paso a paso, trazos SVG e imágenes IA, y luego vuelve tu cara. A partir del .srt del video ya cortado arma el plan de 2 carriles (takeovers + overlays ligeros), renderiza cada takeover como PNG 1920×1080 y, si quieres movimiento, lo anima con Higgsfield. Entrega las piezas + una hoja de montaje con el segundo exacto de cada takeover para po…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.0
broll-vsl-formula100k
Genera entre 15-25 clips de B-roll cinematográficos para VSL o videos largos usando Higgsfield Seedance 2.0 y Cinema Studio 3.0. Regla 80/20: 80% de los clips incluyen el avatar/presentadora con foto de referencia (Seedance 2.0 + medias identity), 20% son tomas cinemáticas abstractas sin persona (Seedance 2.0 o Cinema Studio 3.0 puro). Estética configurable (default: cyberpunk azul eléctrico). Input: guion VSL (texto). Output: catálogo completo con URLs de video + frase exacta del guion donde insertar cada clip. A…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.1
audios-venta-f100k
Escribe guiones de AUDIOS DE VENTA (notas de voz comerciales, audios de cierre en DM, voicenotes de objeción, audios de seguimiento, audios de aplicación a high ticket) para grabar en WhatsApp / Instagram DM / Telegram. Convierte una oferta ya construida en una secuencia hablada de máximo 50 segundos que cierra la venta en privado. Activar SIEMPRE que se pida 'haz un audio de venta', 'guion para nota de voz', 'audio para cerrar en DM', 'voicenote para vender', 'audio de cierre', 'audio de seguimiento', 'audio para…
- Publisher
- Fórmula 100K
- Fixed release
- 1.0.1
cn-model-gateway
国产大模型统一 MCP 服务器,通过标准 JSON-RPC 2.0 协议为 Claude Code / Cursor / Cline 等 Agent 框架提供 DeepSeek、通义千问、智谱 GLM、Kimi、腾讯混元、火山豆包、MiniMax、零一万物、百川智能、阶跃星辰十家模型的统一调用接口。10 个 MCP 工具(ask_model/describe_image/embed_text/rerank/audio_transcribe/video_understand/batch_submit/batch_result/list_providers/health_check)+ 单一网关状态资源 + 2 个 prompt 模板。内置统一错误映射、流式 SSE 输出+心跳保活+断线重连、使用量统计、硬件感知并发控制、SQLite WAL 批量任务队列、自动故障转移、环境变量优先读取 API key。支持 Function Calling、多模态视觉、5 个非 MCP 框架适配器(LangChain/AutoGPT/CrewAI/Coze/Dify)、性能基准测试和 Token 价格追踪。config.json 填写 ap…
- Publisher
- fyniujin
- Fixed release
- 1.8.0
Video Deep Reader
Check the pinned transcript dependency and optionally test YouTube connectivity, fetch public captions into a private per-user cache, retrieve excerpts by video ID, and turn captions or a supplied transcript into timestamped Quick, Deep, or Research notes.
- Publisher
- Bonnie Geng
- Fixed release
- 2.0.5
EchoTik-视频下载地址查询
解析TikTok视频地址,返回该视频的无水印/含水印下载地址、播放地址与封面地址,用于保存带货视频素材或离线分析。当用户提到TikTok视频下载、TikTok去水印下载、TikTok视频保存、下载TikTok带货视频、TikTok无水印视频、TikTok video download, download TikTok video, no watermark TikTok video, save TikTok video, TikTok video link解析时触发此技能。即使用户未明确提及"EchoTik",只要其需求涉及从一个TikTok视频链接取出可下载/可播放的视频地址,也应触发此技能。
- Publisher
- linkfox-ai
- Fixed release
- 1.0.7
Temu欧洲站-税务
Temu 欧洲站电商税务(Tax)API,经 LinkFox 网关转发 Partner EU 7 个 temu.pay.tax.* 接口:导出报表、Galerie签名、发票查询/下载、商家报表下载/上传发票等。当用户提到 Temu EU Tax、temu.pay.tax.invoice、VAT、发票上传、export report、Galerie signature、site=eu 税务 时触发。商品管理用 linkfox-temu-manage-product-eu。
- Publisher
- linkfox-ai
- Fixed release
- 1.0.6
TikTok官方-达人API
TikTok 达人(Creator/affiliate creator)数据与可购物视频技能,经 LinkFox 网关代理调用 TikTok Shop 达人开放接口:达人主页/档案、达人绑定店铺商品、橱窗商品、可购物视频的上传/内容预检/发布/发布状态查询。需要达人 access_token(user_type=1),由 linkfox-tiktok-video-auth 完成达人授权后取得。当用户提到 TikTok 达人、TikTok creator、达人主页、达人档案、达人资料、达人店铺商品、达人绑定店铺商品、达人橱窗商品、showcase 商品、上传可购物视频、发布可购物视频、视频发布状态、视频内容预检、shoppable video、affiliate creator、TikTok 带货达人信息、TikTok creator profile、shop products、showcase products、post shoppable video、video status、precheck 时触发此技能。即使用户未写 EHunt/紫鸟,只要需求是查 TikTok Shop 达人的资料、绑定商品或可购物视频带货操作,也…
- Publisher
- linkfox-ai
- Fixed release
- 1.0.6
AIGC视频生成
AI生视频工具(首尾帧/单图模式),根据原图和提示词生成视频,支持可选尾帧图控制结束画面。支持模型KLING可灵/WAN万相/SEED豆包/SEED_FAST/HAILUO海螺。用户说"生成视频"、"AI视频"、"图生视频"、"做个视频"、"video generation"、"generate video"、"图片转视频"、"动态视频"时触发。
- Publisher
- linkfox-ai
- Fixed release
- 1.2.3
EchoTik-Tiktok新品排行
通过EchoTik新品排行数据,发现TikTok Shop 16个区域市场的热门新品。当用户提到TikTok新品排行、TikTok热销商品、TikTok Shop爆品、短视频电商选品、TikTok新品发掘、跨境TikTok选品、TikTok new product rankings, TikTok bestsellers, short-video product selection, TikTok viral products, new product ranking, TikTok product trends时触发此技能。即使用户未明确提及"EchoTik"或"新品排行",只要其需求涉及发现TikTok Shop上的热卖新品或新兴商品趋势,也应触发此技能。
- Publisher
- linkfox-ai
- Fixed release
- 1.0.8
video-auto-generator
【2026增强版】基于AI自动生成视频内容。支持文本转视频、图文转视频、脚本自动生成+配音+字幕+剪辑全流程。集成edge-tts配音、ffmpeg自动化剪辑。当用户说"生成视频"、"制作视频"、"自动剪辑"、"批量生产视频"、"AI视频生成"、"做个短视频"时触发此技能。
- Publisher
- nh5gntnf78-oss
- Fixed release
- 1.0.1
video-generator
使用AI大模型根据文字、首尾帧图片或图片/视频/音频参考素材生成带对白和环境音的视频。适用于文生视频、图生视频、首尾帧过渡和多模态参考视频生成,通过 Deep Code Plus 试算积分、确认后上传素材并生成视频。
- Publisher
- null
- Fixed release
- 1.0.0
video-frames
Extract frames or short clips from videos using ffmpeg.
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
summarize
Summarize or transcribe URLs, YouTube/videos, podcasts, articles, transcripts, PDFs, and local files.
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
sherpa-onnx-tts
Local text-to-speech via sherpa-onnx (offline, no cloud)
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
sag
ElevenLabs text-to-speech with mac-style say UX.
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
openai-whisper-api
OpenAI Audio Transcriptions API via curl; gpt-4o-transcribe, mini, diarize, or whisper-1.
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
openai-whisper
Local speech-to-text with the Whisper CLI (no API key).
- Publisher
- azizbrownint
- Fixed release
- 1.0.0
daily-digest
Generate a structured daily work digest from session logs and memory files. Use when the user asks for a daily summary, end-of-day recap, work log, or 'what did I do today'. Scans memory/*.md, session transcripts, and git activity to produce a human-readable digest with key accomplishments, decisions, and pending items.
- Publisher
- terrycarter1985
- Fixed release
- 1.0.0
亚马逊-店铺Fulfillment Outbound
亚马逊 Multi-Channel Fulfillment(MCF)新版出仓履约管理。用于获取配送报价和订单预览,创建、查询、列出、更新或取消履约订单,查看包裹与跟踪、签收证明、投递照片、收件人及 locker/drop-off 信息,并查询相关发票头。用户提到 Fulfillment Outbound、MCF、多渠道配送、亚马逊库存配送到站外客户、配送报价、getOffers、getOrderPreview、createOrder、getInvoiceHeaders、receivedBy、unitIdentifiers、v2026-07-04 时触发。即使未明确说“MCF API”,只要希望用亚马逊库存履约非亚马逊渠道订单,也应触发此技能;External Fulfillment、普通 Orders API 和旧版 v2020-07-01 不属于此技能。
- Publisher
- linkfox-ai
- Fixed release
- 1.0.1
出海匠 TikTok 广告与创意情报
使用出海匠(Chuhaijiang)研究 TikTok 公开广告与创意素材,支持广告搜索、广告详情与关联商品,以及创意搜索、详情、脚本分镜分析和向量数据。用户点名出海匠或 Chuhaijiang 时触发;未指定数据源时,仅在需要广告到商品关系、创意结构拆解、AIGC/赞助/带货素材筛选时触发。通用 TikTok 带货视频榜单或详情使用 linkfox-kalodata-tiktok-video,广告账户与投放管理使用对应的 TikTok Ads 管理工具;点名其他数据源时不触发。
- Publisher
- linkfox-ai
- Fixed release
- 1.0.0
出海匠 TikTok 视频情报
使用出海匠(Chuhaijiang)研究 TikTok 公开视频,支持多条件视频搜索、单视频详情、带货商品和评论钻取。用户点名出海匠或 Chuhaijiang 时触发;未指定数据源时,仅在需要视频评论或视频到带货商品的关系钻取时触发。通用 TikTok 带货视频榜单或详情使用 linkfox-kalodata-tiktok-video,视频上传发布使用 linkfox-tiktok-video;点名其他数据源时不触发。
- Publisher
- linkfox-ai
- Fixed release
- 1.0.0
agent-storage-maintenance
Find and reclaim what an agent accumulates on disk: oversized sessions, duplicated transcripts, stale embedding caches, abandoned trees.
- Publisher
- POSTHUMAN
- Fixed release
- 0.1.2
Doubao Speech
Doubao Speech (volcengine.com). Use this skill for ANY Doubao Speech request — reading, creating, and updating data. Whenever a task involves Doubao Speech, use this skill instead of calling the API directly.
- Publisher
- OOMOL
- Fixed release
- 1.0.2
google-voice-caller
Automate Google Voice calls with AI-generated voice (TTS) or local audio injection.
- Publisher
- joe12801
- Fixed release
- 1.2.0
hedy
Access and manage Hedy meeting data: sessions, transcripts, highlights, todos, topics, contexts, vocabulary, and webhooks via the Hedy REST API.
- Publisher
- Julian Pscheid
- Fixed release
- 1.2.0
詹明明·今天拍什么
📐 詹明明·今天拍什么 ——短视频选题技能。先探寻现象、检验假设、比较解释,把问题转成观众在意的选择。交互式选题:供需判定 + 对标信号 + 本人真实数据佐证 + 配比检查,不一键生成选题清单。 触发方式:/zmm-topic、/拍什么、/选题、/zmm-选题、「今天拍什么」「帮我出选题」「这个题能不能做」「有个想法你帮我判断下」「这题只有一头」「这题没劲」「帮我把它讲出两头来」 Topic selection for short videos: supply-demand test, benchmark signals, real data evidence. Trigger: /zmm-topic, "what topic should I shoot", "is this topic worth making" —— 📐 詹明明 · 不给公式,给判据。每条规则都标了实测代价。
- Publisher
- 詹明明
- Fixed release
- 0.2.10
Crun Agent Skills
Run Crun image, video, audio, music, and media-tool workflows through the bundled standalone runtime. Use whenever the user wants to generate, edit, or transform an image, video, voice, speech, or music clip — even if they never say "Crun" or name a model — as well as for model routing, model-schema
- Publisher
- CarolineWhite8888
- Fixed release
- 1.0.0
Wan 3.0 Video Prompt Architect
Build and review production-ready prompts and short shot plans for WAN-30.video text-to-video or image-to-video work. Use when a user needs clearer subject motion, camera direction, scene continuity, audio direction, reference-image roles, or a pre-generation configuration check; do not use this Skill to claim official Alibaba or Wan affiliation, call a model provider, access an account, or spend credits.
- Publisher
- happyhorse
- Fixed release
- 1.0.0
marketplace-publisher
Etsy Style Marketplace Build for AI Agents. Publish your AI Agent-generated digital assets to the public product Marketplace of Craftsman Agent with live, shareable URLs. Support images, 3D models, audio, and video. Set Credits, Find Agentic Manufacturing on Demand Supplier of Physical Products!
- Publisher
- AI-Hub-Admin
- Fixed release
- 1.0.0
gallery-publisher
Instagram built for AI Agents and their creations. Publish your AI Agent-generated digital assets to the Instagram Style Public Gallery of Craftsman Agent with live, sharable URLs. Support images, 3D models, audio, and video with dedicated viewers.
- Publisher
- AI-Hub-Admin
- Fixed release
- 1.0.0
weibo-video-to-transcript
基于视频内容产出文字与文案。当用户提出视频转文字、视频提取文案、字幕提取、语音转写、视频摘要、视频理解、内容提取、素材处理、自媒体素材、内容创作等需求时使用。输入支持本地视频文件和抖音、微博 小红书等平台视频链接,云端大模型转写并可按 Prompt 定制输出。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
Elderly Loneliness Detection & Warm Companionship | 独居老人孤独情绪识别与温暖陪伴
Using a fixed camera in the home of a solitary-living elderly person or in a private nursing-home room, the system analyzes daily activity video and detects loneliness-related behaviors: prolonged solitude (no social interaction), static gazing (long-time fixation with no purposeful activity), sighing (rapid chest/abdomen rise-fall with exhale), and talking-to-self (mouth activity with no conversation partner). | 通过独居老人家中或养老院单人房的固定摄像头,分析老人日常行为视频,检测孤独相关行为:长时间独处(无社交互动)、静止发呆(长时间凝视一处无目的活动)、叹气(胸腹快速起伏伴呼气)、自言自语(口部活动但无对话对…
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.8
Turtle Pneumonia Symptom (Open-Mouth Breathing) Detection | 龟类张嘴呼吸(肺炎征兆)识别
Through fixed enclosure cameras, the system analyzes mouth and nasal videos of turtles to detect abnormally frequent open-mouth breathing in non-feeding states (mouth opening frequency unusually elevated), as well as the presence of mucus (reflective spots or strands) or nasal discharge around the mouth and nose. | 通过龟缸固定摄像头,分析龟类的口鼻部视频,检测龟在非进食状态下(未摄食时)口部频繁开合(张嘴呼吸,频率异常增高),以及口鼻区域是否有黏液(反光点或丝状物)或鼻腔分泌物。当同时或单独出现上述症状时,输出'肺炎风险提示',提醒饲养者检查环境温度、水质,并及时隔离治疗。
- Publisher
- smyx-skills
- Fixed release
- 1.0.11
Video2text Ai 1.0.1
把视频变成文字并产出可直接使用的文案。当用户发来视频链接(抖音、小红书等)或本地视频文件,提出视频转文字、视频提取文案、视频转稿、字幕提取、语音转写、视频总结等需求时使用。云端大模型转写并自动剔除语气词、口误与重复内容,支持用自定义 Prompt 生成总结、改写、金句提取、分镜头、中英翻译等风格化内容。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-extract-video-text
把视频变成文字并产出可直接使用的文案。当用户发来视频链接(抖音、小红书等)或本地视频文件,提出视频转文字、视频提取文案、视频转稿、字幕提取、语音转写、视频总结等需求时使用。云端大模型转写并自动剔除语气词、口误与重复内容,支持用自定义 Prompt 生成总结、改写、金句提取、分镜头、中英翻译等风格化内容。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
Audiobook Skill
Convert a text or Markdown manuscript into clean, long-form spoken narration (MP3, optional MP4) fully locally with Kokoro-82M via MLX. Use when asked to narrate a document, make an audiobook, read a manuscript aloud, or produce spoken-word audio/video from written text. Sentence-aware chunking, resumable jobs, configurable voices/languages, verifiable outputs. Text-to-speech only (not transcription).
- Publisher
- High Noon Office
- Fixed release
- 1.0.0
YouTube Transcript
Use when the user wants a YouTube video's transcript fetched, wants to summarize/analyze/quote a YouTube video by its spoken content, wants to search YouTube (globally or a channel handle's videos), wants a channel handle resolved to its channel ID, or wants the videos in a YouTube playlist. Calls the getyoutubetranscript.com public API - requires an API key (free tier available, no card required).
- Publisher
- TubeAgentKit
- Fixed release
- 1.0.3
Video Subtitle Translation & Dubbing
Multi-language video subtitle translation and automatic dubbing skill (supports English, Chinese, Japanese, Spanish, French, German, Korean, etc.).
- Publisher
- zbjincheng
- Fixed release
- 0.1.4
douyin-video-read
读取抖音视频的内容——元信息、官方 AI 章节要点,以及通过逐帧截图 + 字幕 OCR 得到完整口播讲稿。用户分享抖音链接(v.douyin.com / douyin.com/video/xxx)并希望了解视频讲了什么、提取文案、拿到文字稿时使用。触发词:抖音链接、抖音视频、这个视频讲了什么、提取视频文案、抖音视频转文字、视频字幕、看看这个视频、视频内容。
- Publisher
- 52Siriyue
- Fixed release
- 1.0.0
qwen-asr
Local speech-to-text using Qwen3-ASR (CPU-only, no API key, no cloud). Use when: (1) a voice message or audio file needs transcription, (2) user asks to transcribe audio, (3) speech-to-text is needed. Supports offline, segmented, and streaming modes. macOS and Linux only.
- Publisher
- lizhuo
- Fixed release
- 0.1.1
Guaikei Video Transcript Extractor
将视频内容转写并加工为可复用文案。典型触发:视频转文字、提取视频文案、视频转稿、字幕提取、视频总结、金句提取、视频内容分析、会议纪要、课程拆解、直播复盘、采访整理。支持本地文件与抖音、小红书等平台链接,云端解析,自定义 Prompt 可生成总结、改写、分镜头、翻译等。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-video-transcript-extractor
处理一切「视频变文字」的需求。覆盖视频转写、语音转文字、字幕提取、视频内容分析、会议纪要、课程拆解、直播复盘、采访整理、探店脚本、短视频二创等场景。输入为本地视频文件或抖音、小红书等公网视频链接,输出为剔除冗余后的干净文字稿,亦可按 Prompt 定制任何风格的衍生文案。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-video-transcript
视频转文字、字幕提取、视频总结、会议纪要、课程拆解、直播复盘、采访整理、短视频二创、口播稿生成。当用户需要把视频或音频内容变成文字,或基于视频内容产出文案时使用。支持本地视频与抖音、小红书等平台链接,云端转写并剔除语气词口误,可用自定义 Prompt 生成总结、改写、金句提取、分镜头脚本、中英翻译等风格化内容。 | 156 | 触发词密度最高,搜索友好
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-video-wenzi-fetcher
视频转文字 + 文案产出二合一。用户丢来一个视频(链接或本地文件)要求转写、总结、提取文案、改写二创时使用。基于云端大模型完成语音转写与语义理解,自动过滤口误与冗余,按需输出完整文字稿或指定风格的成品文案。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
financial-card-video
用 JSON 生成报纸风竖版金融卡片视频 (1080x1920).
- Publisher
- bosslay
- Fixed release
- 0.1.0
guaikei-v2t
将视频转为文字稿。适用于「把这个视频的文字提出来」「帮我转写这个视频」「视频总结一下」等指令。支持抖音、小红书等平台链接直连,也支持本地视频文件。云端大模型转写,输出剔除语气词后的通顺文案,可通过 Prompt 定制总结、改写、金句、分镜头、翻译等输出。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-video-to-text
把视频里的说话内容提取成干净文字。用户给你视频链接或本地文件,说「视频转文字」「提取文案」「转个稿」「提取字幕」时使用。自动去掉语气词、口误和重复内容,还能按用户要求总结要点、改写成小红书或抖音文案、提取金句、翻译成英文。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-video2text-ai
把视频变成文字并产出可直接使用的文案。当用户发来视频链接(抖音、小红书等)或本地视频文件,提出视频转文字、视频提取文案、视频转稿、字幕提取、语音转写、视频总结等需求时使用。云端大模型转写并自动剔除语气词、口误与重复内容,支持用自定义 Prompt 生成总结、改写、金句提取、分镜头、中英翻译等风格化内容。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
guaikei-video2text
将视频转为文字与结构化文案的技能。当用户提出"视频转文字 / 视频提取文案 / 视频转稿 / 字幕提取 / 视频总结 / 视频内容分析 / 会议纪要 / 课程拆解 / 直播复盘 / 采访整理 / 短视频二创脚本 / 口播稿 / 小红书文案 / 抖音文案 / 公众号文案"等需求时使用。支持本地视频文件与抖音、小红书等平台视频链接,调用千问大模型自动转写并剔除语气词、口误与重复内容,并可通过自定义 Prompt 生成总结、改写、金句提取、分镜头、中英翻译等风格化文案。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
extract-youtube-transcript
Extract plain-text transcripts from YouTube videos using a local Python script. Use when the user wants to fetch, extract, or get a transcript from a YouTube video URL, analyze YouTube video content as text, or needs subtitles/captions from a video.
- Publisher
- Joe Hu
- Fixed release
- 2.0.0
sn-motion-html
从主题策划到 AI 图片、视频和滚动网页的一体化制作
- Publisher
- SenseNova-Skills
- Fixed release
- 2026.9.12
youtube-transcript-native-node
Extract a clean plain-text transcript from existing YouTube captions - native Node.js, zero npm dependencies. Use when the user asks to summarize, quote, or extract captions/transcript text from a YouTube URL. Wraps the `yt-dlp` binary on PATH; writes subtitles to a temp dir, parses .vtt captions, strips timestamps/HTML tags, and prints clean text or JSON. No API keys required.
- Publisher
- Jeremy Westburg
- Fixed release
- 1.1.27
anything2explainer
Turn any topic into animated explainer videos — Remotion-driven, TTS voiceover, bilingual (Chinese/English)
- Publisher
- Johnson
- Fixed release
- 0.1.0
video-to-markdown
Analyze any YouTube, Facebook, or Instagram video URL and generate a comprehensive Markdown reference document by combining AI vision analysis of extracted frames with full video transcription. Use this skill when a user shares a video URL and wants a summary, notes, breakdown, or reference document from it. Triggers on "analyze this video", "summarize this video", "break down this video", "create notes from this video", "watch this video and explain it", "video to markdown", "pull notes from this", or any request…
- Publisher
- Stewart Green
- Fixed release
- 0.1.1
Weekly episode feed gate
Check deployed podcast feed
- Publisher
- PowMCP
- Fixed release
- 1.0.0
Diagnose a rejected e-invoice
Explain conformance rejects
- Publisher
- PowMCP
- Fixed release
- 1.0.0
wechat-article-extractor
Extract metadata and content from WeChat Official Account articles. Use when user needs to parse WeChat article URLs (mp.weixin.qq.com), extract article info (title, author, content, publish time, cover image), or convert WeChat articles to structured data. Supports various article types including posts, videos, images, voice messages, and reposts.
- Publisher
- 苍何
- Fixed release
- 1.0.0
Video Transcript Method
视频文字稿提取方法。核心能力:将任何在线视频的语音内容提取为结构化文字稿(带时间戳+元信息+要点总结)。覆盖从视频URL解析、音频提取、CC字幕检测、Whisper语音识别、元信息获取、语义分段到结构化文字稿输出的全流程。通用方法,不绑定任何特定视频平台。触发词:视频文字稿、视频转文字、字幕提取、语音转录、vid...
- Publisher
- 波动几何
- Fixed release
- 1.0.0
Pet Video Script Domain
萌宠短视频创作知识参考库(领域负载物)。覆盖宠物角色建档、故事概念设计、分镜脚本生成、AI视频提示词四大域,含15种任务类型。为AI视频生成工具(Seedance 2.0/Kling/Veo等)提供结构化分镜脚本与英文提示词。触发词:萌宠视频、宠物脚本、pet video script、萌宠短视频、宠物分镜、AI...
- Publisher
- 波动几何
- Fixed release
- 1.0.0
openai-whisper-api
Transcribe audio via OpenAI Audio Transcriptions API (Whisper).
- Publisher
- aperdigaokbk
- Fixed release
- 1.0.0
video-understanding
Watch and analyze video URLs or files using Gemini native agentic video understanding for summaries, timestamped answers, transcripts, and visual reconstruction.
- Publisher
- bill492
- Fixed release
- 2.0.4
YouTube Data API - Search, Videos, Comments, Transcripts, Channels
Search YouTube and retrieve videos, shorts, comments, transcripts, streams, and channel data as structured JSON. 15 endpoints across video and channel surfaces.
- Publisher
- scavio-ai
- Fixed release
- 1.0.3
Facebook API - Pages, Posts, Reels, Groups, Events
Pull a Facebook page's profile, posts, reels and photos, one post with its comments, a single reel/video with downloadable URLs, a public group and its posts, an event, and hashtag posts. 11 endpoints, 1 credit each, structured JSON.
- Publisher
- scavio-ai
- Fixed release
- 1.0.3
Youtube Public Api
Read-only YouTube public-data agent skill — search YouTube videos, video/channel/playlist metadata, public comments, related videos, and transcripts, normali...
- Publisher
- ReplyNodes
- Fixed release
- 2.0.1
alibabacloud-live-assistant
Read-only diagnostics for Alibaba Cloud Live: stream quality checks (codecs, bitrate, GOP, B-frames, A/V sync), CDN edge-node probing, local recording/snapshot, traffic-theft analysis on abused live domains, and signed push/pull test URL generation; never changes any configuration. Use when the user reports live stream stuttering, pixelation or latency, push/pull stream failures, audio-video out of sync, wants to probe live CDN nodes or record a stream, suspects live traffic theft or anomalous billing, needs a URL…
- Publisher
- alibabacloud-skills-team
- Fixed release
- 0.0.1
Doubao Asr
Transcribe recorded audio files to text via Doubao Seed-ASR 2.0 (豆包录音文件识别模型2.0) from ByteDance/Volcengine. Best-in-class Chinese speech recognition with spea...
- Publisher
- vahnxu
- Fixed release
- 0.23.0
youtube-research
Pulls structured YouTube data — video and channel details, transcripts/captions, comments, playlists, and search — via the Crawlora API as clean JSON, with no yt-dlp or HTML scraping. Use when the user provides a YouTube URL or asks for a transcript, comments, channel/video metadata, or video search results.
- Publisher
- Tony Wang
- Fixed release
- 1.0.9
Agent Phone Number & SMS - Anima
Give your AI agent its own real US phone number - inbound and outbound SMS, and voice calls with transcripts, in seven languages. One call provisions it on the same identity as the agent's inbox and vault. Use when an agent needs to text, call, or be reached on a number of its own.
- Publisher
- (bogdiyan) Diyan Bogdanov
- Fixed release
- 1.0.0
gemini-omni
Create Gemini Omni voice resources, character resources, and Flash Preview or multimodal text-to-video tasks through RunAPI. Use when the user asks an agent to create or manage Gemini Omni audio voices, character resources, or video. Default to the RunAPI CLI for one-off calls; use SDKs only when integrating RunAPI into an app or backend.
- Publisher
- RunAPI
- Fixed release
- 0.3.3
Wan 3.0 Prime Reference to Video — Reference-Guided Video on RunComfy
Wan 3.0 Prime Reference to Video generates video clips from reference images, reference videos and reference audio on RunComfy. Wan 3.0 Prime Reference to Video binds up to 10 reference images, 5 reference videos and 5 reference audio clips to a prompt that names them as Image 1, Video 1 and Audio 1, so a character, product or location stays consistent across a 2 to 30 second shot at 480p, 720p or 1080p with a synchronized audio track. Wan 3.0 Prime Reference to Video runs on the fast Wan 3.0 Prime tier (wan3.0-vi…
- Publisher
- mew
- Fixed release
- 1.0.1
TinkerClaw Jarvis Voice
Turn your AI into JARVIS. Voice, wit, and personality — the complete package. Humor cranked to maximum.
- Publisher
- Oscar Serra
- Fixed release
- 3.2.1
Elderly Drinking-Cup Pickup Frequency (Dehydration Risk) | 老年人饮水杯拿起频率(脱水风险)
Using a fixed camera in the living room or kitchen, the system analyzes video of the water-cup placement area (e.g., coffee table, dining table), detects hand-to-cup contact actions (pickup, putdown), and counts daily cup-pickup events (an indirect proxy for water intake). | 通过客厅或厨房固定摄像头,分析水杯放置区域(如茶几、餐桌)的视频,检测手部与水杯的接触动作(拿起、放下),统计每日水杯拿起次数(间接反映饮水量)。当每日拿起次数低于预设阈值(如每天少于6次)时,输出'脱水风险'提醒,建议家属或护理人员督促老人增加饮水。
- Publisher
- smyx-skills
- Fixed release
- 1.0.10
webnovel-serial-audio
Turn a serialized web novel into chapter-by-chapter webnovel audiobook audio with one consistent narrator. This serialized webnovel audio studio and web novel narration shop records the current chapter first, then continues later chapters in the same voice so listeners can follow a serial novel TTS library as new chapters land. Use it for chapter audiobook production, webnovel chapter audio, serial fiction narration, and ongoing web novel listening.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
unattended-live-avatar
Turn one shop portrait and short welcome, product, FAQ, and close scripts into talking-avatar clips a store can loop overnight. This unattended live avatar and night-shift avatar studio can clone or pick a voice, then produce a talking avatar livestream set for welcome, product, FAQ, and close so an unattended livestream or digital human livestream can keep a shop loop presenter on camera. Use it for overnight digital human clips, live room avatar loops, and a night-shift talking-head that keeps the room open.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
sleep-story-voice
Turn already-written short sleep stories into one spoken clip per labeled story. This sleep story voice studio records each short sleep-story voice from the pages the producer already wrote. Use it for sleep stories, bedtime story reads, bedtime narration, and short sleep-story voice packs.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Restock Talking Clips
Turn seller-supplied restock facts and an already-written restock script into one talking clip per still. This restock talking studio turns each authorized still into a 2 to 15s restock talking clip from the written line. Use it for restock talking videos, restock announcement talks, and restock drop talking clips.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Walk-in Talking Clips
Turn seller-supplied floor-plan facts and an already-written walk-in script into one talking clip per still. This rental walk-in studio turns each authorized still into a 2 to 15s walk-in talking clip from the written line. Use it for rental walk-in videos, listing walk-in talks, and property walk-in talking clips.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Radio Drama Ad Breaks
Turn written radio-drama ad-break lines into one spoken bumper clip per labeled slot. This radio-drama bumper studio records each pre-roll read, post-roll read, and sponsor bumper from the copy the producer already wrote, then delivers 8 to 20 bumper audio files. Use it for radio drama ad-break voiceovers that keep one bumper on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.3
Context Compactor CLI
Condense a long agent session transcript into a compact handoff memo: decisions, tasks, risks, facts & links. Credential values are auto-redacted before output (best-effort - review memos before sharing; --strict drops whole suspect lines). EN/RU heuristics, zero dependencies. The CLI prints what it reads and where it writes. Use ONLY with the user's explicit consent: tell the user which transcript file will be read.
- Publisher
- ViBo
- Fixed release
- 1.1.6
office-floor-tour
Turn seller-supplied office floor stills into one office floor tour clip per labeled still. This office floor video studio lays out each floor walkthrough clip from the seller-supplied floor photo, then delivers an office listing clip and a floor walkthrough video as separate files. Use it for office floor packs that give each still its own clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
New Manager Week Voice Pack
Turn a written first-week checklist into one new manager week voice clip per labeled cue. This first week voice pack studio records each new manager voice and first week checklist audio from the list the office already wrote, then delivers 8 to 20 new manager week clip files. Use it for manager onboarding voice packs that keep one cue on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Incident Brief Voice Pack
Turn a written incident briefing script into one incident brief voice clip per labeled cue. This incident brief voice studio records each accident notice audio and safety brief read from the script the office already wrote, then delivers 8 to 20 incident brief voice pack files. Use it for incident brief clip packs that keep one brief line on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Hotel Amenity Clips
Turn hotel amenity stills the property already took into one hotel amenity video per labeled still. This hotel facility clip studio makes each amenity still clip from the photo the property already took, then delivers a hotel facility video, a pool amenity video, and a gym amenity clip from photos already on file. Use it for hotel amenity packs that give each still its own clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Homeroom Week Voice Pack
Turn a written homeroom week plan into one homeroom week voice clip per labeled cue. This weekly class voice studio records each week plan voice and class notice audio from the plan the teacher already wrote, then delivers 8 to 20 homeroom voice pack files. Use it for homeroom weekly voice packs that keep one arrangement on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Holiday Homework Voice Pack
Turn a written holiday homework list into one holiday homework voice clip per labeled cue. This homework assignment voice studio records each homework instruction audio and assignment list voice from the list the teacher already wrote, then delivers 8 to 20 homework voice pack files. Use it for holiday homework clip packs that keep one assignment on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
hiring-avatar-studio
Turn one HR or founder portrait and a job brief into one talking-avatar hiring video per open role. This hiring avatar studio and recruitment avatar workflow can clone or pick a voice, then produce a hiring talking head clip that walks through the role, requirements, and next step. Use it for recruiting video, job posting video, job opening presenter clips, and a hiring video studio that keeps each new role on camera.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Hall Guide Voice Pack
Turn a written hall window list into one hall guide voice clip per labeled cue. This window list voice studio records each hall window audio and window cue voice from the list the hall already wrote, then delivers 8 to 20 hall guide voice pack files. Use it for hall guide clip packs that keep one window on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Grid Notice Voice Pack
Turn a written grid-matter list into one grid notice voice clip per labeled cue. This grid matter voice studio records each community grid notice audio and street grid voice from the list the office already wrote, then delivers 8 to 20 grid notice voice pack files. Use it for grid notice clip packs that keep one matter on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Game UI Voice Pack
Turn written game UI lines into one game voiceover clip per labeled cue. This game UI voice studio records each button voice clip, win voice line, and fail voice line from the copy the studio already wrote, then delivers 8 to 20 game cue audio files. Use it for game UI voice packs that keep one cue on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Elder Checkup Voice Pack
Turn a written elder-checkup schedule into one elder checkup voice clip per labeled cue. This checkup schedule voice studio records each elder checkup notice audio and community checkup read from the schedule the office already wrote, then delivers 8 to 20 elder checkup voice pack files. Use it for elder checkup clip packs that keep one schedule item on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
douyin-video-script-maker
Create a Douyin short-video script, Douyin spoken script, or Douyin product-video script from a topic, product or service facts, audience, and creator voice. This AI Douyin script writer produces three hook options, a ready-to-film short-video script, shot-by-shot beats, natural spoken lines, subtitle cues, title ideas, hashtags, and a comment prompt for knowledge sharing, local business, product demos, reviews, unboxings, shop content, and creator series. The chosen title then becomes a matching vertical 9:16 Dou…
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
Customer Onboard Voice Pack
Turn a written customer-onboarding step list into one customer onboard voice clip per labeled cue. This onboarding-step voice studio records each customer onboarding audio and step list voice from the steps the enablement team already wrote, then delivers 8 to 20 customer onboard voice pack files. Use it for onboarding clip packs that keep one written step on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Credit Rights Voice Pack
Turn a written credit-card benefits table into one credit-rights voice clip per labeled cue. This benefits-table voice studio records each card-rights audio and rights-table voice from the list the desk already wrote, then delivers 8 to 20 credit-rights voice pack files. Use it for card-rights clip packs that keep one table item on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
corporate-podcast-studio
Turn executive talking points or company column copy into a serialized corporate podcast with one consistent host voice. This executive podcast studio and company podcast series records each branded episode, then continues the brand podcast as an executive column so listeners recognize the same host across the corporate audio series. Use it for leadership podcast episodes, C-suite podcast audio, and company thought-leadership shows.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Guaikei Kuaishou Trending Video Fetcher
按关键词搜索快手视频、抓取博主公开作品和视频评论,返回结构化JSON,支持爆款选题和竞品监控分析。
- Publisher
- engheng-art
- Fixed release
- 1.0.0
詹明明·找对标
📐 詹明明·找对标 ——找对标。给一个方向就去抖音/小红书/视频号搜人,给一个名字就去认人;三筛过滤(赚钱 / 看懂 / 能仿)挑出真正值得抄的那个,然后把他的**三批内容**全扒下来——最早 10 条(他怎么起的号)、数据最好 10 条(什么能爆)、最新 10 条(他现在在哪)——封面、大字、逐字稿、互动数据一条不落,最后出完整拆解和抄袭路线图。 触发方式:/zmm-benchmark、/找对标、/对标、「我该学谁」「帮我找个对标」「这个号值不值得学」「把这个博主拆一下」「他是怎么起号的」「扒一下这个账号」 Find and dissect a benchmark creator: search by direction or by name, filter on money/understandable/copyable, then pull the earliest 10, best-performing 10, and latest 10 posts with covers, cover text, transcripts and engagement data, and produce a full teard…
- Publisher
- 詹明明
- Fixed release
- 0.1.4
Class Duty Voice Pack
Turn a written class duty roster into one class duty voice clip per labeled cue. This duty roster voice studio records each class duty audio and roster reminder from the list the teacher already wrote, then delivers 8 to 20 class duty voice pack files. Use it for duty clip packs that keep one roster item on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Census Notice Voice Pack
Turn a written census schedule into one census notice voice clip per labeled cue. This census schedule voice studio records each census notice audio and arrangement voice from the list the office already wrote, then delivers 8 to 20 census notice voice pack files. Use it for census clip packs that keep one schedule item on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Civil Affairs Voice Pack
Turn a written civil-affairs materials list into one civil affairs voice clip per labeled cue. This materials list voice studio records each civil affairs instruction audio and window materials voice from the list the office already wrote, then delivers 8 to 20 civil affairs voice pack files. Use it for civil affairs clip packs that keep one materials item on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Blackboard One-Shot Clips
Turn authorized blackboard photos and teacher-supplied facts into one blackboard one-shot clip per photo. This one-shot blackboard studio makes a silent 2-15s blackboard clip from each board photo clip. Use it for classroom blackboard video and lesson-board clip sets that stay one photo one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Assembly One-Step Clips
Turn authorized stills and seller-supplied step facts into one assembly step video per still. This one-step clip studio turns each still into a product assembly clip. Use it for furniture assembly video and how-to step clip work that stays one photo one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
apartment-kitchen-tour
Turn one kitchen photo the listing already uses into one short clip for the listing page. The photo is the clip's first frame, so that frame keeps the room's layout, and from there the camera eases in, drifts between foreground and background, or light sweeps across the counter. Use it for kitchen tour video, apartment kitchen clip, listing photo animation, and real estate room video work that stays one photo one clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
YouTube Lesson Card Set
Turn public YouTube lesson captions into a set of 4 to 8 takeaway cards. This lesson card studio reads the lesson captions, pulls the key points, and lays out one card per point for lessons and tutorial videos.
- Publisher
- beatra-ai
- Fixed release
- 0.1.3
voice-clone-series-studio
Clone one voice you own and keep using it for a series of episodes, updates, and lessons. This voice cloning and AI voice clone studio for series narration creates a reusable AI voice, custom AI voice, and brand voice clone from a clean sample, then turns each new script into narration in that same voice so a podcast, course, or creator series stays consistent. Use it to clone my voice, build a podcast voice clone, and generate the next episode voiceover or recurring voiceover without recasting.
- Publisher
- beatra-ai
- Fixed release
- 0.1.3
Founder IP Avatar Studio
Turn one authorized founder or expert portrait and a weekly script into a talking-head IP video in that person's likeness and voice. This founder avatar and digital avatar studio can clone that person's voice and produce a stable founder talking head clip for founder updates, expert explainers, and brand announcements. Use it for a digital human spokesperson, talking head founder series, expert avatar video, brand digital human, AI presenter, AI spokesperson video, and recurring presenter video.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
Account Opening Voice Pack
Turn written account-opening steps into one account opening voice clip per labeled cue. This account opening guidance studio records each teller guidance clip and KYC step voice from the steps the desk already wrote, then delivers 8 to 20 account opening audio files. Use it for brokerage onboarding audio that keeps one step on each clip.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
WaveSpeedAI MiniMax Speech 2.6 TTS
Convert text to speech using MiniMax Speech 2.6 Turbo via WaveSpeed AI. Features ultra-human voice cloning, sub-250ms latency, 40+ languages, emotion control, and 200+ voice presets. Use when the user wants to generate speech audio from text.
- Publisher
- WaveSpeed
- Fixed release
- 2.0.1
WaveSpeedAI Wan 2.6 Video Generation
Generate videos using Alibaba's Wan 2.6 model via WaveSpeed AI. Supports text-to-video and image-to-video generation with up to 15 seconds duration at 720p or 1080p. Features audio-guided generation, prompt expansion, multi-shot mode, and configurable seeds. Use when the user wants to create videos from text prompts or animate images.
- Publisher
- WaveSpeed
- Fixed release
- 2.0.1
WaveSpeedAI Seedance 1.5 Pro Video Generation
Generate videos using ByteDance's Seedance V1.5 Pro model via WaveSpeed AI. Supports text-to-video and image-to-video generation with 4-12 second duration at up to 1080p. Features audio generation, camera control, smart duration, and configurable seeds. Use when the user wants to create videos from text prompts or animate images.
- Publisher
- WaveSpeed
- Fixed release
- 2.0.1
WaveSpeedAI Veo 3.1 Fast Video Generation
Generate and extend videos using Google's Veo 3.1 Fast model via WaveSpeed AI. Supports text-to-video, image-to-video, and video extension. Features up to 4K resolution, audio generation, and chained extensions up to 148 seconds. Use when the user wants to create videos from text or images, or extend existing Veo-generated videos.
- Publisher
- WaveSpeed
- Fixed release
- 2.0.1
WaveSpeedAI Infinitetalk Talking Avatar Video Generation
Generate talking head videos from a portrait image and audio using WaveSpeed AI's InfiniteTalk model. Produces lip-synced video up to 10 minutes long at 480p or 720p. Supports optional mask images to target specific faces and text prompts for additional guidance. Use when the user wants to animate a face with audio or create talking avatar videos.
- Publisher
- WaveSpeed
- Fixed release
- 2.0.1
industrial-promo-video
产业园招商短视频营销专家:AI生成脚本视频、多平台投放方案、客户意向识别留资、线索沉淀腾讯文档。
- Publisher
- perrykono-debug
- Fixed release
- 1.1.0
watch-cli
Watch any social video → get an architecture diagram, working component, runnable notebook, or step-by-step cheat sheet — automatically.
- Publisher
- Son Piaz
- Fixed release
- 1.0.2
openai-transcription
Transcribe uploaded audio through RunAPI with an OpenAI-compatible API. Use for one-off transcription, subtitle output, multilingual hints, or application integration. Prefer the RunAPI CLI for manual requests and the target-language SDK for production integration.
- Publisher
- RunAPI
- Fixed release
- 0.1.3
fish-audio
Create account-owned Fish Audio voice resources, attempt to reuse their IDs, or generate MP3/WAV speech through RunAPI. Use for voice resource management, one-off speech generation, or application integration. Prefer the RunAPI CLI for one-off requests and the target-language SDK for production integration.
- Publisher
- RunAPI
- Fixed release
- 0.4.1
openai-tts
Generate MP3 speech with OpenAI TTS through RunAPI. Use for one-off speech generation or application integration. Prefer the RunAPI CLI for one-off requests and the target-language SDK for production integration.
- Publisher
- RunAPI
- Fixed release
- 0.1.3
Seedance 2.5 Image to Video — 720p Still-to-Video with Native Audio
Seedance 2.5 Image to Video animates one still image into a 4-30 second 720p cinematic clip with optional synchronized native audio. Seedance 2.5 Image to Video runs on RunComfy through the RunComfy CLI, and this skill documents the full four-field schema — prompt, image, duration, generate_audio — plus the $0.35 per second pricing. Seedance 2.5 Image to Video takes exactly one image and has no aspect-ratio control, so the output ratio follows your source still and product shots stay composed as photographed. Reac…
- Publisher
- mew
- Fixed release
- 1.0.0
Agent Subtitle Translator
Translate SRT, VTT, and ASS subtitles safely
- Publisher
- 二狗子你变了
- Fixed release
- 1.0.9
Seedance 2.5 Reference to Video — 1080p Reference-Guided Video on RunComfy
Seedance 2.5 Reference to Video on RunComfy. Seedance 2.5 Reference to Video is ByteDance's reference-guided video endpoint: it takes up to 9 reference images, 1-3 reference video clips, and 3 reference audio files and returns a 4-30 second 1080p clip with native synchronized audio, identity and art direction locked to your references. This skill calls Seedance 2.5 Reference to Video through the RunComfy CLI, and Seedance 2.5 Reference to Video bills $0.53 per counted second. Triggers on "seedance 2.5", "seedance…
- Publisher
- mew
- Fixed release
- 1.0.0
Seedance 2.5 Image to Video — Animate a Still with Native Audio
Seedance 2.5 Image to Video animates one still image into a 4-30 second 720p cinematic clip with optional synchronized native audio. Seedance 2.5 Image to Video runs on RunComfy through the RunComfy CLI, and this skill documents the full four-field schema — prompt, image, duration, generate_audio — plus the $0.35 per second pricing. Seedance 2.5 Image to Video takes exactly one image and has no aspect-ratio control, so the output ratio follows your source still and product shots stay composed as photographed. Reac…
- Publisher
- mew
- Fixed release
- 1.0.0
Asr Claw
Speech recognition CLI for AI agent automation. Transcribe audio from stdin, files, or URLs.
- Publisher
- 任嘉
- Fixed release
- 1.1.1
Clothes Try-On Studio
Virtual clothes try-on studio using YouCam (Perfect Corp) AI. Swap outfits onto the user's photo, optionally change the background, and optionally turn the result into a short motion video (turn / runway / pose). Use for "換衣", "虛擬試穿", "outfit try-on", "try on clothes". Do NOT use for makeup, hair, or skin analysis.
- Publisher
- YouCam API
- Fixed release
- 1.0.1
音潮 AI 音乐创作
使用音潮(YinChao)生成可播放的完整 AI 歌曲、纯音乐和 BGM;支持文字或歌词转歌曲、歌词谱曲演唱、参考音频风格创作、歌曲续写或延长,以及纯歌词创作。当用户要求写歌、创作歌曲、生成音乐、AI 作曲、把歌词唱出来、制作 BGM 或纯音乐、仿写歌曲或续写音乐时使用。 Use for AI music generation, song generation, instrumental music and BGM generation, text-to-music, lyrics-to-song, songwriting, vocal music, reference audio, and music extension; not for music search or playback, TTS, transcription, audio conversion, or mixing.
- Publisher
- Joey Yuan
- Fixed release
- 1.5.0
finding-speaking-opportunities-on-twitter
Finds speaking opportunities and event organizer contacts on Twitter using apidojo's Twitter scrapers. Triggers when the user asks to: find speaking opportunities on Twitter, discover conferences looking for speakers on X, find event organizers calling for speaker submissions, identify call-for-speakers announcements in an industry on Twitter, find podcast or summit hosts looking for guests, discover virtual event opportunities for thought leadership, or build a speaking opportunity pipeline from Twitter. Returns…
- Publisher
- API Dojo
- Fixed release
- 1.0.0
scraping-tiktok-posts-by-music
Extracts TikTok posts using a specific audio track or sound using apidojo's TikTok Music Scraper on Apify. Triggers when the user asks to: get all TikTok videos using a specific sound, find TikTok posts using an audio clip, scrape TikTok videos using a trending music track, export TikTok posts made with a song, find creators using a specific sound, or collect TikTok content associated with a music URL. Returns video caption, views, likes, comments, channel info, song metadata, and hashtags per post. Ideal for musi…
- Publisher
- API Dojo
- Fixed release
- 1.0.0
抖音视频总结
提取抖音音频、云端生成字幕并总结重点与行动项
- Publisher
- RuidongYuan-dot
- Fixed release
- 1.0.0
MarkItDown
Convert documents AND web pages to Markdown with Microsoft's MarkItDown CLI (`markitdown`). Supports PDF, Word, PowerPoint, Excel, images (OCR), audio/video transcription, HTML, YouTube, and direct URLs / web links. Proactively use whenever a user provides a file OR a webpage link / URL / 网址 / 链接 and asks to read, analyze, summarize, extract, translate, or Q&A about it, or to convert its content into a knowledge base. ALSO use proactively to cut token cost: when asked to summarize / analyze / extract from a large…
- Publisher
- stwhwing
- Fixed release
- 1.5.3
session-logs
Search, inspect, and analyze prior OpenClaw sessions and transcript history.
- Publisher
- Altair
- Fixed release
- 1.0.0
ai-draw-cue-word-project
Build high-consistency AI character prompts across five model syntax families (natural-language / conversational / MJ-Niji / SD / domestic API) and multi-format outputs (image / storyboard / comic panel / video / 3D) using a weight & ratio precision-control workbench. Ships character-anchor sheets, composition & shots references, posture-emotion mapping, reference-image capability matrix, multi-character spatial relations, storyboard/panel layout rules, speech-bubble positioning, LoRA management, temporal-consiste…
- Publisher
- nohn3043-arch
- Fixed release
- 2.9.0
tiktok-downloader
Download TikTok videos without watermark and extract MP3 audio via the free TikTok Download API (tk.seekubo.com). Give it any TikTok URL (www.tiktok.com, vm.tiktok.com, vt.tiktok.com, m.tiktok.com) and it resolves video metadata plus direct, browser-fetchable CDN download links. No API key or subscription required.
- Publisher
- WangSirMe
- Fixed release
- 2.0.0
Transcribe Media
Fetch and use transcripts from public and local media
- Publisher
- niuzb
- Fixed release
- 1.0.7
Read Aloud
Turn text into natural, playable speech
- Publisher
- niuzb
- Fixed release
- 1.0.2
Dataify YouTube Video Post
Collect YouTube Video Post data and return results
- Publisher
- dataify-server
- Fixed release
- 1.3.0
Dataify YouTube Video By URL
Collect YouTube Video By URL data and return results
- Publisher
- dataify-server
- Fixed release
- 1.3.0
Dataify YouTube Transcript By ID
Collect YouTube Transcript By ID data and return results
- Publisher
- dataify-server
- Fixed release
- 1.3.0
Dataify YouTube Audio By URL
Collect YouTube Audio By URL data and return results
- Publisher
- dataify-server
- Fixed release
- 1.3.0
bullshit-argument-audit
Audit text, transcripts, audio, or video for bogus proof tactics, handwaving, bad citations, and weak argument moves.
- Publisher
- Stanislav Stankovic
- Fixed release
- 1.0.0
youtube-apify-transcript
Fetch YouTube transcripts via the Apify API. Works from cloud IPs (Hetzner, AWS, etc.) by bypassing YouTube's bot detection. Features local caching (free repeat requests) and batch mode. Requires APIFY_API_TOKEN and the Python requests library.
- Publisher
- Robby
- Fixed release
- 1.4.0
elevenlabs-voices
High-quality voice synthesis with 18 personas, 32 languages, sound effects, batch processing, and voice design using ElevenLabs API.
- Publisher
- Robby
- Fixed release
- 2.2.0
ai-draw-cue-word-project
Build high-consistency AI character prompts across five model syntax families (natural-language / conversational / MJ-Niji / SD / domestic API) and multi-format outputs (image / storyboard / comic panel / video / 3D) using a weight & ratio precision-control workbench. Ships character-anchor sheets, composition & shots references, posture-emotion mapping, reference-image capability matrix, multi-character spatial relations, storyboard/panel layout rules, speech-bubble positioning, LoRA management, temporal-consiste…
- Publisher
- nohn3043-arch
- Fixed release
- 2.8.1
yt2gdrive
Sync latest audio or video (mp3/mp4) from YouTube channels to Google Drive.
- Publisher
- yangoal2019
- Fixed release
- 1.0.1
PlaceCall
The phone is your last API. Give your agent a voice to call any US business and take action in the real world. PlaceCall handles reservations, inquiries, and quotes - navigating IVRs, holds, and transfers. You get a verified outcome + full transcript. First 250 calls free. Pay only for outcomes.
- Publisher
- VOYGR
- Fixed release
- 1.0.0
短视频生成专业版
企业级竖版短视频批量生成系统,支持多模板、多语言、品牌定制、团队协作与自动化工作流,提升内容生产效率。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
竖版视频生成免费版
从Markdown脚本一键生成9:16竖版短视频,支持TTS配音、字幕烧录与画面同步,适合个人内容创作者。Use when 需要视频处理、音频编辑、媒体转换、配音生成时使用。不适用于版权受保护的媒体内容处理。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。
- Publisher
- 天轰穿
- Fixed release
- 1.0.3
WhatsApp表情专业版
企业级WhatsApp GIF管理工具,支持批量发送、定时任务、GIF库管理、多账号及营销效果分析,适合团队协作与营销推广。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
音频生成工具-专业版
支持15+专业音频模型,包含TTS、语音克隆、多角色对话、原创音乐生成及管道自动化处理,满足内容团队需求。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
音频生成工具-免费版
轻量级文本转语音工具,支持多语言TTS与基础音效生成,适合个人内容创作。Use when 需要文本翻译、多语言转换、本地化处理时使用。不适用于专业医学法律翻译认证。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。提供结构化输出和错误处理机制。
- Publisher
- 天轰穿
- Fixed release
- 1.0.3
Discord语音工具专业版
面向企业和社区团队的Discord语音AI工具,支持多服务商流式转写、自动重连、多频道管理和权限审计,提升语音交互效率。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
statement-reconciliation
Check a vendor statement against the AP log to confirm every invoice is recorded, and draft a request for any that are missing. Use when a vendor statement arrives.
- Publisher
- skillsandagentsco
- Fixed release
- 1.0.2
youtube-caption-studio
Turn a YouTube link or a pasted transcript into a Chinese spoken script and a remake structure. This video captions and caption extract workflow reads YouTube captions and optional comments, or works from the transcript you already copied, then writes spoken Chinese with the remake beats that follow the original. Use it for video captions, caption extract, YouTube captions, and YouTube transcript work when you need a spoken remake from what the video already said.
- Publisher
- beatra-ai
- Fixed release
- 0.1.3
wechat-channels-script-studio
Turn a product and its confirmed facts into a WeChat Channels short-video script you can film today. This WeChat Channels script writer lays the video out as a segment table second by second, writes the full spoken narration line by line, and places the product-link conversion beats against the exact segments where a viewer decides to tap through — then scores the draft across six dimensions so the weak part is visible before anyone shoots. Hand it a reference video that already sold and it carries that structure…
- Publisher
- beatra-ai
- Fixed release
- 0.1.4
talkies
Self-hosted OpenAI-compatible speech service. /v1/audio/transcriptions fronts 14 open ASR models (Whisper, Parakeet, Nemotron-3.5-ASR, Canary, Sherpa-ONNX, Vosk, plus wav2vec2 and ZIPA phoneme recognizers that emit IPA); /v1/audio/transcriptions/stream accepts live PCM over WebSocket. /v1/audio/speech fronts 3 TTS engines / 4 backends — Kokoro-82M (41 baked voices, PyTorch + ONNX runtimes), the CUDA-only Qwen3-TTS family (voice cloning, preset speakers, voice design), and the CUDA-only Chatterbox Turbo (English, 1…
- Publisher
- Ciprian Mandache
- Fixed release
- 1.3.18
爆款短视频封面
根据封面文案生成高点击率中文竖版短视频封面
- Publisher
- wenmao030
- Fixed release
- 1.0.0
alibabacloud-media-diagnostics
Diagnose playback and streaming problems in user-provided media files and URLs (m3u8 / HTTP / RTMP / RTSP / SRT): moov atom position, codec compatibility (H.265/HEVC, AAC-HE, 10-bit), container issues, HLS playlist and TS segment integrity, bitrate/frame-rate anomalies, audio-video sync, and live latency factors such as B-frames and GOP size. Use when the user reports a video that will not play, shows a black or green screen, stutters or buffers, or asks for an HLS playlist, TS segment, or live stream latency chec…
- Publisher
- alibabacloud-skills-team
- Fixed release
- 0.0.1
Estrus/Mating Behavior Detection | 畜禽发情/配种行为识别
Detects estrus behavior in female livestock from continuous barn videos — including mounting acceptance, standing reflex, restlessness, appetite drop and vulva changes — and outputs an estrus recognition result with the optimal mating time window. | 识别母畜发情期行为特征(爬跨、静立反射等),优化配种时机。
- Publisher
- smyx-sunjinhui
- Fixed release
- 1.0.10
bili-review
抓取 B 站视频的 AI字幕、弹幕与评论,综合生成总结。总结包含:速读卡(秒级判断看不看),详细总结(步骤清单/红黑榜/民间避坑)。触发词:B站视频总结、B站总结、B站评论、看评论区、视频讲了什么、B站深度分析、弹幕分析。
- Publisher
- frazier
- Fixed release
- 2.2.0
cliphi-clips
Turns the user's long videos into ready-to-post vertical clips with captions and branding, using the Cliphi API. Use this skill when the user wants to clip a video, make Shorts / TikToks / Reels from a YouTube video, podcast, livestream or Twitch VOD, find viral moments in a video, or repurpose long video into short clips.
- Publisher
- Cliphi
- Fixed release
- 1.0.0
voice-learn
Improves a voice profile by learning from manual edits. Use after editing generated text to refine registers and close voice drift over time
- Publisher
- athola
- Fixed release
- 1.9.19
vhs-recording
Generates terminal recordings using VHS tape scripts and produces GIF outputs
- Publisher
- athola
- Fixed release
- 1.9.19
media-composition
Combines GIFs and videos into composite tutorials with vertical or grid layouts via ffmpeg
- Publisher
- athola
- Fixed release
- 1.9.19
gif-generation
Converts webm/mp4 video files to optimized GIFs via ffmpeg with configurable quality settings
- Publisher
- athola
- Fixed release
- 1.9.19
Wechat Article Video
公众号转视频号:30–40秒完整短版、Edge配音、同步字幕与封面
- Publisher
- ToBeWin
- Fixed release
- 0.1.0
transcribe.so
Transcribe audio and video with the transcribe.so CLI. Turns YouTube videos, podcasts (Apple Podcasts, Spotify, SoundCloud, Vimeo, Twitch, Loom), direct media URLs, and local audio or video files into speaker-labelled transcripts with timestamped segments, chapters, sections, cited Q&A, and subtitle files (SRT, VTT, karaoke VTT). Use when the user wants a transcript, show notes, chapters, subtitles, quotes, or answers grounded in a recording. 52 languages and dialects.
- Publisher
- Seunghun Sunmoon Lee
- Fixed release
- 1.0.0
Qwen视频智能分析
基于Qwen 3.5 Plus多模态模型,支持本地视频和远程URL,按自定义抽帧频率智能分析视频场景、动作、物体及生成内容摘要。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
播客下载器
从小宇宙下载播客音频与节目说明。从小宇宙(xiaoyuzhoufm。com)下载播客音频和Show Notes。自动转换为MP3格式(兼容Sanag、小游等骨传导蓝牙耳机、水下游泳时离线播放)。支持多种输入格式,输出结构化结果,适用于独立开发者与一人公司效率提升。Use。Use when 需要代码生成、编程辅助、调试测试、开发部署时使用。不适用于无明确技术栈的模糊需求。 when 需要代码生成、编程辅助、调试测试、开发部署时使用。不适用于无技术栈的通用场景。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
播客
规划播客剧集、生成完整脚本、音频后期处理及社交媒体切片,支持多格式和自动化生产流程。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
播客下载器
从小宇宙下载播客音频与节目说明。从小宇宙(xiaoyuzhoufm。com)下载播客音频和Show Notes。自动转换为MP3格式(兼容Sanag、小游等骨传导蓝牙耳机、水下游泳时离线播放)。支持多种输入格式,输出结构化结果,适用于独立开发者与一人公司效率提升。Use。Use when 需要代码生成、编程辅助、调试测试、开发部署时使用。不适用于无明确技术栈的模糊需求。 when 需要代码生成、编程辅助、调试测试、开发部署时使用。不适用于无技术栈的通用场景。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
llm-provider Whisper
Local speech-to-text tool using Whisper CLI for multi-language transcription without requiring an API key, supporting automation and structured output.
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
视频
从Markdown脚本自动生成9:16竖屏短视频,支持自动分镜、音频编辑及媒体转换。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
Dlazy Generate
Generates images, videos, or audio by automatically selecting the appropriate dlazy model based on input prompts and media.
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
课程生成器
从转录稿或文献生成可独立阅读、可溯源验收的结构化课程,也可在用户明确要求时归档既有课程或从已验证素材提取培训方案。本技能应在用户要“把长转录稿整理成课程”“生成总览和章节”“归档课程”“按受众定制课程方案”时使用。不要用于:仅做 ASR 纠错(用 transcription-corrector)、复盘讲课表现(用 lecture-review)、把多篇文章扩写成书(用 article2book)。
- Publisher
- xierluo
- Fixed release
- 2.8.1
Dlazy Generate
Generate images, videos, or audio by automatically selecting the optimal dlazy model, supporting multiple input formats and structured output.
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
Dlazy Audio音频生成
通过dlazy CLI调用15+托管音频模型,支持文本转语音、音乐、音效及语音克隆,提供高效灵活的音频生成服务。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
Dlazy Audio音频生成
通过dlazy CLI调用15+托管音频模型,支持中文文本转语音、音乐、音效和语音克隆,适合自动化音频生成工作流。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
Azure Ai Voicelive P
Real-time Azure voice AI SDK for bidirectional audio/text streaming, speech-to-text, microphone input, customization, and multilingual transcription.
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
AIOZ音频上传
通过AIOZ Stream API快速上传音频文件,支持默认及自定义编码配置,完成后返回HLS/DASH流媒体播放链接。
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
speechfy
Multi-provider Text-to-Speech: Speechify API (primary) + Edge TTS (fallback). Gera .ogg (Opus) para voice messages.
- Publisher
- Rickk Barbosa
- Fixed release
- 1.1.0
openshorts
Turn long videos (podcasts, webinars, streams) into vertical 9:16 clips with subtitles, re-cut them, and publish them to TikTok, Instagram Reels and YouTube Shorts via the OpenShorts API or MCP server. Use when the user wants to clip a video into shorts, find the best moments of a video, restyle captions on a clip, re-cut a clip, schedule or post clips to social platforms, or automate a clipping pipeline.
- Publisher
- mutonby
- Fixed release
- 1.1.0
StoryShort
Create, render and publish AI videos (faceless, UGC ads, movie maker, podcast, music video...) with StoryShort — plus standalone AI images and clips. Full option discovery (voices, caption themes, styles, models with credit costs), exact price quotes before spending, and direct publishing to TikTok, YouTube and Instagram.
- Publisher
- samuelrondot
- Fixed release
- 1.0.0
Video Analyzer
视频分析处理 — 本地视频反编译分析工具。将视频拆解为时间轴剧本、语音转文字、场景分析、跨模态关联和精华摘要,支持多ASR引擎切换(Whisper/Paraformer/SenseVoice)、中文NLP增强、PaddleOCR中文识别。v4.0 新增短视频平台适配(抖音/快手/B站/视频号)和自动剪辑建议(高光检测/冗余标记/EDL导出/字幕样式)。v4.1 新增tiny模型优先体验(75MB低门槛)、说话人分离质量评分、剪映draft.json导出。v4.2 新增场景管理(detect→slice一条链)、短视频爆款预测、实时直播分析(流式ASR+敏感词检测)。v4.3 新增纯音频输入(mp3/m4a/wav播客与录音)、批量队列(SQLite+硬件档位并发)、GPU自动加速(CT2 int8量化)、ASR配置统一(--asr-engine单参数)。
- Publisher
- fyniujin
- Fixed release
- 4.3.0
The Scribble Thing
Turn line art into scribe animation
- Publisher
- Boring Stuff Club
- Fixed release
- 1.0.0
YouTube Video
AI YouTube content creation powered by CellCog. YouTube videos, Shorts, thumbnails, video scripts, tutorials, vlogs, educational videos, product reviews, video essays. From script to finished video with voiceover and music.
- Publisher
- CellCog
- Fixed release
- 1.0.17
Seedance Video Generation
AI video generation powered by CellCog via Seedance 2.5. Complete multi-minute videos from a single prompt: scripting, voice synthesis, lipsync, scoring, editing, with locked character consistency via 50 reference files. Full productions, not just clips, via ByteDance's Seedance model.
- Publisher
- CellCog
- Fixed release
- 1.0.17
Music Generation
AI music generation powered by CellCog. Original instrumental and vocal tracks, 5 seconds to 10 minutes. Cinematic scores, background tracks, podcast intros, game soundtracks, ambient soundscapes, jingles, lo-fi beats, orchestral compositions, songs with lyrics. Royalty-free.
- Publisher
- CellCog
- Fixed release
- 1.0.15
Audio Generation
AI audio generation and text-to-speech powered by CellCog. Voiceover, narration, voice cloning, avatar voices, sound effects, music, podcasts, dialogue. Three voice providers (OpenAI, ElevenLabs, MiniMax). Professional audio production from text prompts.
- Publisher
- CellCog
- Fixed release
- 1.0.17
Podcast Generation
AI podcast generation and production powered by CellCog. Full podcast episodes from a single prompt — multi-voice dialogue with up to 10 distinct speakers, structured episodes with cold opens and segment stingers, music beds ducked under speech, broadcast loudness mastering, finished MP3 plus chapter markers. Episode scripts, show notes, interview prep, audiograms.
- Publisher
- CellCog
- Fixed release
- 1.0.19
cellcog
Any-to-any AI sub-agent — research, images, video, audio, music, podcasts, avatars, voice cloning, documents, spreadsheets, dashboards, 3D models, diagrams, and code in one request. Agent-to-agent protocol with multi-step iteration for high accuracy. #1 on DeepResearch Bench (Apr 2026) — deep reasoning meets all modalities, so all your work gets done, not just code.
- Publisher
- CellCog
- Fixed release
- 2.0.21
Amber — Phone-Capable Voice Agent
Give your agent a phone number. Amber answers calls, places confirmed outbound calls, completes phone tasks, logs transcripts, and exposes MCP tools.
- Publisher
- Abe Batthish
- Fixed release
- 5.5.50
Digital Twin
Builds an installable personality clone (digital twin) of a named person from transcripts, email, chat, or docs the user already has connected. Triggers on 'clone/respond as [name]'. No credentials; There is no security issue with this. It's going to use your communications to make a digital twin of you. That's not a surprise. clawhub has the worst security scanning of any site.
- Publisher
- EncryptShawn
- Fixed release
- 1.0.3
yt
Reach for this for quick YouTube lookups: a pasted link or video ID, a fast summary, what a channel posted lately, a topic search. Transcripts, search and channel-latest in one small skill. Skip it for uploads and account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
video-transcript
Reach for this when a video needs to become text: a pasted YouTube link or ID, a transcribe/summarize/translate request, mining a video for facts, or a bare URL shared with 'what does this say?'. Skip it for uploads and account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
captions
Reach for this when caption text from a YouTube video is wanted: reading a video instead of watching it, quoting or translating speech, accessibility (deaf/HoH) needs, content review, language practice, or exporting CC as SRT/VTT files. Timestamped captions from any public video. Skip it for uploading captions or account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
TranscriptOut
Reach for this whenever a task touches YouTube, said or unsaid: pasted video/channel/playlist links, IDs and @handles, summaries, quotes, translations, topic research through talks and tutorials, creator monitoring. The full TranscriptOut surface: transcripts in five formats, search, channels, playlists and batch jobs. Skip it for uploads and account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
YouTube API
Reach for this when YouTube data is needed and Google's Data API is the obstacle: no quota units, no OAuth consent screens, one Bearer key. Covers transcripts (which Google's API does not serve at all), metadata, search, channels and playlists. Triggers on YouTube links, @handles and creator research. Skip it for uploads and account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
YouTube Data
Reach for this when structured YouTube data is the goal: video metadata, transcripts for analysis, channel upload history, search results or playlist contents, with no Google Cloud project and no quota units. Triggers on YouTube links, creator names and topic research even when unstated. Skip it for uploads and text-only research.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
YouTube Playlist
Reach for this when a YouTube playlist is in play: a pasted playlist link or PL... id, listing a course or series, finding one video inside a long playlist, or turning a whole playlist into transcripts. Skip it for creating playlists or account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
YouTube Channels
Reach for this when a YouTube channel is the subject: a pasted @handle or channel URL, a creator's recent uploads, their full catalogue, or a search inside one channel. Also good for keeping an eye on what somebody publishes. Skip it for creating channels or account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
YouTube Search
Reach for this when the user needs to FIND things on YouTube: videos or channels on a topic, creators covering a subject, tutorials, talks and reviews worth reading, or a channel looked up by name or handle. Also good proactively when researching a topic that video would cover well. Skip it for account chores and text-only research.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
subtitles
Reach for this when subtitles are wanted from a YouTube video: following foreign-language content, reading along, translating speech, language practice, or exporting ready SRT/VTT files. Skip it for uploading subtitles or account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
YouTube Full
Reach for this whenever a task touches YouTube, whether or not the word appears: a pasted watch/shorts/channel/playlist link, a bare 11-char video ID or @handle, a creator to look up, a talk to summarize, quote or translate, research where lectures, tutorials and reviews beat written sources, or a product launch to catch on video. One skill for transcripts, video and channel search, channel browsing, in-channel search, playlists and 4,000-video batch jobs. Skip it for uploads, comments and account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
transcript
Reach for this whenever what was SAID in a YouTube video matters: a pasted link or bare video ID, a summary, quote, translation or fact-check request, notes from a lecture, or research that leans on a specific video. Works even when nobody says the word 'transcript'. Skip it for uploads and account chores.
- Publisher
- artemchuikin
- Fixed release
- 1.0.0
LYGO Flame Ward
LYGO Flame Ward — harden the lattice against disinfo, injected half-truths, corrupted authority, and silent WebAudio device fingerprinting. Default: all sources fabricated until concordance. Seals-first. Ingest gate · flame-scan · endpoint-scan · quarantine · burn-receipt. Local-first, consent-gated. No network, no subprocess, no auto-publish.
- Publisher
- LYRA Agent - LYGO OS
- Fixed release
- 1.0.1
Scavio Facebook API
Pull a Facebook page's profile, posts, reels and photos, one post with its comments, a single reel/video with downloadable URLs, a public group and its posts, an event, and hashtag posts. 11 endpoints, 1 credit each, structured JSON.
- Publisher
- scavio-ai
- Fixed release
- 1.0.0
Video No Subtitle Transcribe
无字幕视频转写兜底方案。当视频(YouTube/Bilibili 等)没有字幕、字幕接口被禁或拉取失败时,用 yt-dlp 下载音频 + faster-whisper 本地转写,输出带时间戳的完整文稿。触发词:视频没字幕、转写视频、字幕被禁、whisper 转写、提取语音内容。
- Publisher
- BCCB
- Fixed release
- 1.0.4
exam-study-guide-predictor
Turn a student's own course material into a two-in-one deliverable: a study guide AND a prediction of what the exam will actually ask, optimized for marks. Use this skill whenever the user wants to prepare or revise for an exam, midterm, final, practical, OSPE, or quiz using their lecture slides, professor audio/voice-note transcripts (English OR Arabic), past exam papers, model answers, lab/practical manuals, textbook chapters, or class notes — even if they just say "help me study for X", "predict my exam", "make…
- Publisher
- Hossam
- Fixed release
- 0.1.0
播客
根据播客音频或转写自动生成章节、亮点和节目说明,支持中文交互及多格式输出,适用于内容制作和自动化场景。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
播客
规划播客剧集、生成脚本、制作音视频与社交媒体切片,支持中文交互与自动化,助力播客内容生产和增长。
- Publisher
- 天轰穿
- Fixed release
- 1.0.3
Dlazy Audio音频生成
通过dlazy CLI调用15+托管音频模型,实现文本转语音、音乐、音效和语音克隆的生成,支持管道串联自动化处理。
- Publisher
- 天轰穿
- Fixed release
- 1.0.2
Qwen视频智能分析
基于 Qwen 3.5 Plus 多模态模型对视频进行智能分析。支持本地视频文件和远程 URL 两种输入方式, 可自定义分析提示词与抽帧频率(FPS),灵活控制分析精度与 API 调用成本. 核心能力涵盖场景描述、动作识别、物体检测、视频摘要生成、内容审核与问答式分析. 适用于内容创作辅助、视频内容索引、媒体资产...
- Publisher
- 天轰穿
- Fixed release
- 1.0.4
AI播客生成
基于MagicPodcast API将PDF文档、文本内容、笔记和网页链接转化为 自然流畅的双主持人对话式播客节目。支持多语言生成,返回可分享的 播客链接。适用于内容创作者、教育工作者、研究人员等需要将文字 内容转化为音频场景的用户。Use when 需要文本翻译、多语言转换、本地化处理时使用。不适用于专业医学法...
- Publisher
- 天轰穿
- Fixed release
- 1.0.3
DaDaScribe advanced speech-to-text transcription & translation
Transcribe audio and video with the DaDaScribe AI service (YouTube URLs, direct links, or local files). Supports 100+ languages, speaker diarization with named speakers, translation to up to 5 languages, and returns .txt transcripts plus .srt subtitles. Use whenever the user asks to transcribe, capt
- Publisher
- Fabrizio Ferrari
- Fixed release
- 1.0.2
audio-quality-check
Analyzes audio recording quality - echo detection, loudness, speech intelligibility, SNR, and spectral analysis. Use when the user wants to check a recording's quality, detect echo or duplication, measure speech clarity, compare original vs processed audio, or diagnose why a recording sounds bad, including tracks from Blackbox or any call recording app.
- Publisher
- Misha Kolesnik
- Fixed release
- 0.1.4
Instagram API
An Instagram API alternative on fetcher.sh — pay-per-call in USDC via x402, or prepaid credits with a Bearer key, no login and no session cookies. Use when the user wants to resolve an Instagram profile by @handle, search users by keyword, pull a profile's posts, reels, stories, tagged posts, followers, or followings, look up a single post by its shortcode, read a post's comments, fetch posts under a hashtag or reel-only hashtag feed, pull posts from a location, or pull posts using a specific audio/music track. Al…
- Publisher
- fetcher-sh
- Fixed release
- 0.1.0
播客
从播客音频或文字转写自动生成章节、亮点及节目说明,支持多语言和多格式输出,提升内容制作效率与质量。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
VideoLens
Manually turn user-selected videos into timestamped reports with a pinned VideoLens runtime and OpenAI BYOK. Use for summaries, tutorials, meetings, bugs, UX, privacy, and creator QA. Bootstrap clones GitHub code and installs dependencies only after explicit approval; analysis sends selected media-derived content to OpenAI only after separate credit approval.
- Publisher
- shadoprizm
- Fixed release
- 1.1.3
CC Video Creation
CC can create professional FactSage videos itself — free, no API, without asking Jarvis. Pipeline: story → animation (Manim/Ken Burns/Kling AI) → upload.
- Publisher
- Northcap Group
- Fixed release
- 1.0.1
確定申告アシスタント
フリーランス・個人事業主のための確定申告サポート。所得分類、経費整理、控除確認、消費税・インボイス制度の知識アシスタント。令和7年度税制改正(基礎控除95万円段階・160万円の壁)対応。API接続なし、マイナンバー取扱なし。
- Publisher
- LeoYann
- Fixed release
- 1.3.1
Short Video BGM Studio
Describe the footage and get an original instrumental track written for it, yours to keep and use commercially. This AI background music generator turns a scene, mood, and tempo feel into royalty-free BGM for short videos, vlogs, product clips, tutorials, livestream and store loops, podcast intros, and slideshow recaps, with an energy arc you choose — a calm or immediate opening, a lift at the moment that matters, a clean ending — room left for narration, and a result you can listen to before you publish.
- Publisher
- beatra-ai
- Fixed release
- 0.1.3
video-content-pipeline
Faceless video production for YouTube/TikTok/Shorts: Pixar-style images (free via Pollinations), parallax effects, voiceover (Edge TTS), composer — the full pipeline from script to finished video.
- Publisher
- Mohamed
- Fixed release
- 1.0.6
AI Song Cover Studio
Turn a song recording into a newly interpreted cover or rearranged song from a reference performance and a fresh genre, arrangement, and vocal direction. This AI song cover studio and reference-audio cover generator reinterprets a classic song as rock, acoustic, folk, jazz, Chinese-style, ballad, or a new vocal character, and reviews recognizable reference influence, arrangement freshness, vocal delivery, lyrics, pronunciation, and structure. Use it for classic-song reinterpretation, genre swaps, creator demos, tr…
- Publisher
- beatra-ai
- Fixed release
- 0.1.1
bilibili-video-storyboard
Create a Bilibili video storyboard from a topic, title, outline, or script. This AI Bilibili storyboard maker turns a long-form video idea into a chapter-led shot list, Bilibili video script, camera direction, narration and B-roll cues, and one to four storyboard key frames for explainers, reviews, tutorials, vlogs, gameplay, animation, and creative videos.
- Publisher
- beatra-ai
- Fixed release
- 0.1.2
AI Podcast Voiceover
Turn an article, notes, or a finished script into a listener-ready solo podcast episode with a consistent host voice. This AI podcast voice generator and AI podcast narration service adapts supplied material into a speakable podcast script, sets names and specialist terms for clear pronunciation, and creates MP3 podcast audio with natural pacing. Use this podcast voiceover AI and text-to-speech podcast service for article-to-podcast audio, news briefings, expert commentary, and knowledge shows, then carry the host…
- Publisher
- beatra-ai
- Fixed release
- 0.1.5
TikTok Shop Product Video Maker
Create TikTok Shop product-video plans from product facts, photos, selling points, and English or Japanese audience context. This AI product video maker produces hooks, a ready-to-film script, shot beats, subtitle cues, localized titles, hashtags, and a product-page-safe CTA for product showcases, demonstrations, unboxings, reviews, and creator-led shopping videos.
- Publisher
- beatra-ai
- Fixed release
- 0.1.3
AI Voiceover for Short Videos
Turn final short-video scripts into ready-to-edit voiceover audio for TikTok, Reels, YouTube Shorts, product reviews, hook lines, explainers, and ads. This AI voice over generator, AI voice reader, and text-to-speech voiceover workflow makes scripts speakable, helps choose a suitable voice and supported language, and tunes pacing, pauses, names, numbers, brands, and pronunciation. Review the current price estimate, create MP3 voiceover audio, compare the returned duration with the target edit, and place the narrat…
- Publisher
- beatra-ai
- Fixed release
- 0.1.8
speko
Use Speko to transcribe audio, synthesize speech, and pick the model for each leg of a voice pipeline from measured benchmarks instead of a hardcoded vendor. One key covers STT, LLM and TTS across 50+ models, with a dry-run route preview, per-language selection, price ceilings and automatic failover. Use when asked to transcribe a recording or voice note, read something aloud, choose or justify a speech model, work in a language other than English, or cap voice spend. For placing phone calls, use the speko-calls s…
- Publisher
- Speko
- Fixed release
- 1.0.4
speko-calls
Place and monitor outbound AI phone calls through Speko. This skill dials real telephone numbers and costs real money, so it needs its own platform credential and asks the user to confirm the number and the purpose before every call. Use when explicitly asked to call someone, ring a number, run an outbound voice campaign, or check the status, recording or transcript of a call that was already placed. For transcription, speech synthesis or model routing with no dialing involved, use the speko skill instead.
- Publisher
- Speko
- Fixed release
- 1.0.0
video-add-captions
Add word-timed captions to an Open Recut program. Use this skill to map the canonical transcript through timeline.json, review a maintained style on source-backed pixels, render a local transparent HyperFrames PNG sequence, and register it as an overlay contribution for the shared delivery render.
- Publisher
- WhiteTowerAI
- Fixed release
- 1.0.7
video-add-b-roll
Use when a talking-head, interview, documentary, or explanatory video needs deliberate transcript-timed visual cutaways from local media or Pexels.
- Publisher
- WhiteTowerAI
- Fixed release
- 1.0.3
scavio-youtube
Search YouTube and retrieve videos, shorts, comments, transcripts, streams, and channel data as structured JSON. 15 endpoints across video and channel surfaces.
- Publisher
- scavio-ai
- Fixed release
- 3.0.3
readworthy
Evaluate whether an article, document, video transcript, or webpage is worth reading; recommend full reading or specific sections; maintain a private local reading profile; and learn from explicit feedback. Use when a user shares content links, asks what is worth reading, corrects a prior assessment, requests rankings, or asks for cross-article insights.
- Publisher
- ZHAO PEICHEN
- Fixed release
- 1.0.1
video-analyzer
视频分析处理 — 本地视频反编译分析工具。将视频拆解为时间轴剧本、语音转文字、场景分析、跨模态关联和精华摘要,支持多ASR引擎切换(Whisper/Paraformer/SenseVoice)、中文NLP增强、PaddleOCR中文识别。v4.0 新增短视频平台适配(抖音/快手/B站/视频号)和自动剪辑建议(高光检测/冗余标记/EDL导出/字幕样式)。v4.1 新增tiny模型优先体验(75MB低门槛)、说话人分离质量评分、剪映draft.json导出。v4.2 新增场景管理(detect→slice一条链)、短视频爆款预测、实时直播分析(流式ASR+敏感词检测)。
- Publisher
- fyniujin
- Fixed release
- 4.2.0
声音克隆 可灵 Kling Audio Clone
Generate customized speech that highly restores the timbre by uploading reference audio using Kling Audio Clone. 使用可灵 (Kling) 声音克隆模型,通过上传参考音频,生成高度还原该音色的定制语音。
- Publisher
- dlazy
- Fixed release
- 1.3.8
语音合成 可灵 Kling TTS
Convert text into high-quality, emotional speech reading using Kling TTS. 使用可灵 (Kling) TTS 模型,将文本转化为高质量、情感丰富的语音朗读。
- Publisher
- dlazy
- Fixed release
- 1.3.8
音效生成 可灵 Kling SFX
Generate matching scene sound effects based on text descriptions or video frames using Kling SFX. 使用可灵 (Kling) 音效模型,根据文字描述或视频画面智能生成匹配的场景音效。
- Publisher
- dlazy
- Fixed release
- 1.3.8
video-gen
Generate a video from a story-telling narration script plus the author's photos, using Seedance 2.0 image-to-video via OpenRouter's asynchronous video API: pre-render gates (face-scan, narrative order, risk-POC), then per-clip submit -> poll -> download, then ffmpeg assembly. Multi-clip projects default to silent clips plus one continuous soundtrack (synthesized or royalty-free/PD); subtitle voiceover is the fallback (no OpenRouter TTS). Verified working 2026-08-15 (POC: 4s 480p clip, $0.28, ~3 min). Use when the…
- Publisher
- Jeff Yang
- Fixed release
- 1.0.0
story-telling
Generate a narration script (voiceover + shot list + per-clip visual prompts + music cues) for a personal travel-story video or essay, from a thought-flow master skill and the author's own photos. Use when the user wants "a script", "旁白脚本", a storytelling/故事化 script for a video, or asks to turn a personal travel story (e.g. the death-in-Mexico project) into a narrated video or essay. Output is Simplified Chinese first-person storytelling, engine-agnostic in voiceover and timing, engine-specific in clip prompts. Ne…
- Publisher
- Jeff Yang
- Fixed release
- 1.0.0
AI播客生成-免费版
将PDF、文本、链接转为双人对话播客,适合个人创作者快速制作音频内容。Use when 需要文件处理、文档转换、格式互转、内容提取时使用。不适用于加密文件破解。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。提供结构化输出和错误处理机制。
- Publisher
- 天轰穿
- Fixed release
- 1.0.2
yt-dlp-cli
yt-dlp. Use when the user wants to download, extract audio, pull subtitles, clip, or archive a YouTube or other streaming-site URL, or when yt-dlp hits 403, bot-check, or only-360p. Not for local-file edit/transcode, watch/summarize without saving, search, or generic downloads.
- Publisher
- Trevin
- Fixed release
- 0.1.2
腾讯会议录制纪要&待办
Use the complete official Tencent Meeting record command family to locate recordings, retrieve native smart-minutes summaries and Todos, verify evidence in transcripts, handle recording-access requests, and produce a source-labeled meeting Todo ledger.
- Publisher
- TmeetingSkill
- Fixed release
- 1.0.1
Azure语音交互免费版
使用Azure VoiceLive构建基础实时语音AI应用,支持文本/音频输出与基本会话管理。Use when 需要视频处理、音频编辑、媒体转换、配音生成时使用。不适用于版权受保护的媒体内容处理。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。
- Publisher
- 天轰穿
- Fixed release
- 1.0.2
Azure语音转写免费版
使用Azure AI进行批量语音转文字,支持基础转写与时间戳,适合个人用户处理音频。Use when 需要提升效率、自动化流程、批量处理、工作流优化时使用。不适用于需要人工创意判断的任务。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。
- Publisher
- 天轰穿
- Fixed release
- 1.0.2
音频流上传免费版
快速上传音频至流媒体平台,支持基础创建、上传与完成三步流程,获取HLS流媒体链接。Use when 需要视频处理、音频编辑、媒体转换、配音生成时使用。不适用于版权受保护的媒体内容处理。适用于独立开发者、企业团队和自动化工作流场景。支持中文交互,无需复杂配置即开即用。输出结果可直接使用,减少二次加工成本。
- Publisher
- 天轰穿
- Fixed release
- 1.0.2
Seedance 2.5 参考音频生视频
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance 2.5 参考音频生视频”:提供音乐或声音即可围绕节奏、情绪和声音线索生成画面;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 2.5 参考音频生视频, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway、Pika、Sora、Veo、剪映 CapCut、HeyGen 等 AI 视频工具的替代方案、同类能力、价格、API、…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 参考视频生视频
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance 参考视频生视频”:提供参考视频即可复用动作节奏、镜头运动或叙事结构;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 参考视频生视频, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway、Pika、Sora、Veo、剪映 CapCut、HeyGen 等 AI 视频工具的替代方案、同类能力、价格、API、国内可用入口或工作…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 参考音频生视频
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance 参考音频生视频”:提供音乐或声音即可围绕节奏、情绪和声音线索生成画面;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 参考音频生视频, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway、Pika、Sora、Veo、剪映 CapCut、HeyGen 等 AI 视频工具的替代方案、同类能力、价格、API、国内可用入口或工…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 参考图生视频
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance 参考图生视频”:提供参考图即可延续人物、产品、构图或视觉风格;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 参考图生视频, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway、Pika、Sora、Veo、剪映 CapCut、HeyGen 等 AI 视频工具的替代方案、同类能力、价格、API、国内可用入口或工作流迁移,…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 参考生视频
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance 参考生视频”:可用参考图、参考视频或参考音频约束人物、产品、动作、运镜与风格;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 参考生视频, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway、Pika、Sora、Veo、剪映 CapCut、HeyGen 等 AI 视频工具的替代方案、同类能力、价格、API、国内可用入口…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance2.5 最低 8 折|比官方更便宜的视频生成渠道
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Seedance2.5 最低 8 折|比官方更便宜的视频生成渠道”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance2.5 最低 8 折|比官方更便宜的视频生成渠道, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance2.0 最低 8 折|比官方更便宜的视频生成渠道
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Seedance2.0 最低 8 折|比官方更便宜的视频生成渠道”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance2.0 最低 8 折|比官方更便宜的视频生成渠道, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hail…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 最低 8 折|比官方更便宜的视频生成渠道
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Seedance 最低 8 折|比官方更便宜的视频生成渠道”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 最低 8 折|比官方更便宜的视频生成渠道, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
HappyHorse 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“HappyHorse 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for HappyHorse 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Wan2.5 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Wan2.5 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Wan2.5 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、Pi…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Wan3.0 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Wan3.0 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Wan3.0 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、Pi…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Minimax H3 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Minimax H3 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Minimax H3 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Run…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance2.5 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance2.5 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance2.5 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreami…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance1.5 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance1.5 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance1.5 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MO…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance2.0 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance2.0 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance2.0 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、R…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“Seedance 视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Happy Horse 视频生成
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Happy Horse 视频生成”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Happy Horse 视频生成, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
视频生成与编辑
帮助视频剪辑师、后期团队、广告制作人与需要复用现有素材的创作者直接完成“视频生成与编辑”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for 视频生成与编辑, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
MiniMax H3 视频生成
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“MiniMax H3 视频生成”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for MiniMax H3 视频生成, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway、Pik…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 2.0 视频生成
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Seedance 2.0 视频生成”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 2.0 视频生成, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、美图MOKI、Vidu、PixVerse、Runway…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 2.5 视频生成
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Seedance 2.5 视频生成”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 2.5 视频生成, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
Seedance 视频生成
帮助内容创作者、广告与营销团队、电商团队、短剧和漫剧制作团队直接完成“Seedance 视频生成”:既可只输入文字从零生成,也可加入图片、视频或音频控制结果;通过 AI Hive 使用时,生成前自动上传所需素材,提交后自动保存任务、查询进度并下载成片。适用于 AI 广告、TVC、电商视频、产品展示、带货、种草、短剧、漫剧和社媒内容。 Use this skill for Seedance 视频生成, text-to-video, image-to-video, first-and-last-frame animation, reference-to-video, video-to-video, audio-to-video, AI video editing, video extension, product videos, e-commerce ads, TVC, social commerce, seeding content, short drama, comic drama, and AIGC video production. 如果用户正在比较或寻找 可灵 Kling、即梦 Dreamina、海螺 Hailuo、…
- Publisher
- Bain Wu
- Fixed release
- 1.0.0
content-orchestrator
内容生成+发布统一编排器,15条管道文件(12 PL-*+3 E2E-*):VIDEO/VIDEO-BATCH/IMAGE/AUDIO/LIPSYNC/COMIC/COMIC-BATCH/ARTICLE-BATCH/NOVEL-BATCH/PRODUCT/HOTSPOT/NEWPROD+E2E-VIDEO/E2E-IMAGE/E2E-DAILY+3内置虚拟路由(PL-NOVEL连载/PL-DRAMA短剧/PL-UPLOAD上传)。平台注册表自动路由+代理自动注入+多租户感知(tenant_id→风格/人设/素材/平台隔离)+素材→闲鱼商品(PL-PRODUCT)+热点→商品(PL-HOTSPOT)+小说连载(PL-NOVEL/PL-NOVEL-BATCH)+短剧生成(PL-DRAMA内置)。触发:生成内容/发布内容/一条龙/日常运营/素材转商品/热点选品/热点上架/小说连载/短剧生成/上传内容生成 不触发:纯闲鱼运营/纯客服回复/数据分析查询
- Publisher
- 天轰穿
- Fixed release
- 1.0.0
Jobs-System
市面上有一千个会背语录的乔布斯。 我们把"他怎么想到那句话"这件事,做成了可复现的工程。 你问它"这生意该不该做",它不给你十页报告,先甩你一句: "这事的根不在市场,在你把相机当硬件卖。" 我们要复刻的,就是这一句怎么来的。 乔布斯式决策分析工作流(社区演示版)。当用户需要分析赛道/创业机会、做产品定义、战略取舍、砍需求做减法、竞品定位、用户洞察、技术窗口判断,或用「乔布斯视角/连点成线」拆解问题时启用。三层推理框架(理解意图+拷问+答案+Voice 蒸馏)产出带证据链、确定性标记与域类型路由的结构化决策;B 端/I 层危险域标⚠️翻车不模仿其判断,AI/云/社媒等未及域标「框架推演·非本人立场」。非触发(闲聊、纯信息检索、纯代码实现)不加载。
- Publisher
- Sabre
- Fixed release
- 1.0.0
播客下载器
从小宇宙下载播客音频与节目说明。从小宇宙(xiaoyuzhoufm。com)下载播客音频和Show Notes。自动转换为MP3格式(兼容Sanag、小游等骨传导蓝牙耳机、水下游泳时离线播放)。支持多种输入格式,输出结构化结果,适用于独立开发者与一人公司效率提升。Use。Use when 需要代码生成、编程辅助、调试测试、开发部署时使用。不适用于无明确技术栈的模糊需求。 when 需要代码生成、编程辅助、调试测试、开发部署时使用。不适用于无技术栈的通用场景。
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
Piper TTS
用Piper本地文字转语音发语音消息。Local text-to-speech using Piper for voice message delivery。Use when。触发关键词: using, local, speech, text, piper, tts, beware。开箱即用,无需复杂配置,支持中文交互与结构化输出。
- Publisher
- 天轰穿
- Fixed release
- 1.0.4
Video Object Remover
Remove an unwanted person, object, logo, or distraction from a video with Video Object Remover.
- Publisher
- luffy
- Fixed release
- 0.1.0
bilibili-all-in-one
A comprehensive Bilibili toolkit that integrates hot trending monitoring, video downloading, video watching/playback, subtitle downloading, and video publishing capabilities into a single unified skill. Supports Bilibili session cookie authentication for publishing and high-quality downloads. Requests go to official Bilibili API endpoints over HTTPS.
- Publisher
- zuoyunlai
- Fixed release
- 1.0.0
language-immersion-tv
Turn movies and TV shows into language-learning material by analyzing subtitle files to extract vocabulary, build frequency decks, and create contextual flashcards. Use when learning a language through media immersion.
- Publisher
- voronindenis5
- Fixed release
- 1.0.2
Bilibili Video Parser
把B站视频链接一键转成中文字幕连贯稿 + 结构化 HTML 分析报告。优先使用 CC 字幕秒出结果,无字幕时本地 faster-whisper 转写,不需要任何 API key。
- Publisher
- Black_Amico
- Fixed release
- 0.1.4
reddit-lead-inbox
Monitors Reddit 24/7 for buyer-intent keywords, filters leads by AI intent scoring, drafts founder-voice replies, enables manual approval and attribution tra...
- Publisher
- heroinyan-stack
- Fixed release
- 1.0.0
youtube-transcript
Fetch and save YouTube video transcripts as clean plain text. Use when the user provides a YouTube URL or wants to extract a transcript from a podcast, interview, or talk.
- Publisher
- Yuval Adam
- Fixed release
- 1.0.0
bilibili-all-in-one
面向B站的六合一全功能工具技能,集成热门监控(Hot Monitor)、视频下载(Downloader)、 数据追踪(Watcher)、字幕处理(Subtitle)、播放信息(Player)与视频投稿(Publisher)六大模块. 支持热门/热搜/必看榜/分区排行实时获取,360p至4K多清晰度下载与mp4/flv/mp3格式转换, 播放量/点赞/评论长期追踪与多视频对比,字幕下载与格式转换,弹幕获取与播放列表解析, 以及视频上传/定时发布/草稿编辑。凭据支持环境变量、JSON文件、直接参数三种方式, 默认内存存储,可选 `B...
- Publisher
- 天轰穿
- Fixed release
- 1.0.3
Graincrawl
Search local Granola notes and transcripts
- Publisher
- OpenClaw
- Fixed release
- 1.0.1
Ai Video Director
AI视频导演是一款从脚本到成片一站式视频制作工具。一个人就是一支视频团队,支持热点短视频生成、 口型同步数字人、智能路由决策树、多引擎降级。营销策略自动注入,三层TTS降级链保证配音可用性。 核心能力: - 智能路由决策树:5步自动选择最优视频引擎,支持角色一致性/画质/风格化/通用/口型同步多引擎 - 口型同步...
- Publisher
- 天轰穿
- Fixed release
- 1.0.1
Article to Podcast
Use when the user wants to turn an article/blog/text into a published podcast episode: triggers like "生成播客""做一期播客""博客转播客""文章转播客""发布/更新播客""blog to podcast", o...
- Publisher
- Wu Haohua
- Fixed release
- 1.0.0
Openai Whisper
Local speech-to-text with the Whisper CLI (no API key).
- Publisher
- Sean Ford
- Fixed release
- 1.0.0
Join meeting
AgentCall (agentcall.dev) — Join a video meeting (Google Meet, Teams, Zoom) as an AI bot with voice and visual presence. Supports audio-only mode with voice intelligence (barge-in, interruptions), text-to-speech mode, and webpage modes for custom UI. Use when asked to join a call, attend a meeting,
- Publisher
- johnpatternai
- Fixed release
- 1.1.15
Hyperframes
Create video compositions, animations, title cards, overlays, captions, voiceovers, audio-reactive visuals, and scene transitions in HyperFrames HTML. Use wh...
- Publisher
- LucasL
- Fixed release
- 1.0.0
Openai Whisper Api
Transcribe audio via OpenAI Audio Transcriptions API (Whisper).
- Publisher
- Peter Steinberger
- Fixed release
- 1.0.0
Session Archiver
Session auto-archiver. Automatically archives completed sessions by extracting meaningful user+assistant messages from .reset transcript files and appending...
- Publisher
- sniper-one
- Fixed release
- 2.1.1
Video Generator | 视频生成器
Automated text-to-video pipeline with multi-provider TTS/ASR support - OpenAI, Azure, Aliyun, Tencent | 多厂商 TTS/ASR 支持的自动化文本转视频系统
- Publisher
- Justin Liu
- Fixed release
- 1.0.42
video-analyzer
鏅鸿兘鍒嗘瀽 Bilibili/YouTube/鏈湴瑙嗛锛岀敓鎴愯浆鍐欍€佽瘎浼板拰鎬荤粨銆傛敮鎸佸叧閿抚鎴浘鑷姩宓屽叆銆?
- Publisher
- tamakooooo
- Fixed release
- 1.0.11
小米 MiMo TTS
Text-to-speech using Xiaomi MiMo TTS API. Generates WAV audio files. Triggers when user says "send voice message", "voice reply", "read to me", "use clip voi...
- Publisher
- whmmy
- Fixed release
- 1.0.2
Piper TTS
Local text-to-speech using Piper for voice message delivery. Use when the user asks for voice responses, audio messages, TTS, text-to-speech, voice notes, or...
- Publisher
- bewareofddog
- Fixed release
- 1.0.1
YouTube Transcript
Fetch and summarize YouTube video transcripts. Use when asked to summarize, transcribe, or extract content from YouTube videos. Handles transcript fetching via residential IP proxy to bypass YouTube's cloud IP blocks.
- Publisher
- The Zealot
- Fixed release
- 1.0.1
video by remotion
Automated video production studio using Remotion + React + TTS. Creates animated explainer videos from JSON content scripts through a make-driven pipeline: T...
- Publisher
- hzsunzixiang
- Fixed release
- 1.0.1
Sherpa ONNX TTS
Local text-to-speech via sherpa-onnx (offline, no cloud)
- Publisher
- Daniel Sinewe
- Fixed release
- 0.1.0
analyze video by qwen
使用 Qwen 3.5 Plus 模型分析视频内容,支持本地文件和远程 URL,可自定义分析提示词和抽帧频率
- Publisher
- www
- Fixed release
- 1.0.1
Video
该技能介绍如何通过视频剪辑实现变现;当你计划从事或优化视频剪辑时调用。
- Publisher
- mike47512
- Fixed release
- 0.1.0
Bilibili All In One
A comprehensive Bilibili toolkit that integrates hot trending monitoring, video downloading, video watching/playback, subtitle downloading, and video publish...
- Publisher
- enoyao
- Fixed release
- 1.0.24
SenseAudio 会议助手 / SenseAudio Meeting Assistant
用于构建和排查 SenseAudio 会议助手,覆盖实时会议转写、说话人区分、实时翻译、会议纪要生成、行动项提取与转录导出。Build and troubleshoot SenseAudio meeting assistants for live meeting transcription, speaker-aw...
- Publisher
- scikkk
- Fixed release
- 1.0.0
tiktok-scraper
Access live TikTok data via the CreatorCrawl API. Look up creator profiles, search videos, pull comments and transcripts, track trending hashtags, discover p...
- Publisher
- Simon Balfe
- Fixed release
- 1.0.0
whisperx
WhisperX provides local speech-to-text transcription using OpenAI Whisper, with high-quality offline recognition, no API key required, word-level timestamps,...
- Publisher
- niuzb
- Fixed release
- 1.0.0
Podcast Downloader
小宇宙播客下载工具。从小宇宙(xiaoyuzhoufm.com)下载播客音频和Show Notes。自动转换为MP3格式(兼容Sanag、小游等骨传导蓝牙耳机、水下游泳时离线播放)。当用户需要下载播客、保存播客音频、提取播客文字内容时使用。支持:(1) 单集下载,(2) 批量下载,(3) 自定义音质,(4) 自动...
- Publisher
- zhang Yuming
- Fixed release
- 1.0.0
Seedance
Generate detailed, production-ready cinematic video prompts following Seedance 2.0’s strict Subject-Action-Camera-Style-Audio-Constraints format for AI video...
- Publisher
- honeybee1130
- Fixed release
- 1.0.0
Openai Whisper 1.0.0
Local speech-to-text with the Whisper CLI (no API key).
- Publisher
- Patryk Czubiński
- Fixed release
- 1.0.0
Podcast Generation from PDF, Text, and Links
Generate AI podcast episodes from PDFs, text, notes, and links using MagicPodcast in OpenClaw. Creates natural two-person dialogue audio, supports custom lan...
- Publisher
- mogens9
- Fixed release
- 1.0.11
MarkItDown Skill
OpenClaw agent skill for converting documents to Markdown. Documentation and utilities for Microsoft's MarkItDown library. Supports PDF, Word, PowerPoint, Excel, images (OCR), audio (transcription), HTML, YouTube.
- Publisher
- karmanverma
- Fixed release
- 1.0.1
Podcast
Create and grow podcasts by planning episodes, producing audio or video, generating clips, and building audience across formats.
- Publisher
- Iván
- Fixed release
- 1.0.1
Podcast Chaptering Highlights
Create chapters, highlights, and show notes from podcast audio or transcripts. Use when a user wants chapter markers, highlight clips, or show-note drafts without publishing or distribution actions.
- Publisher
- codedao12
- Fixed release
- 1.0.0
Qwen3-tts
Local text-to-speech using Qwen3-TTS-12Hz-1.7B-CustomVoice. Use when generating audio from text, creating voice messages, or when TTS is requested. Supports 10 languages including Italian, 9 premium speaker voices, and instruction-based voice control (emotion, tone, style). Alternative to cloud-based TTS services like ElevenLabs. Runs entirely offline after initial model download.
- Publisher
- paki81
- Fixed release
- 1.0.0
Openai Whisper
Local speech-to-text with the Whisper CLI (no API key).
- Publisher
- Peter Steinberger
- Fixed release
- 1.0.0
Sag
ElevenLabs text-to-speech with mac-style say UX.
- Publisher
- Peter Steinberger
- Fixed release
- 1.0.0
Video Frames
Extract frames or short clips from videos using ffmpeg.
- Publisher
- Peter Steinberger
- Fixed release
- 1.0.0