音视频技术员

Audio And Video Technicians

音视频技术员(Audio And Video Technicians)相关 AI Agent 技能。支持 Claude Code / Cursor 一键安装。

⭐ 该职业下的热门技能

transcribe.md ★ 20.1k
from "openai"
Transcribe audio files to text with optional diarization and known-speaker hints. Use when a user asks to transcribe speech from audio/video, extract text from
查看详情 →
clip-hand-skill.md ★ 17.6k
from "RightNow-AI"
Expert knowledge for AI video clipping — yt-dlp downloading, whisper transcription, SRT generation, and ffmpeg processing
查看详情 →
videocaptioner.md ★ 14.7k
from "WEIFENG2333"
Process video subtitles — transcribe speech, optimize/translate text, burn styled subtitles into video. Use when you need to add subtitles to a video, transcrib
查看详情 →
acestep-simplemv.md ★ 10.6k
from "ace-step"
Render music videos from audio files and lyrics using Remotion. Accepts audio + LRC/JSON lyrics + title to produce MP4 videos with waveform visualization and sy
查看详情 →
obsidian-publish-all.md ★ 1.8k
from "obsidianmd"
Publish all active obsidian-help locales (en, ja, pt-br, es, zh, fr) via ob publish. Use this skill whenever the user wants to publish or deploy the help site,
查看详情 →
spotify.md ★ 602
from "sundial-org"
Control Spotify playback on macOS. Play/pause, skip tracks, control volume, play artists/albums/playlists. Use when a user asks to play music, control Spotify,
查看详情 →
gif-generation.md ★ 289
from "athola"
Post-process video files and generate optimized GIFs. Converts webm/mp4.
查看详情 →
screenshots.md ★ 196
from "kevinpbuckley"
Capture screenshots of the editor window, viewports, and blueprints for AI vision analysis
查看详情 →
media-processing.md ★ 193
from "nicepkg"
Video/audio/image processing with FFmpeg and ImageMagick. Tools: FFmpeg (video/audio), ImageMagick (images). Capabilities: format conversion, encoding (H.264/H.
查看详情 →
c-video.md ★ 141
from "daxaur"
Download videos, extract audio, convert formats, and clip segments using `yt-dlp` and `ffmpeg`. Supports YouTube, Vimeo, and hundreds of other sites.
查看详情 →
image-to-video.md ★ 135
from "NeverSight"
Still-to-video conversion guide: model selection, motion prompting, and camera movement. Covers Wan 2.5 i2v, Seedance, Fabric, Grok Video with when to use each.
查看详情 →
talking-head-production.md ★ 135
from "NeverSight"
Talking head video production with AI avatars, lipsync, and voiceover. Covers portrait requirements, audio quality, OmniHuman, PixVerse lipsync, Dia TTS. Use fo
查看详情 →
background-removal.md ★ 135
from "NeverSight"
Remove backgrounds from images with BiRefNet via inference.sh CLI. Model: BiRefNet (high accuracy background removal). Use for: product photos, portraits, e-com
查看详情 →
talking-head-production.md ★ 135
from "NeverSight"
Talking head video production with AI avatars, lipsync, and voiceover. Covers portrait requirements, audio quality, OmniHuman, PixVerse lipsync, Dia TTS. Use fo
查看详情 →
image-upscale.md ★ 134
from "NeverSight"
Upscales an image using AI super-resolution to increase resolution with detail generation. Use when you need to enlarge images, improve low-resolution photos, o
查看详情 →
© 2026 skillsABC.com — 14.8万+ AI Agent 技能集市
职业分类 数据洞察 提交技能 隐私声明