Documents skills

Browse reusable Agent Skills, each with a clear purpose and practical guidance.

byted-las-long-video-understand

Performs deep AI-powered analysis and understanding of long-form videos (up to 3 hours, 10GB) using Volcengine LAS large language models. Video analysis, video comprehension, and video summarization — generates comprehensive video summaries, video recaps, chapter breakdowns, event timelines, key moments, and structured content indexing. Supports behavior detection, action detection, video annotation, video tagging, and video content recognition. Enables intelligent video question answering — ask questions about video content and get AI answers. Handles long meeting recordings, lecture videos, webinars, surveillance and security footage, movies, tutorials, and any long video that needs detailed understanding. Async processing with submit-poll workflow. Use this skill when the user wants to analyze or understand long videos (up to 3h/10GB) with LLM-based deep comprehension, summarize video content or generate recaps, extract chapters/key moments/timelines from videos, do video Q&A (ask questions about video content), detect actions/behaviors in videos, review meeting recordings/lectures/webinars, or get structured video descriptions.

371 repo starsObserved in 1 repos
Documents

byted-las-pdf-parse-doubao

PDF 解析(Doubao):将 PDF/扫描件转成结构化 Markdown/文本,支持表格与多栏版式。当用户要从 PDF 提取文字/表格或做 OCR 时触发。

371 repo starsObserved in 1 repos
Documents

byted-las-video-edit

智能视频剪辑:用自然语言描述要保留的片段/人物/事件,自动从长视频抽取剪辑(可用参考图做人物/物体匹配)。当用户要按描述找片段、提取高光或做智能剪辑时触发。

371 repo starsObserved in 1 repos
Documents

byted-las-video-inpaint

Removes unwanted visual elements from videos using AI-powered inpainting via Volcengine LAS. Video watermark removal, subtitle removal, logo removal, and text overlay removal — erases and cleans up watermarks, hardcoded subtitles, logos, text, and other fixed-region artifacts and obstructions from video footage. Video inpainting, video repair, video restoration, and video cleanup. Supports specifying exact bounding boxes for targeted removal and erasure. Handles videos up to 4 hours and 30GB. Use this skill when the user wants to remove watermarks from videos, erase hardcoded subtitles/captions, delete logo or text overlays, clean up video footage by removing unwanted objects, repair or restore videos by removing obstructions, do video inpainting on fixed regions, or remove visual artifacts from video content.

371 repo starsObserved in 1 repos
Documents

byted-las-video-resize

Resizes, scales, and adjusts video resolution and dimensions using GPU-accelerated NVENC encoding via Volcengine LAS. Video resizing, video scaling, video upscaling, and video downscaling — change video resolution, enlarge or shrink video dimensions, and adjust video size for different platforms and screens. Video compression by reducing resolution and video transcoding and re-encoding with specific dimension constraints. Supports flexible min/max width and height ranges with aspect ratio preservation strategies (increase, decrease, or disable), including landscape to portrait conversion. Async submit-poll workflow with batch support. Use this skill when the user wants to resize or scale video resolution (upscale/downscale), change video dimensions for different platforms, compress videos by reducing resolution, transcode/re-encode videos with GPU NVENC, adjust aspect ratio including landscape-to-portrait conversion, adapt videos for mobile/web/social media, or batch process multiple videos.

371 repo starsObserved in 1 repos
Documents

byted-las-vlm-video

视频多模态理解:对视频生成描述/摘要/标签,并支持基于视频内容问答。当用户要理解视频内容、生成描述或对视频提问时触发。

371 repo starsObserved in 1 repos
Documents

byted-mediakit-process-tools

火山引擎 AI MediaKit 音视频处理工具集,提供视频理解、音频提取、视频剪辑、音视频拼接、画质增强、文生视频、音视频合成,以及本地翻转、调速、字幕、水印、转码等能力。当用户提及音频剪辑、视频剪辑、音视频拼接、文生视频、音频提取、画质增强、视频理解、音视频合成、媒体裁剪、视频翻转、调速、加字幕、加水印、转码等需求时必须调用本Skill。当用户需要视频理解时,宿主agent必须自动解析用户的具体要求作为prompt参数传入,同时传入视频URL和fps参数;max_frames 为可选参数。

371 repo starsObserved in 1 repos
Documents

byted-mediakit-voiceover-editing

Volcano Engine AI MediaKit talking-head video editing Skill: a one-stop workflow from environment setup through media management, audio processing, talking-head cuts, video export, review UI, and iterative refinement. You MUST invoke this Skill when the user mentions talking-head editing, cutting talking video, video editing, removing pauses, processing audio, exporting talking video, automatic editing, removing verbal slips, or similar. Also invoke when the user uploads video or audio and asks for editing.

371 repo starsObserved in 1 repos
Documents

byted-music-generate

Generate music using Volcengine Imagination API. Supports vocal songs, instrumental BGM, and lyrics generation. Use when the user wants to create songs, background music, soundtracks, write lyrics, or mentions "music generation", "BGM", or "songwriting".

371 repo starsObserved in 1 repos
Documents

byted-seedance-video-generate

Generate videos using Seedance models. Invoke when user wants to create videos from text prompts, images, or reference materials.

371 repo starsObserved in 1 repos
Documents