Extract audio from short videos (Douyin/TikTok) and transcribe to text with timestamps. Use when user provides video URL and needs audio transcription.
Alibaba Cloud Bailian Video Analysis Skill. Use for intelligent video comprehension and analysis via the Bailian (QuanMiaoLightApp) API.
**Required API Product**: QuanMiaoLightApp (version 2024-08-01)
**Required API Actions**: SubmitVideoAnalysisTask, GetVideoAnalysisTask
**DO NOT use**: videorecog, Mts, or any other product for video analysis
Triggers: "analyze video", "understand video", "analyze the local video /temp/xxx.mp4", "analyze the local video https://xxx.com/temp/xxx.mp4", "what is this video about", "summarize this video", "split video into shots", "video comprehension", "extract video insights", "transcribe video", "extract video captions", "generate video title", "generate video outline", "video mindmap".
This skill should be used when the user asks to "build a frame extraction job" / "视频抽帧 / 抽关键帧", "label driving images with a VLM" / "图像打标 / image labeling with Qwen-VL", "compute image embeddings" / "图像向量化 / multi-modal embedding", "build a video_table / image_table / clip_dir_table for AI FUNC", "扫 OSS 建 video meta 表", or mentions driving-scene / ADAS / 智驾 / 智能驾驶 / 自动驾驶 / 路测 / 行车记录仪 / 座舱 video or image pipelines on MaxFrame + OSS + ODPS. Not for audio (use driving-audio-maxframe-job).
Process images, audio, and video files stored in Alibaba Cloud OSS.
Supports 14+ image operations (resize, crop, rotate, watermark, blur, format conversion, etc.),
image-intelligent features via IMM (blind watermark, face/body/car detection, QR recognition, labeling, scoring),
and audio/video processing (transcoding, screenshot, animation, sprite sheet, concatenation, metadata extraction, HLS streaming).
Results can be returned as signed URL, downloaded locally, or saved as new OSS object.
Also supports plain file upload/download.
Use when the user needs to process or transform media files in OSS, such as generating thumbnails, transcoding video,
extracting audio, adding watermarks, detecting faces, compressing images, or converting formats.
Triggers on media processing requests in English or Chinese
(resize, crop, thumbnail, transcode, video convert, audio convert, watermark,
face detection, 缩略图, 裁剪, 压缩, 转码, 视频转换, 音频处理, 水印, 盲水印, 人脸检测, 截帧, 拼接).
Quick BI-SmartQ skill with multiple data analysis capabilities:
1. **File Q&A**: Upload Excel/CSV files for intelligent analysis via Quick BI API
2. **Dataset Q&A**: Natural language queries on Quick BI platform datasets, with automatic intelligent table selection and matching
3. **Document Parsing**: Parse PDF/Word/Excel/CSV/images, extract text, and support extracting key fields to generate structured Excel
4. **Dashboard Skill Generation**: Auto-convert QuickBI dashboards into data query skills
5. **Data Insight**: Deep data insight analysis on Quick BI datasets
6. **Data Report**: Auto-generate professional data reports based on analysis results
Use when users mention data analysis, smart Q&A, querying data, file analysis,
document parsing, dashboard skills, data insight, or data reports.
Alibaba Cloud Media Processing Service (MPS) one-stop video processing skill. Use when users need video processing, transcoding, snapshot generation, content moderation, or video upload. For video distribution scenarios, complete video upload, snapshot, multi-resolution transcoding, and content moderation in a single workflow for efficient standardized video asset production.