Short drama script generation | | - The short drama/movie script generation operator is an automated script reverse engineering tool designed for short dramas as well as long-form or serialized video content such as movies. The operator leverages a visual multimodal large model (VLM) to automatically extract all characters from the entire drama or film, analyze character relationships, and generate high-quality textual scripts and character lists with scene, action, expression, and dialogue details based on visuals and dialogue, supporting secondary creation, overseas translation, and copyright protection for video content.
- Key features:
- Consistent character identification: Overcomes the limitations of isolated understanding in single episodes or segments, enabling stable tracking of main characters in long-running series or feature-length films. Ensures a high degree of consistency in character identity and settings across episodes, even in cases of costume changes, side profiles, or complex scene transitions, ultimately building a complete global character list.
- High-fidelity script reconstruction: Combines video visuals and dialogue to generate professional-grade storyboard scripts (with precise timestamps in movie mode). Accurately restores scene layout, character emotions, body movements, and key dialogue, providing high-quality textual drafts ready for secondary development or for translation or proofreading.
- Dual-mode adaptive architecture:
- Short drama mode: Supports batch input of multiple short drama episodes, processes them strictly in input order, and maintains continuity of the serialized plot and character consistency.
- Movie mode: For movies or long recordings lasting several hours per episode, automatically initiates adaptive processing strategies for long videos, effectively alleviating the issue of detail loss caused by large models with long context windows.
- Flexible output format customization: Provides an open custom instruction (Prompt) interface. You can freely adjust the script generation style for each episode according to specific business requirements (such as focusing on psychological descriptions, specific storyboard layout formats, specific text markers, and more), meeting the direct integration needs of various downstream businesses.
- Convenient result delivery: Supports securely writing the generated character list and full script directly to your designated cloud storage (TOS), and can also generate packaged pre-signed download links.
|
Short drama script parsing | | - The short drama script parsing operator automatically converts short drama script text into a structured asset table. The operator reads one or more script files, automatically identifies the script format and output language, extracts three types of assets—characters, props, and scenes—and adds visual details to character appearance, character styling, prop visual descriptions, and scene spatial descriptions. The final results are saved as JSON files.
- Key features:
- Multi-script input parsing: Supports input of one or more script files, concatenates and processes them in input order, suitable for parsing single episodes, multiple episodes, or entire series scripts.
- Character asset table extraction: Identifies character names, gender, basic appearance, and the scenes in which the character appears, and provides multiple styling descriptions for main characters that match the script background.
- Prop asset table extraction: Identifies key props, their locations, their role in the plot, and visual features to assist prop design and asset management.
- Scene asset table extraction: Identifies main scenes, space types, time setting, atmosphere, and visual layout to assist scene construction and storyboard generation.
- Chinese and English script adaptation: Supports Chinese short drama numbering format and English Hollywood/Fountain slug-line format; in
auto mode, automatically identifies the script format and determines the asset table output language. - Stable processing of ultra-long scripts: Automatically segments scripts exceeding the length threshold along scene boundaries to reduce timeout and truncation risks caused by ultra-long context requests.
- TOS deliverables: The character table, prop table, scene table, and script background information are all saved as JSON files, and the corresponding TOS paths are returned.
|
Short drama script fission | | - Designed to perform targeted rewriting based on existing episodic scripts. The operator preserves the original script's core character structure, conflict chain, emotional arc, and key pacing points, while systematically rewriting character identities, world settings, regional context, industry background, and key terminology to generate a new script version tailored to the target market and subject matter. The output includes the directory of the newly split episodic scripts, the complete script, outline, and before-and-after mapping table, making it suitable for short drama overseas distribution, replicating hit shows, and testing content directions.
- Key features:
- Targeted rewriting with skeleton retention: Preserves the original script's character functions, character relationship structure, sequence of key events, and main storyline logic, while only rewriting surface settings and the context of expression.
- Localization rewriting: Supports generating scripts for different regional contexts, including simultaneous adaptation of character naming styles, institutional backgrounds, dialogue expressions, and social relationship logic.
- Subject-based rewriting: Supports overall style migration according to the target subject matter, rewriting the same plot skeleton into new scripts reusable across different content genres.
- Batch processing of episodic scripts: Supports submitting multiple episodic scripts at once via directory or list, processes them in input order, and maintains narrative coherence across the complete script.
- Structured delivery of results: Outputs the split episodic scripts, complete script, outline, and mapping table after transformation, facilitating subsequent video production, proofreading, manual refinement, and effectiveness review.
|
Video script planning | | - Enter the video topic and creative requirements to generate a video production plan for subsequent production. The plan breaks down the creative idea into a clear content structure, shot arrangement, material suggestions, and timing and pacing, making it suitable for early-stage planning of educational and marketing videos.
- Key features:
- From creative idea to production plan: Organizes the topic, audience, and communication objectives into an actionable video production plan.
- Shot and pacing planning: Break down video content and plan the shot direction, duration, and transitions for each section.
- Material usage suggestions: Arrange the use of video, images, and other materials according to content requirements to facilitate subsequent production.
- Support for different content types: Supports planning for course instruction and marketing or e-commerce videos.
|
Video scene segmentation | | - The video scene segmentation operator uses a multimodal large model to segment input videos into shots/scenes, perform global character recognition, associate characters at the scene level, and extract character images. The operator outputs scene summary results, a character registry, video clips for each scene, and image files organized by character for subsequent retrieval, editing, and content understanding.
- Key features:
- Supports scene segmentation based on VLM, as well as equal-duration segmentation when
min_segment_duration == max_segment_duration. - Supports global character extraction and deduplication aggregation to generate a character registry.
- Supports associating characters within scenes, outputting the time intervals of character appearances, key frame timestamps, and bbox information in the scene.
- Supports automatically extracting independent video files for each scene.
- Supports extracting and filtering representative images for each character, and outputs them, organized by character.
- Supports outputting token usage and LLM request counts to facilitate cost assessment.
|