{"id":55633,"date":"2026-08-09T10:02:21","date_gmt":"2026-08-09T02:02:21","guid":{"rendered":"https:\/\/www.1ai.net\/?p=55633"},"modified":"2026-08-09T10:02:21","modified_gmt":"2026-08-09T02:02:21","slug":"%e6%9c%ac%e5%9c%b0%e9%83%a8%e7%bd%b2%e6%9c%80%e5%bc%ba%e5%bc%80%e6%ba%90%e8%a7%86%e9%a2%91%e6%a8%a1%e5%9e%8bminimax-h3%ef%bc%8c%e4%bb%8e%e9%9b%b6%e8%a3%85%e6%9c%ba%e5%88%b0%e5%86%99%e5%87%ba%e4%b8%93","status":"publish","type":"post","link":"https:\/\/www.1ai.net\/en\/55633.html","title":{"rendered":"Locally deployed most powerful open source video model Mini Max H3, from a zero to a professional tip"},"content":{"rendered":"<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55635\" title=\"2b038a9fj00tjhbnr00knd000v90kcip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/2b038a9fj00tjhbnr00knd000v900cip.jpg\" alt=\"2b038a9fj00tjhbnr00knd000v90kcip\" width=\"1125\" height=\"450\" \/><\/p>\n<p>THE GOAL OF THIS LESSON: A COMPUTER WITH ONLY A SYSTEM AND A GRAPHIC CARD, FOLLOWED BY AN AI VIDEO WITH STEREO, AND LEARNED THE OFFICIAL HINT -- H3 IS GOOD AND BAD, AND MOST OF IT IS ON IT. AT EACH STEP, IT WAS WRITTEN \u201cWHAT TO DO, HOW TO DO, WHAT TO SEE AFTER\u201d, WITHOUT ANY CODE BASIS\u3002<\/p>\n<p><strong>one thing you have to know before you start: the native canvas of the local model is 768p\u3002<\/strong>\u00a0The model was only trained on the short side of a canvas of 768, and the official 2K came out of another closed-source module. But 768p is not a local ceiling: the resolution can be hard and bigger (someone runs 1920 x 1088 on 5090, the painting is really better at a cost of 10 seconds and 58 minutes) and the community already has a ready-made local overscore stream of 768p to pull 1080p - a two-way cut-off cut-off. First, the division of the three modules is clear, and the official source is only the middle:<\/p>\n<ul>\n<li><strong>H3-Context-IR<\/strong>\u00a0- UNDERSTANDING AND STRUCTURING OF THE HINTS (REFORM YOUR DESCRIPTION INTO A MODEL-FRIENDLY FORMAT). \u274c CLOSED SOURCE, API ONLY<\/li>\n<li><strong>H3-Base<\/strong>\u00a0- 33B generate the subject, output 768p level audio video. Zenium\u00a0<strong>Open Source, principal of the program<\/strong><\/li>\n<li><strong>H3-Regenerate-2K<\/strong>\u00a0- Reborn 768p to 2K. \u274c Closed source, API only<\/li>\n<\/ul>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55636\" title=\"999f98c7j00tjhbph006zd000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/999f98c7j00tjhbph006zd000m800cip.jpg\" alt=\"999f98c7j00tjhbph006zd000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<p>So the phrase \"local running H3 at no cost\" basically works: drafts, 768 p is free, 1080 p can be solved locally by community super-points; only \"official 2K\" must go API. When to use it, there's a clear account at the end\u3002<\/p>\n<p>The data is from a RTX 4090 48GB, but it's not the threshold - the officially confirmed floor is\u00a0<strong>3060 12GB + 32GB RAM<\/strong>I don't know. From 3060, 5070 to 3090, 5090, how much can be quantified, how fast can you run, there is a table for step 0. The small amount of the presence only affects how much you can specify and how long you can wait, and does not affect whether you can run\u3002<\/p>\n<p>catalogs<\/p>\n<ol>\n<li>Knows a few terms (30 seconds)<\/li>\n<li>Step 0: Web version inspection + hardware self-check<\/li>\n<li>Step 1: Install ComfyUI<\/li>\n<li>Step 2: Download Model File<\/li>\n<li>STEP 3: RUN THROUGH THE FIRST VIDEO (T2V)<\/li>\n<li>Step 4: Understand parameters - hard rules of frame and resolution<\/li>\n<li>STEP 5: TUSHENG VIDEO (I2V)<\/li>\n<li>STEP 6: REFERENCE VIDEO (R2V)<\/li>\n<li>Step 7: Full Guide to Phrasing<\/li>\n<li>Step 8: Accelerate - SageAttention 24%, stack Cache up to 58%<\/li>\n<li>Check for common problems<\/li>\n<li>Sound effect, alignment ahead of expectations<\/li>\n<li>Local vs API, account<\/li>\n<\/ol>\n<p>1. Understanding several terms (30 seconds)<\/p>\n<ul>\n<li><strong>ComfyUI<\/strong>: RUN FREE OPEN-SOURCE WORKSTATION FOR AI GENERATION MODEL. THE INTERFACE IS NODE-CONNECTED, BUT YOU DON'T HAVE TO DO IT YOURSELF -- OFFICIAL TEMPLATES ARE SET, YOU CHANGE PARAMETERS\u3002<\/li>\n<li><strong>Workflow<\/strong>: a lined node map, equivalent to a \u201cformula\u201d. Loading template = Opens a ready-to-use formulation\u3002<\/li>\n<li><strong>Weight \/ Model File<\/strong>: Model body, suffix\u00a0<strong>.safetensors<\/strong>\u00a0Big paper. H3 needs four: DiT (painted), text encoder (readers with hints), video VAE and audio VAE (translating internal data into pixels and sound waves)\u3002<\/li>\n<li><strong>Quantitative<\/strong>: Technology to minimize the model\u3002<strong>bf16<\/strong>\u00a0It's full of precision<strong>int8<\/strong>,<strong>nvfp4<\/strong>\u00a0It is a compressed version, with very small loss of paint quality and a significant decline in the demand for visibility. Local deployments are mostly in quantitative form\u3002<\/li>\n<li><strong>T2V \/ I2V \/ R2V<\/strong>Three modes of generation - pure text generation, giving pictures to move, and targeting characters or styles for reference material (chart\/video\/audio)\u3002<\/li>\n<\/ul>\n<p>Step 0: Web version inspection + hardware self-check<\/p>\n<p>Before we do it, it takes 10 minutes to check on the web page<\/p>\n<p>40 GB weight first. H3 In the conch AI page version (domestic hailuoai.com, overseas hailuoai.video) you can direct the test: login \u2192 video generation \u2192 model selection <a href=\"https:\/\/www.1ai.net\/en\/tag\/minimax\" title=\"[View articles tagged with [MiniMax]]\" target=\"_blank\" >MiniMax<\/a> H3 POACH A PICTURE OF THE PERSON IN THE PICTURE WAVING AT THE CAMERA, WITH THE SOUND OF THE STREET ENVIRONMENT, AND GENERATE IT\u3002<\/p>\n<p>Look at three things:<strong>The subject's unstable, the camera's moving naturally, the sound and the image, right<\/strong>I DON'T KNOW. THESE THREE ARE SATISFIED, THEN GO DOWN; IF NOT, YOU DON'T HAVE TO READ THE LESSON, SAVE THE REST OF THE DAY. THE SIZE OF THE PAGE VERSION AND THE API ARE TWO SETS OF ACCOUNTS, AND THE PAGE SIZE DOES NOT AFFECT THE MOVEMENT OF API\u3002<\/p>\n<p>Hardware self-check<\/p>\n<ul>\n<li><strong>Graphics<\/strong>Minimum NVIDIA 12GB display, recommended 24GB+. How to find: Windows press\u00a0<strong>Ctrl+Shift+Esc<\/strong>\u00a0\u2192 PERFORMANCE GPU \u2192 \"PILOT GPU RAM\"; OR COMMAND LINE LOSS\u00a0<strong>nvidia-smi<\/strong>.<\/li>\n<li><strong>Memory<\/strong>MINIMUM 32GB, RECOMMENDED 64GB. TASK MANAGER &amp; PERFORMANCE &amp; MEMORY\u3002<\/li>\n<li><strong>Disk Free<\/strong>MINIMUM 60 GB, RECOMMENDED 150 GB+. THE WEIGHT OF ABOUT 40 GB, PLUS THE GENERATION AND R2V WEIGHTS, WILL CONTINUE TO RISE\u3002<\/li>\n<li><strong>systems<\/strong>: Windows 10\/11 or Linux.<\/li>\n<li><strong>reticulation<\/strong>: For access to Hugging Face or its mirror image, see step 2 mirror scheme\u3002<\/li>\n<\/ul>\n<p>Right: What's your graphic card in<\/p>\n<p>The CofyUI core developer officially confirmed the floor line:<strong>3060 12GB + 32GB RAM + a nice NVME solid and can run 480p<\/strong>I don't know. 12GB has a 12GB run, find your own line of action:<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55638\" title=\"bd8fd8dfj00tjhb000aid000v9009sp\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/bd8fd8dfj00tjhbq000aid000v9009sp.jpg\" alt=\"bd8fd8dfj00tjhb000aid000v9009sp\" width=\"1125\" height=\"352\" \/><\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55634\" title=\"6fab2041j00tjhbq00gtd000m800tnp\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/6fab2041j00tjhbqb00gtd000m800tnp.jpg\" alt=\"6fab2041j00tjhbq00gtd000m800tnp\" width=\"800\" height=\"1067\" \/><\/p>\n<p>Data source: 5090 and 12GB cards from community measurements, 4090 mobile versions from 20-step measurements on SageAttention<\/p>\n<p>Two generic cross-slotting reminders:<\/p>\n<ul>\n<li><strong>System storage and solidity are as important as graphic cards\u3002<\/strong>\u00a0H3 Runs on a consumption-grade card only with a layer load of \"notable memory \" , 32 GB memory is the bottom line, 64 GB from the face \u2014 16 GB memory + large memory is not moving. The model goes from disk to memory, and the mechanical hard drive will keep the first run waiting for a few minutes, and the weight must be on NVME\u3002<\/li>\n<li><strong>30\/40 and 50 are a hidden difference: NVFP4 is only 50 (Blackwell) supported by hardware\u3002<\/strong>\u00a0When loading nvfp4 weights on 3090\/4090, ComfyUI follows a \"simulator\" - saves disks and visible occupancy, but before counting, depresses back to high accuracy, without taking time. So 50 users are bold enough to choose a community NVFP4 version of DiT (smaller and faster), and 30\/40 users are more cost-effective to choose INT8\/FP8\u3002<\/li>\n<\/ul>\n<p>Step 1: Installation of ComfyUI<\/p>\n<p><strong>Version shall be thallium 0.30.0<\/strong>, this is a hard threshold: the original support of H3 (4 dedicated nodes + official templates) was merged with Comfy-Org\/CommyUI #15224 on 3 August 2026, without these nodes in the old version and without any working flow\u3002<\/p>\n<p>STATUS A: COMPLETE NEW INSTALLATION (RECOMMENDED FOR WHITE)<\/p>\n<ol>\n<li>Open the ComfyUI official network and download the corresponding system<strong>desktop version<\/strong>Installation package\u3002<\/li>\n<li>Double-click installation, all the way by default. The installation automatically handles Python, PyTorch, CUDA dependency, which is why the desktop is friendly to new hands\u3002<\/li>\n<li>Finish loading start. The first start will be initialized, and when it's finished\u3002<\/li>\n<\/ol>\n<p>Case B: ComfyUI<\/p>\n<ul>\n<li><strong>desktop version<\/strong>: Check in the menu for updates to the latest stable version\u3002<\/li>\n<li><strong>manual guit deployment<\/strong>_Other Organiser\u00a0<strong>don't pull<\/strong>once again\u00a0<strong>pip install -r requirements.txt<\/strong>.<\/li>\n<li><strong>Integration pack (Autumn leaf pack, etc.)<\/strong>: When the integration package is updated by the author or the kernel is updated in the starter. Note: Desktop and integration packages follow<strong>Steady<\/strong>release, individual nightly features may arrive a few days later\u3002<\/li>\n<\/ul>\n<p>Verify installation successful<\/p>\n<p>Open Post Startup Browser\u00a0<strong>http:\/\/127.0.0.1:8188<\/strong>(desktop directly pops up window) See node canvas. Confirmed version: Settings (low left corner gear) \u2192 On, version number \u2265 0.30.0. If you're lower, go back and update. Don't try<strong>The most high-frequency error of the MiniMaxH3ImageToVideo node, 99% is a problem of the version<\/strong>.<\/p>\n<p>Step 2: Download model files<\/p>\n<p>The model hosts the Comfy-Org\/ MiniMax-H3 repository in Hugging Face\u3002<\/p>\n<p><strong>warning: do not close the entire warehouse\u3002<\/strong>\u00a0THE ORIGINAL WAREHOUSE IS ABOUT 318 GB, AND YOU ONLY NEED FOUR FILES\u3002<\/p>\n<p>Domestic download speed up<\/p>\n<p>Link any Hugging Face\u00a0<strong>i don't know<\/strong>\u00a0Replace\u00a0<strong>cf-miror.com<\/strong>And it's a mirror image of the country, and it's a magnitude difference. The large 20GB file proposes a download tool (IDM, aria2, or browser to bring back the download) to support the break-up, without having to repeat it\u3002<\/p>\n<p>We'll use the order line. One order will be precise, and four files will be executed\u00a0<strong>set HF_ENDPOINT=https:\/\/hf-miror.com<\/strong>, Linux\/macOS\u00a0<strong>export<\/strong>):<\/p>\n<p>pip install-U hugglingface_hub<\/p>\n<p>hf download Comfy-Org\/ MiniMax-H3 \\<br \/>\ndiffusion_models\/minimax_h3_fl2va_pruned_int8_convrot.safetensors\\<br \/>\ntext_encoders\/qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors \\<br \/>\nvae\/minimax_h3_video_vae_fp16.safetensors \\<br \/>\nvae\/minimax_h3_udio_vae_fp32.safetensors \\<br \/>\n\u2013local-dir ComfyUI\/models\/<\/p>\n<p>4 REQUIRED T2V \/ I2V (APPROXIMATELY 39.6 GB)<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55637\" title=\"88f8deedj00tjhbsb005rd000v9007ap\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/88f8deedj00tjhbsb005rd000v9007ap.jpg\" alt=\"88f8deedj00tjhbsb005rd000v9007ap\" width=\"1125\" height=\"262\" \/><\/p>\n<p>TWO VAES\u00a0<strong>All must go down<\/strong>- VIDEO VAE OUT, AUDIO VAE OUT, LESS VOICE VAE YOU'LL GET A SILENT VIDEO\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55639\" title=\"a69c1b9dj00tjhbsg004nd000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/a69c1b9dj00tjhbsg004nd000m800cip.jpg\" alt=\"a69c1b9dj00tjhbsg004nd000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<p>After the release, the directory should be long:<\/p>\n<p>I don't know<br \/>\nideas - models\/<br \/>\n_diffusion_models\/<br \/>\n\u2502 -minimax_h3_fl2va_pruned_int8_convrot.safetensors<br \/>\nideas -text_encoders\/<br \/>\n\u2502 qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors<br \/>\n&gt; vae\/<br \/>\n_minimax_h3_video_vae_fp16.safetensors<br \/>\n\u2500-minimax_h3_audio_vae_fp32.safetensors<\/p>\n<p>desktop user note: the model directory may be searched for \"model paths \" in the settings under the path selected for installation\u3002<\/p>\n<p><strong>More lazy<\/strong>: Skip manual download, direct step 3 - When loading the template, ComfyUI will pop a window to list the missing model and give the download button and click automatically down to the right position. The value of manual downloads is that they can be sent through mirrors and breakpoints, and people with a volatile network recommend it manually\u3002<\/p>\n<p>An overview of all official variants (options as required, most not available)<\/p>\n<p>Diffusion model (choose one; want to play R2V to play ref2va):<\/p>\n<ul>\n<li><strong>minimax_h3_fl2va_pruned_int8_convrot<\/strong>(19.5 GB) - \u2705\u00a0<strong>recommend<\/strong>,T2V\/I2V BEST BALANCE<\/li>\n<li><strong>minimax_h3_fl2va_pruned_fp8_scaled<\/strong>(19.5 GB) - ALTERNATIVE QUANTITATIVE FORMAT<\/li>\n<li><strong>minimax_h3_fl2va_int8_convrot<\/strong>(31.7 GB) - STANDARD INT8 (UNCUTED)<\/li>\n<li><strong>minimax_h3_fl2va_bf16<\/strong>(61.7 GB) - FULL PRECISION, FINE CALL<\/li>\n<li><strong>minimax_h3_ref2va_pruned_int8_convrot<\/strong>(19.5 GB) -\u00a0<strong>R2V MODE SPECIAL<\/strong>Step 6 will be used<\/li>\n<li><strong>minimax_h3_ref2va_bf16<\/strong>(61.7 GB)\u2014 R2V FULL PRECISION<\/li>\n<\/ul>\n<p>Text encoder (select):<\/p>\n<ul>\n<li><strong>qwen3vl_32b_minimax_h3_nvfp4_awq<\/strong>(14.6 GB) - \u2705\u00a0<strong>recommend<\/strong>ANY GPU CAN RUN<\/li>\n<li><strong>qwen3vl_32b_minimax_h3_int8_convrot<\/strong>(25.3 GB) - INT8<\/li>\n<li><strong>qwen3vl_32b_minimax_h3_bf16<\/strong>(48.0 GB) - FULL PRECISION<\/li>\n<\/ul>\n<p><strong>Why do you choose this<\/strong>(jumping does not affect operations):<\/p>\n<ol>\n<li><strong>H3 Eats a visible head of encoder, not 33B's Dit\u3002<\/strong>\u00a0The encoder directly uses the full weight of Qwen3-VL-32B (62.13 GiB under bf16, larger than the DiT of 61.73 GiB), where a 32B visual language model is only a \"problem.\" So the encoder has to select the unvfp4 version to be quantified. The online circulation of \u201ccodifier\u201d 51.5GB is incorrect, based on the size of the official warehouse file\u3002<\/li>\n<li><strong>pruned = cutting down the AdaLN branch of 13B\u3002<\/strong>\u00a06173 - 37.46 \u2248 24.3 GiB, exactly the 13B. The official confirmation that this part of the modem output can be expected to be a cache and that pure reasoning does not need to be loaded at all. Conclusion: Pruned for reasoning alone and fine-tuned for full weight\u3002<\/li>\n<\/ol>\n<p>Quantification of community selection by card slot (optional extra meals)<\/p>\n<p>Official\u00a0<strong>pruned_int8+nvfp4_awq<\/strong>\u00a0All-eat combinations. Any slot is recommended to run first. After running, think faster, save it and come here for dinner. The following are community-based third-party conversions (unofficial publication, drawings and permissions for self-checking of warehouses README), which are directly loaded by the original ComfyUI:<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55640\" title=\"bcbfd979j00tjhbt70d4d000v90eop\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/bcbfd979j00tjhbt700d4d000v900eop.jpg\" alt=\"bcbfd979j00tjhbt70d4d000v90eop\" width=\"1125\" height=\"528\" \/><\/p>\n<p>When changing community weights, remember one:<strong>Two VAEs, always in the official original<\/strong>- They're small, and quantifying them can only hurt paint and sound\u3002<\/p>\n<p>5. STEP 3: RUN THROUGH THE FIRST VIDEO (T2V)<\/p>\n<p>Load official templates<\/p>\n<ol>\n<li>Open CommyUI Top Menu<strong>Workflow<\/strong>\u00a0\u2192\u00a0<strong>Browse Templates<\/strong>\u00a0\u2192\u00a0<strong>video<\/strong>Classification, search \"Mini Max H3\"\u3002<\/li>\n<li>You'll see\u00a0<strong>6 templates<\/strong>Don't get me wrong<\/li>\n<\/ol>\n<ul>\n<li><strong>Mini Max H3 Text to Video<\/strong>\u00a0- Local, Vincent video<strong>Start with this<\/strong><\/li>\n<li><strong>Mini Max H3 Image to Video<\/strong>\u00a0- Local, Tusheng video (first frame\/tail frame)<\/li>\n<li><strong>Mini Max H3 Reference to Video<\/strong>\u00a0\u2013 local, reference-based video, with an additional ref2va weight<\/li>\n<li><strong>api_minimax_h3_t2v \/ r2v \/ flf2v<\/strong>\u00a0\u2013 API Edition to Mini Max Cloud, to fill API key, ignore first<\/li>\n<\/ul>\n<ol>\n<li>click on\u00a0<strong>Text to Video<\/strong>I don't know. If the bullet window hint is missing the model, download it by hint; step 2 is manually set and is ready (to restart the ComfyUI - model list scanned on startup without recognition)\u3002<\/li>\n<\/ol>\n<p>You know the key nodes in the work stream<\/p>\n<p>There's a line of nodes on the canvas when the template is loaded. All you need to know is these:<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55641\" title=\"b7644696j00tjhbth0006od000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/b7644696j00tjhbth006od000m800cip.jpg\" alt=\"b7644696j00tjhbth0006od000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<ul>\n<li><strong>UNETLoader<\/strong>Select the DiT weight to confirm the bottom box as\u00a0<strong>minimax_h3_fl2va_pruned_int8_convrot.safetensors<\/strong>.<\/li>\n<li><strong>Text Encoder Load Node<\/strong>: Confirm Selected\u00a0<strong>qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors<\/strong>.<\/li>\n<li><strong>Two VAE Loaders<\/strong>VIDEO VAE AND AUDIO VAE ARE SELECTED SEPARATELY\u3002<\/li>\n<li><strong>Resolutory Selector<\/strong>: Control output resolution (the width ratio, Megapixels, two parameters calculated as wide). Template default is a quick preview specification<strong>Don't move for the first time<\/strong>.<\/li>\n<li><strong>promptword box (positive)<\/strong>Write what you want\u3002<\/li>\n<li><strong>Length\/ Frame Input<\/strong>: determine the length of the video, see step 4\u3002<\/li>\n<li><strong>SaveVideo<\/strong>: OUTPUT END OF MP4\u3002<\/li>\n<\/ul>\n<p>Run<\/p>\n<p>In the prompting box, write an English scene (not to mention grammar, step 7 is a positive lesson):<\/p>\n<p>A little walkers along a long street at night, no light lights reflecting in puddles, soft rain sounds.<\/p>\n<p>point (in space or time)\u00a0<strong>Run<\/strong>.<\/p>\n<p>You should see something<\/p>\n<ul>\n<li>THE PROGRESS BAR IS FIRST LOADED IN THE MODEL (APPROXIMATELY 34 GB FROM DISK TO DISPLAY\/RAM) AND THEN GRADUALLY SAMPLED\u3002<\/li>\n<li><strong>It's normal for the first time to be very slow<\/strong>: 4090 48G LIVE HEAD RUN, 644 SECONDS (INCLUDING LOADING), 425 SECONDS WITH THE SAME SPECIFICATION; 768 X 432 SMALL SIZE HEAT RUN, ONLY ABOUT 99 SECONDS. OTHER SLOT ALIGNMENTS ARE EXPECTED: 5090 OUT OF 5 SECONDS APPROXIMATELY 108 SECONDS; 16 GB CARDS (E.G. 4090 MOBILE VERSION) OUT OF 960 X 540 OUT OF 5 SECONDS APPROXIMATELY 182 SECONDS; 12 GB CARDS ARE LOADED BY MEMORY STRATIFICATION AND 864 X 480 OUT OF 90 FRAMES ABOUT 6 MINUTES. THE FIRST RUN IS TO ADD A FEW MORE MINUTES OF LOADING TIME TO THESE NUMBERS, MAKING TEA, ETC\u3002<\/li>\n<li>After the video preview of the SaveVideo node:<strong>There's pictures, there's voices<\/strong>I DON'T KNOW. SOUND AND IMAGE ARE GENERATED IN THE SAME FRONT-TO-FACE TRANSMISSION, WHICH IS THE FUNDAMENTAL POINT OF THE H3 DISTINCTION FROM A POST-GENERATION VOICE MODEL\u3002<\/li>\n<li>If the console is brushed\u00a0<strong>Input tensors must be in dtype of torch<\/strong>\u2014\u2014<strong>It's not a mistake<\/strong>i don't know. the official document confirms that this is a normal hint for some layers back to standard attention, and that it is not affected\u3002<\/li>\n<\/ul>\n<p>It's over. The deployment phase is over. It goes from \"can run\" to \"can use.\"\u3002<\/p>\n<p>Step 4: Hard rules for understanding parameters - frames and resolutions<\/p>\n<p>These two rules read the code of H3 node of ComfyUI directly, not experiential\u3002<\/p>\n<p>rule i: the frame number must fall on the 17k+5 grid and be silently adsorbed<\/p>\n<p>Source logic is a line:<\/p>\n<p>\u266a When n n % 17! \u266a<br \/>\nn + = 1<\/p>\n<p>The frame you fill out, if it's not legal, it will<strong>Sniff up to the nearest legal value, no error, no hint<\/strong>I don't know. Measure 61 frames, get 73 frames\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55643\" title=\"fd83ff2aj00tjhbty004wd000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/fd83ff2aj00tjhbty004wd000m800cip.jpg\" alt=\"fd83ff2aj00tjhbty004wd000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<p>the training area is 124-362 frames, 24fps below\u00a0<strong>5.17 to 15.08 seconds<\/strong>;all legal slots in this compartment:<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55642\" title=\"bdc41f79j00tjhbu8006qd000v9000g4p\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/bdc41f79j00tjhbu8006qd000v900g4p.jpg\" alt=\"bdc41f79j00tjhbu8006qd000v9000g4p\" width=\"1125\" height=\"580\" \/><\/p>\n<p><strong>60 seconds and only 15.08 seconds<\/strong>, the longer content needs to be filmed and edited. the node tooltip is ~124-362, longer is unsettled - less than 124 frames can run (the grid starts from 5, 22, 39 ...), but quality is not assured outside the training distribution. the day-to-day picks it out of this table, so don't let it suck -- the six seconds you think and the actual 6.58 seconds, in a card-faced scene, is a disaster\u3002<\/p>\n<p>Rule II: The resolution's \"native canvas\" is short edge 768<\/p>\n<p>Model Press<strong>Short edge 768, area limit 768 x 1344, multiple of 32 per axis<\/strong>\"Train.\" Operational recommendations:<\/p>\n<ul>\n<li><strong>Try the hint<\/strong>: The template default fast preview specification (e.g. 768 x 432), approximately 99 seconds\/bar, fast iterative\u3002<\/li>\n<li><strong>Spectrum<\/strong>: Resolute Selector\u2019s Megapixels\u00a0<strong>1.0<\/strong>,16:9\u00a0<strong>1344 x 768<\/strong>- It's the original canvas, the best quality\u3002<\/li>\n<li><strong>floor 384p<\/strong>:256p and lower<strong>Total failure<\/strong>It's official confirmation, not your configuration\u3002<\/li>\n<li><strong>It's not impossible, it's uneconomical<\/strong>: the model has not been trained in a larger painting, but the hard drive can come up with a piece \u2014 a 10-second piece of 1920 x 1088, which is actually better than 768p, at a cost of 58 minutes. community consensus is..<strong>Low Resolution Generation + Local Excess<\/strong>: 768p scale is followed by a LTX 2.3 or Wan 2.2 5B ultra-spectrum workflow to 1080p, and the same 1920 x 1080 finished product is reduced from approximately 700 seconds to 161 seconds, 6 times faster, and the paint is close to the status profile. Want the official original 2K, take the API regeneration -- three paths to the end\u3002<\/li>\n<\/ul>\n<p>Two realistic patterns of presence and time-consuming<\/p>\n<p><strong>It's a model, not a painting\u3002<\/strong>\u00a0The peaks of 768 x 432 and 1152 x 640 were almost identical (the difference was &lt; 1%). Process-level detail: The sampling phase is stable on 19 GB, the coding\/decode phase is running at once to 32 GB - this is the reading of 48 G card unrun, the dynamic load of ComfyUI adapts to the empty visibles, and the small visible cards can run by switching to and from, but only slower. Meaning: The model can be installed, the resolution can be driven to the original canvas; it can&#039;t be installed, the lower resolution can&#039;t save you, but a smaller quantitative version. OOM does not waste time changing sizes\u3002<\/p>\n<p><strong>Time-consuming patterns: Pixel dimensions are calculated and time-consuming\u3002<\/strong>\u00a01152 x 640 relative 768 x 432 is 2.22 pixels, which takes only 1.86 times more time, sublinearity, and high resolution is more intuitive than intuitive; but the length of time is the opposite - attention is ultra-linear, double the length, more than double the time\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55645\" title=\"8471db3fj00tjhbuj0071d000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/8471db3fj00tjhbuj0071d000m800cip.jpg\" alt=\"8471db3fj00tjhbuj0071d000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<p>Specification scale measurements (4090 48G, SageAttention cuda++ + acceleration):<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55644\" title=\"1cb358c5j00tjhbv0002fd000of00eep\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/1cb358c5j00tjhbv0002fd000of00eep.jpg\" alt=\"1cb358c5j00tjhbv0002fd000of00eep\" width=\"879\" height=\"518\" \/><\/p>\n<p>Read the table three times:<strong>Five seconds, five minutes, 10 seconds, 12 minutes, 15 seconds, one 19 minutes<\/strong>I don't know. Note also that 1536x864 relative 1152x640 is 1.8 pixels, time-consuming close to double the same time \u2014 the next linear dividend of the painting is gone, and it is confirmed once again that the supernatural canvas is not available. The replicability of the data has been tested: reruns with the specifications on the next day, 320.11s against 32.05s, error 0.6%\u3002<\/p>\n<p>STEP 5: TUSHENG VIDEO (I2V)<\/p>\n<p>Purpose: To move a ready-made image, or to replace the middle motion with two at the end\u3002<\/p>\n<ol>\n<li>Template Library Loading\u00a0<strong>Mini Max H3 Image to Video<\/strong>I don't know. The weight and T2V are identical (fl2va) and do not need to be downloaded again\u3002<\/li>\n<li>use\u00a0<strong>LoadImage<\/strong>\u00a0We'll upload your map\u00a0<strong>MiniMaxH3ImageToVideo<\/strong>\u00a0Node\u00a0<strong>first_frame<\/strong>\u00a0Enter\u3002<\/li>\n<li><strong>last_frame<\/strong>\u00a0<strong>It's optional<\/strong>: ONLY FOR THE FRAME = FROM THIS MAP ONWARDS; ALL AT THE END = THE MODEL DISPLAYS A CONSISTENT MOVEMENT BETWEEN THE TWO GRAPHS (THE \"A TO B\" LENS SUITABLE FOR PRODUCT ROTATION, ATTITUDE CHANGE, ETC.); ONLY FOR THE FRAME = THE MODEL REVERSES A REASONABLE OPENING, AND EVENTUALLY FALLS ON YOUR GRAPH\u3002<\/li>\n<\/ol>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55646\" title=\"f342ba96j00tjhbva004wd000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/f342ba96j00tjhbva004wd000m800cip.jpg\" alt=\"f342ba96j00tjhbva004wd000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<ol>\n<li>THE HINT IS WRITTEN IN THE 7-STEP I2VA FORMAT (NEEDS A LINE FIRST ALIGNMENT COMMAND)\u3002<\/li>\n<li>The input diagram will be adapted to generate resolution and too much difference will be processed --<strong>Before uploading, customize the map to match the width of the output<\/strong>Saved by accidental cutting\u3002<\/li>\n<\/ol>\n<p>STEP 6: REFERENCE VIDEO (R2V)<\/p>\n<p>USE: LOCKING A CHARACTER, A PAINTING STYLE, AN ACTION, A MIRROR OR A SOUND TO MAKE THEM APPEAR IN A NEW VIDEO. THIS IS THE STRONGEST AND MOST COMPLEX PATTERN OF H3\u3002<\/p>\n<p>Ready<\/p>\n<p>R2V<strong>Another weight<\/strong>\u00a0<strong>ref2va<\/strong>, AND T2V\/I2V\u00a0<strong>fl2va<\/strong>\u00a0Not universal:<\/p>\n<ol>\n<li>Go back to Comfy-Org\/ MiniMax-H3 download\u00a0<strong>minimax_h3_ref2va_pruned_int8_convrot.safetensors<\/strong>(19.5 GB), SAME\u00a0<strong>Photo by Flickr user ComfyUI\/models\/diffusion_models\/<\/strong>.<\/li>\n<li>ENCODERS AND TWO VAES ARE REUSED WITHOUT RESET\u3002<\/li>\n<li>Template Library Loading\u00a0<strong>Mini Max H3 Reference to Video<\/strong>, confirm UNETLoader cuts the ref2va weight\u3002<\/li>\n<\/ol>\n<p>Use rules<\/p>\n<ul>\n<li><strong>Quantity ceiling<\/strong>: Up to 9 reference charts, 3 reference videos (each with its own track), 3 independent reference audio\u3002<\/li>\n<li><strong>Use tab references in connect order<\/strong>: The first is, the second is\u00a0<strong>\u00a0<\/strong>Take that kind of push. These labels must be used to name names in the message\u3002<\/li>\n<li><strong>For each reference \"Piracy\"<\/strong>: specify which reference line is what \u2014 look, style, action, mirror or sound. The authorities have made it clear that the visible assignment is much more effective than putting a pile of material behind it\u3002<\/li>\n<li><strong>{\\bord0\\shad0\\alphah3d}ref_image_size<\/strong>\u00a0<strong>parameter<\/strong>:<strong>watch<\/strong>(Default) zoom in to generate resolution, fast<strong>max<\/strong>\u00a0keeps a maximum of 2048 px short edges, and the role is more solid<strong>but token, every step of the sample is complete<\/strong>I don't know. Daily\u00a0<strong>watch<\/strong>Only when your face can't be locked\u00a0<strong>max<\/strong>.<\/li>\n<\/ul>\n<p>STRUCTURAL DIFFERENCES IN R2V HINTS<\/p>\n<p>R2V 'S COMPLETE HINT IS A SIX-PART STYLE (THREE MORE THAN THE T2V FIELD) WITH A FIXED ORDER:<strong>subject_definitions<\/strong>(Defines labels and features for each reference)\u00a0<strong>summary<\/strong>(Summary paragraph beginning with the type of task in square brackets)\u00a0<strong>retention_analysis<\/strong>(label-by-label declaration of retention)\u00a0<strong>detailed_description<\/strong>(Major, 350-500 English, opening style before [Shot 1])\u00a0<strong>overall_sundscape<\/strong>\u00a0\u2192\u00a0<strong>no, no, no<\/strong>.<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55648\" title=\"ee0469aej00tjhbvn00b2d000m800gop\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/ee0469aej00tjhbvn00b2d000m800gop.jpg\" alt=\"ee0469aej00tjhbvn00b2d000m800gop\" width=\"800\" height=\"600\" \/><\/p>\n<p>Each of the four labels:<\/p>\n<ul>\n<li>\u00a0- What to repeat in a film: people, scenes, costumes, styles, actions<\/li>\n<li>\u00a0- A map is directly used as a frame or structure anchor<\/li>\n<li><strong>\u00a0<\/strong>\u00a0- Total level of relationship: edited source, starting point of renewal, source of cut rhythm<\/li>\n<li>\u00a0- Audios copied or referenced<\/li>\n<\/ul>\n<p>One of the easiest things to get wrong:<strong>Pictures are only used to define roles or styles<\/strong>\u00a0<strong>, write in the source<\/strong>\u00a0\u00a0<strong>In the definition<\/strong>I don't know. This is the \"This is the frame\" scene\u3002<\/p>\n<p>For the other two paragraphs:<strong>summary<\/strong>\u00a0Click task type in square brackets at the beginning, multiple\u00a0<strong>+<\/strong>\u00a0Company--<strong>keepframe command<\/strong>\u00a0\/\u00a0<strong>i'm sorry<\/strong>\u00a0\/\u00a0<strong>i'm sorry<\/strong>\u00a0\/\u00a0<strong>video conversion<\/strong>\u00a0\/\u00a0<strong>audio return<\/strong>\u00a0\/\u00a0<strong>audio reference<\/strong>i don't know. adjudication: reference video provides only mirrors and rhythms = reference promotion, really changing the video eding\u3002<strong>retention_analysis<\/strong>\u00a0Gives each label a fixed relationship word: the visible content is used\u00a0<strong>full_preserved<\/strong>\u00a0\/\u00a0<strong>i'm sorry<\/strong>\u00a0\/\u00a0<strong>attribute_transfer<\/strong>\u00a0\/\u00a0<strong>weak_reference<\/strong>Audio\u00a0<strong>full_copy<\/strong>\u00a0\/\u00a0<strong>i'm sorry<\/strong>\u00a0\/\u00a0<strong>reference<\/strong>\u00a0\/\u00a0<strong>weak_reference<\/strong>.<\/p>\n<p>It's the first time that this format has receded, but it's the output format of the official pay module Context-IR -- you're a handwritten whore. The rookies don't have to come up with six paragraphs at a time: first, replace the material with the illustrative hints that the template contains, then sew the entire case-by-case version of the official R2V guide, and write a template at a time, with a few words at a time. Just remember:<strong>Every reference in<\/strong>\u00a0<strong>subject_definitions<\/strong>\u00a0<strong>It is marked with a label and a characterization, which is used in all subsequent paragraphs to describe how much it has been retained\u3002<\/strong><\/p>\n<p>Step 7: A guide to the integrity of the narrative<\/p>\n<p>H3 IS A STRUCTURED SET OF OFFICIAL DEFINITIONS (ORIGINAL VERSION OF THE OFFICIAL GUIDE). A SINGLE WORD CAN ALSO COME OUT, BUT MULTIPLE LENSES, CHARACTER PAIRS, SPECIFIED MIRRORS, CARD TIME POINTS MUST BE FORMATTED\u3002<strong>All written in English except for the original language of white and graphic text\u3002<\/strong><\/p>\n<p>9.1 Bones: three fields<\/p>\n<p>[Shot 1]..<\/p>\n<p>overall_sundscape:<\/p>\n<p>no, no, no<\/p>\n<ul>\n<li><strong>integrated_multimedia_description<\/strong>\u00a0- Subject. It's time-lined, action, camera, talking person, confidant, inside sound\u3002<\/li>\n<li><strong>overall_sundscape<\/strong>\u00a0- 1-4 sentence, sound of the environment and action (wind and rain, footsteps, clothing friction, breathing, laughter). Don't write: white, singing, painting music\u3002<\/li>\n<li><strong>no, no, no<\/strong>\u00a0- 1 to 3 words, music (a character who can't hear, only an audience): instruments, speed, rhythm, power and power. Don't write: emotional words, \"sorted\" \"historic\" and explain the role of music\u3002<\/li>\n<\/ul>\n<p>Write the last two fields without content\u00a0<strong>N\/A<\/strong>(Soundscape only if the user expressly requests full silence\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55647\" title=\"58719b1aj00tjhbvy0073d000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/58719b1aj00tjhbvy0073d000m800cip.jpg\" alt=\"58719b1aj00tjhbvy0073d000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<p>9.2 Camera: [Shot N] and switch time<\/p>\n<ul>\n<li><strong>[Shot 1]<\/strong>\u00a0Initial answer<strong>Whole style + initial diagram<\/strong>, without a time stamp. Style words:<strong>Cinematic<\/strong>,<strong>live-action<\/strong>,<strong>2D-animatized<\/strong>,<strong>3D CG<\/strong>,<strong>playmation<\/strong>,<strong>watercolor<\/strong>,<strong>video field<\/strong>.<\/li>\n<li>Follow-up camera tape strict incremental transition time:<strong>[Shot 2] At 00:03.500, the camera cuts to..<\/strong><\/li>\n<li>Toggle verb:<strong>i mean, the camera cups to<\/strong>\u00a0\/\u00a0<strong>the shot transfers to<\/strong>\u00a0\/\u00a0<strong>you know, the shot switches to<\/strong>;cross-dissolve, fade, wipe.<\/li>\n<li><strong>When do you cut the camera<\/strong>: Switch should introduce new information (new subject, new space, new perspective, new time). Just pull or fine-tune the angle, use the mirror, don't cut it\u3002<\/li>\n<\/ul>\n<p>9.3 Mirror: Type + Range + Speed<\/p>\n<p>Writes a natural English sentence in the lens and does not stack tags at the end of the sentence\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55649\" title=\"76eada63j00tjhbw70c0d000jt00v8p\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/76eada63j00tjhbw700c0d000jt00v8p.jpg\" alt=\"76eada63j00tjhbw70c0d000jt00v8p\" width=\"713\" height=\"1124\" \/><\/p>\n<p>Range:<strong>with small amplitude<\/strong>\u00a0\/\u00a0<strong>\u266a with broad angle \u266a<\/strong>;velocity:<strong>at slow speed<\/strong>\u00a0\/\u00a0<strong>i don't know, at last<\/strong>I don't know. Medium range, normal speed<strong>Directly omitted<\/strong>.<\/p>\n<p>The camera pushes in with small movements at low speed told the fed better in her hands.<br \/>\nThe game is right with broad expression at last, revealing the open door.<br \/>\nThe game holds a status shot as the runner exits the fire.<\/p>\n<p>9.4 CONVERSATION: SPEAKER ID + TAG<\/p>\n<p>List of rules:<\/p>\n<ol>\n<li>ID:<strong>(S1)<\/strong>,<strong>(S2)<\/strong>;many-person\u00a0<strong>(S1, S2)<\/strong>;<strong>CROSS LENS ID REMAINS UNCHANGED<\/strong>; never a silent character is numbered\u3002<\/li>\n<li>For the first time, the speaker is present at the foot anchor: type of role, age, sex, presence in the picture, sound, sound, speed, accent\u3002<\/li>\n<li>IDENTITY, ID, ACTION, TONE \u00a0<strong>Outside<\/strong>; Only language tags and lines themselves<strong>No change, no translation, no retention of original markers<\/strong>.<\/li>\n<\/ol>\n<p>The young woman with a quiet, Breathy voice says:<br \/>\n[English] Wait for us<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55650\" title=\"ec7a7231j00tjhbwi005xd000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/ec7a7231j00tjhbwi005xd000m800cip.jpg\" alt=\"ec7a7231j00tjhbwi005xd000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<ol>\n<li><strong>An external voice<\/strong>It has to be fixed\u00a0<strong>says in an off-screen voiceover<\/strong>And the people's lips are not moving:<\/li>\n<\/ol>\n<p>The man (S1) says in an off-screen voiceover: [English] I still remember that road.<\/p>\n<ol>\n<li>Line<strong>Cross the cut point<\/strong>: to be placed in both sections of the connection and clearly write the audio series (e. g\u00a0<strong>the carries over from the previous shot<\/strong>); lines are spoken<strong>Snippets<\/strong>With\u3002<\/li>\n<li><strong>Image Text<\/strong>(Signatures, subtitles, neon lights): Put in double quotation sign in English and leave the original untranslated -<strong>A red neon sign reading \"in business\" brings above the doorway.<\/strong><\/li>\n<\/ol>\n<p>9.5 A COMPLETE COPYABLE EXAMPLE (OFFICIAL T2VA CASE)<\/p>\n<p>live-action: [Shot 1] Live-action, a medium-widge shot justices a bullet opening the system of a small street before sunrise.<\/p>\n<p>wooden begins open over a quiet set as lessons clear only inside the bakery.<\/p>\n<p>no, no, no, no, no.<\/p>\n<p>REPLACED WITH YOUR SCENE, IT'S A GOOD H3 HINT\u3002<\/p>\n<p>9.6 Three variants with graphs: start with one additional line alignment command<\/p>\n<p>I2V FAMILY (I2VA \/ FL2VA \/ L2VA) IN THREE FIELDS<strong>Before<\/strong>An additional line of fixed format alignment commands, followed by an empty line:<\/p>\n<p><strong>I2VA ONLY<\/strong>- Fixed:<\/p>\n<p>For the target video, at 0.00 seconds into the target video, is fully referred.<\/p>\n<p>Description of the structure: First the style, the main body, the structure in the anchor map, then how the action will be carried out (first frame anchor embezzled action start, continuous development results). Role clothing, colour, key items, spatial relationship and graphic consistency\u3002<\/p>\n<p><strong>FL2VA<\/strong>\u2014 Declare the point at which each of the two charts corresponds:<\/p>\n<p>How the reference pictures aligns with the target video \u2014 locals with the 0.00-second mark of the target video; locals with the 8.00-second mark of the target video.<\/p>\n<p>Don't describe the two images in the text<strong>Write the path that connects them<\/strong>(THE INITIAL FRAME STATE IS THE VISIBLE INTERMEDIATE CHANGE THE GAP GRADUALLY NARROWS THE END FRAME STATE). FL2VA TRYS TO BE SINGLE AND ALLOWS THE MODEL TO INSERT A CONTINUOUS VALUE\u3002<\/p>\n<p><strong>L2VA ONLY<\/strong>_Text of the image:<\/p>\n<p>How the reference pictures align with the target video  along with the 6.00-second mark of the target video.<\/p>\n<p>Describe the structure: deduce a reasonable pre-state, a clear movement and a transition path, a last mirror gradually shrunk down on the map\u3002<\/p>\n<p>time\u00a0<strong>S.SS<\/strong>\u00a0It must be accurate to two decimals and consistent with your actual video time (converted against the 6-step frame table)\u3002<\/p>\n<p>9.7 Implicit laziness<\/p>\n<p>This is how the official design is: the API version of H3 has a closed-source Context-IR module dedicated to \"reforming people's words into the above-mentioned formats,\" which is not available locally, but you can have any LLM on your behalf. Throw the official guide link with your idea:<\/p>\n<p>Learn how to write T2V papers from https:\/\/huggingface.co\/MiniMaxi\/MiniMax-H3\/blob\/main\/docs\/VIDEO_PROMPT_WRITING_GUIDE_base_en.md<br \/>\nthen my idea into the three-field H3 policy format:<br \/>\n[Your Chinese thought]<\/p>\n<p>The results are checked in three places: if the lines have been rewrited (must be in the original language), if the time stamp of the lens is increasing, and if there are mood words in the total length of the time and in the play field\u3002<\/p>\n<p>I DON'T WANT TO RELY ON THE LLM'S TOUCH, BUT THERE'S A CORRECT LAZY PATH:<strong>Cost-IR API<\/strong>(\u00a55.80\/million token input) Throw in your big whites and material, and it returns the enhanced hint in standard format -- this is the closed-source module of the API version that rewrites your tip. Use the return result as a template, and change the words to a few words at a time, much faster than learning from zero, as is recognized by local ComfyUI\u3002<\/p>\n<p>Step 8: Accelerating - SageAttention 24%, stacking Cache up to 58%<\/p>\n<p>Run for it and get this far. Accelerator in two layers, superseding:<strong>SageAttention counts every step faster, and the Cache node skips the unnecessary step\u3002<\/strong>\u00a0First tier - same seed and workstream (1152 x 640\/124 frame) A\/B Measurement 24%:<\/p>\n<ol>\n<li><strong>LoadSageAttention<\/strong>: download from SageAttention returns the wiel that matches your PyTorch\/ CUDA version (torch and cu versions in the file name, match then down), and then\u00a0<strong>pip install\u00a0<\/strong>I don't know. Desktop user executes the Python environment that is used by ComfyUI in its built-in terminal\u3002<strong>version approval 2.x<\/strong>- The back fp8 mode is only available in 2.x, 1.x has no kernel at all; these wiels are pre-compiled by 2.x and are directly the most economical. It really has to come from the source code itself, with a hard threshold: sm_89(40) requires CUDA \u226512.4 (written in setup.py), and we can't make it in 12.0, and we can't get it up to 12.8\u3002<\/li>\n<li><strong>Load KJNodes Node Pack<\/strong>:CommyUI Manager Search\u00a0<strong>Photo by ComfyUI-KJNodes<\/strong>\u00a0installing;or manual\u00a0<strong>it's not like you're in love with me, it's not like you're in love with me<\/strong>\u00a0until (a time)\u00a0<strong>I'm sorry<\/strong>Start again\u3002<\/li>\n<li><strong>Connect<\/strong>: Workstream Riga\u00a0<strong>Patch Page Attention KJ<\/strong>\u00a0Node in\u00a0<strong>UNETLoader<\/strong>\u00a0and\u00a0<strong>BasicGuider<\/strong>\u00a0Between - UNETLoader 's model output \u2192 Patch node model input, Patch node model output \u2192 BasicGuider input. Only guider needs a patch, scheduler don't move\u3002<\/li>\n<\/ol>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55651\" title=\"269b5c16j00tjhbxe004sd000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/269b5c16j00tjhbxe004sd000m800cip.jpg\" alt=\"269b5c16j00tjhbxe004sd000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<ol>\n<li><strong>Select Mode<\/strong>Direct use\u00a0<strong>autumn<\/strong>\u00a0all right. it automatically falls on the speediest path by graphics -- 4090 on the auto is equal to\u00a0<strong>qk_int8_pv_fp8_cuda++<\/strong>(checked source code: the sm89 branch returns exactly the same fp8 kernel, manual selection is not bad but unnecessary)<strong>30 is card (3090\/3060 et al. Ampere) without FP8 hardware units<\/strong>, auto will reach the int8 path available to Ampere, and there is still a considerable acceleration. I really need to lock it manually\u00a0<strong>qk_int8_pv_fp8_cuda++<\/strong>, provided that step 1 is indeed SageAttention 2.x-1.x without this kernel, the selection is a misstatement or a silent retreat\u3002<\/li>\n<li>Normal Run\u3002<strong>Lazy replacement<\/strong>: do not want node to start ComfyUI plus\u00a0<strong>_other organiser<\/strong>\u00a0Parameter global opening (the desktop version is added in the set startup parameter). The two pits must be clear:<strong>THIS PARAMETER AND KJ TWO AND ONE, NEVER OPEN AT THE SAME TIME<\/strong>\u2014 Both sides repeat patches, while community feedback slows; and no start-up mode is available and only the default path is followed. To control the nodes, to save things, to use parameters\u3002<\/li>\n<\/ol>\n<p>Second level: Cache Node, not hard to count<\/p>\n<p>SageAttention is the result of faster counting every step, and Cache-like nodes are another idea: the adjacent output of the diffuse sample is often almost the same, and changes to a certain extent are directly replicating the previous step<strong>It doesn't count<\/strong>I DON'T KNOW. TWO LAYERS SUPERHEAVY, THE SAME 362 WITH A FULL LENGTH PIECE (1152 X 640, 4090 48G):<\/p>\n<ul>\n<li><strong>SageAttention Only<\/strong>\u00a0\u2013 1183.6s (no acceleration baseline 21.9%; and 1158s in step 4 are measured in different batches and 2% inter- batch fluctuations are normal)<\/li>\n<li><strong>Sage + EASYCache<\/strong>\u00a0\u2013 899.5s (reducing 24.0% based on Sage)<\/li>\n<li><strong>Sage + TeaCache<\/strong>\u00a0\u2013 500.5s (sage save 57.7%<strong>More than just driving Sage<\/strong>\uff09<\/li>\n<\/ul>\n<p><strong>EasyCache is the ComfyUI native node<\/strong>No third-party bag: add one to the canvas\u00a0<strong>EasyCache<\/strong>\u00a0Node, serial to model link (and Sage patch node lined around), key parameter is\u00a0<strong>i'm sorry<\/strong>(Other\u00a0<strong>start_percent<\/strong>\u00a0\/\u00a0<strong>end_percent<\/strong>\u00a0\/\u00a0<strong>verbose<\/strong>\u00a0Three, normally, do not move. Default 0.2 can be used directly. Note that the 899.5s above are measured under a more conservative 0.1 - the more the threshold jumps, the faster the 124 frame shorts are measured 0.2 run 215.7s, 0.1 run 260.0s, with a default value that will be faster than the figures in the text and a slightly greater loss. TeaCache\u00a0<strong>H3 SPECIALIZED<\/strong>Icyoung\/CommyUI-Mini MaxH3-TeaCache\u00a0<strong>Mini MaxH3TeaCache<\/strong>), the other option is lihaoyun6\/CommyUI-MiniMaxH3-Cache\u00a0<strong>Mini MaxH3Cache<\/strong>).<strong>Don't pretend to be Manager's universal<\/strong>\u00a0<strong>ComfyUI-TeaCache (welltop-cn)<\/strong>\u2013 Its support list stops at the FLUX\/ HiDream generation, without H3, and when you're finished, you find no point to connect to the H3 link. Key parameters are:\u00a0<strong>rel_l1_thresh<\/strong>The more you jump, the more you jump\u3002<\/p>\n<p>There is no free lunch - Cache's acceleration comes from skipping real calculations, but the price depends on the place. We checked the frame by frame under R2V reference pattern:<strong>Faces don't change<\/strong>(Same face, same hair, same dress structure, steady from head to end) What really falls is the \u201cactivity\u201d of the picture \u2014 the near-stilling \u201cdeep\u201d frame ratio from 13.8% to 41.5%, which is easier to break. So when Cache is in a film, don't look at her face, see if the picture's up and down and down\u3002<\/p>\n<p>The other one is ahead:<strong>Cache changed to produce the results themselves, not \u201crun faster on the same note\u201d\u3002<\/strong>\u00a0As seen, TeaCache output to SIM with no Cache baseline is only 0.78 (EasyCache is 0.96) - so the \"draft on TeaCache pick-up, face-to-face reassembly\" intuitive play is not working: instead of configuration, the image you picked changed. The right one:<strong>Sage + TeaCache, do not expect the draft image to be reproduced as it is; positive film that requires it, from selection to finalization, with the same conservative configuration<\/strong>(Sage + EASYCache or simply Sage). Thresholds are down, they're more conservative, they're smaller\u3002<\/p>\n<p>11. Inventory of frequently asked questions<\/p>\n<p><strong>Q: Tip missing Mini MaxH3ImageToVideo \/ EmttyMini MaxH3LatentAV<\/strong>\u00a0A: ComfyUI version &lt; 0.30.0, upgrade. Most high-frequency problems. Desktop\/Integrator Packages follow the steady release and are updated and retested\u3002<\/p>\n<p><strong>Q: Bomb \/ CUDA out of memory<\/strong>\u00a0A: ORDERED \u2014 1\u00a0<strong>pruned_int8_convrot<\/strong>\u00a0DIT+\u00a0<strong>nvfp4_awq<\/strong>\u00a0Encoders do not miss the smallest official combination, bf16;2 turn off other visible memory-eating programs (games, another ComfyUI, running model browser labels);3 saves in the system at above 32GB, 12GB cards are loaded by memory fractions;4 not yet, replaces community INT4 quantification (end of step 2). Remember the pattern:<strong>LOWER RESOLUTION WON'T SAVE OOM<\/strong>It's the model itself\u3002<\/p>\n<p><strong>Q: OUTPUT VIDEO WITHOUT SOUND<\/strong>\u00a0A: BOTH VAES MUST BE LOADED..<strong>video_vae_fp16<\/strong>\u00a0and\u00a0<strong>audio_vae_fp32<\/strong>And confirm that there's work in it\u00a0<strong>VAEDECodeAudio<\/strong>\u00a0Node attached\u00a0<strong>SaveVideo<\/strong>I don't know. Official templates are usually not missing, and it is easier to delete errors when you change your own workflow\u3002<\/p>\n<p><strong>Q: GENERATE COMPLETELY FAILED, RESOLUTION IS VERY SMALL<\/strong>\u00a0A: H3 lowest 384p, 256p and the following are inevitable failures. Preset with the resolution of an official template\u3002<\/p>\n<p><strong>Q: THE VIDEO IS DIFFERENT FROM THE ONE I FILLED OUT<\/strong>\u00a0A: Frames are adsorbed to the 17k+5 grid (step 6) with a single cap of 15.08 seconds. Pick directly from the legal frame list and do not fill in any value\u3002<\/p>\n<p><strong>Q: R2V TEMPLATE \"NO MODEL FOUND\"<\/strong>\u00a0A: R2V\u00a0<strong>ref2va<\/strong>\u00a0WEIGHTS, AND T2V\/I2V\u00a0<strong>fl2va<\/strong>\u00a0It's two, downloading alone (step 6)\u3002<\/p>\n<p><strong>Q:R2V PARTICULARLY SLOW<\/strong>\u00a0A: INSPECTION\u00a0<strong>{\\bord0\\shad0\\alphah3d}ref_image_size<\/strong>\u00a0Did you choose\u00a0<strong>max<\/strong>\u2014refer token, take every step of the sample<strong>max<\/strong>\u00a0It's several times slower. Daily\u00a0<strong>watch<\/strong>.<\/p>\n<p><strong>Q: REFERENCE VIDEO\/INPUT CHART OUT OF SHAPE, DECORATION<\/strong>\u00a0A: THE REFERENCE VIDEO WILL BE PRESSURIZED TO THE \"SHORT EDGE 768, SIZE LIMIT 768 X 1344, ALIGNMENT 32\" CANVAS, AND OUTPUT RESOLUTION IS TWO SETS OF RULES. THE MATERIAL WAS FIRST DESIGNED TO BE A WIDE-RANGING RATIO\u3002<\/p>\n<p><strong>Q: Console brush dtype warning<\/strong>\u00a0A: Normal phenomenon (step 5), partial retreat standard attitudinal, without affecting generation\u3002<\/p>\n<p><strong>Q: MY 3090 IS SEVERAL TIMES BEHIND THE OTHERS' 4090<\/strong>\u00a0A: Two reasons for supersing - 1,390 is the Ampere architecture, without the FP8 hardware unit, INT8\/FP8 power is important to turn back to high accuracy before computing, and the comparison between generations is naturally slow; 2 community members have observed that the dynamic H3 display is conservative, and that the 24GB card may only be used at around 18GB, and that the known underutilization is not your fault. Capable: OpenSageAttention<strong>autumn<\/strong>\u00a0Modes, INT8 Lean weight for mass-oriented, add memory to 64GB less layer to wait<\/p>\n<p><strong>Q: 16GB card (5070 Ti \/ 4080) goes off, \"Device memory is nearly full\"<\/strong>\u00a0A: THIS IS A DISPLAY OF THE LOADING PHASE + MEMORY DOUBLE-TIGHTENING. SEQUENCED: 1 SYSTEM 32GB IS THE BOTTOM LINE, ADDED TO 64GB MOST EFFECTIVE; 2 PLUS START PARAMETERS\u00a0<strong>\u2013cache-none-disable-smart-memoory<\/strong>\u00a0Force the model to be fully off-loaded;3 turn off the browser hardware accelerator (the browser itself will account for 1-2GB);4 replace the INT4\/NVFP4 small weight at the end of step 2. There are also 5070 Ti cases of unusual fluctuations in the speed of user reporting, which are significantly slower than the same card and are updated to the latest stable version of ComfyUI and graphic card-driven comparison<\/p>\n<p><strong>Q: GENERATED PERSON DOES NOT SPEAK CHINESE ENOUGH \/ THE LINE HAS BEEN CHANGED<\/strong>\u00a0A: CHECK LABEL - LANGUAGE TAG SYNTH<strong>[ Chinese ]<\/strong>) with the original line in the label and the identity and tone description outside the label. The lines are usually rewritten because they're written outside\u3002<\/p>\n<p>12. Sound effects, alignment ahead of expectations<\/p>\n<p>The official campaign \u201cPersonal 32kHz stereo\u201d, which measured decomposition: 32kHz, double-sounding, does have different left and right channels (L\/R 0.2\u20130.9 dB, which is lower than the main signal 14-24 dB) and is not a one-channel copy of two. But it is expected to be right:<strong>It's a \"space-sensitive\" narrow field stereo, not a separation of powerful images<\/strong>, do not count on the right- and right-crossing sound positioning. It's one of the most valuable things: white, sound, music and images are generated in a front-to-face transmission -- lips are born, not lateral\u3002<\/p>\n<p>13. Local vs API, one account<\/p>\n<p>OFFICIAL API PRICES (RMB):<strong>2K ~0.80\/S, 768P ~0.50\/S, 768P ~2K GENERATE ~0.30\/S<\/strong>; audio input is free of charge, and pictures 5 are free of charge (over \u00a50.20\/)\u3002<\/p>\n<p>Note 0.50 + 0.30 = 0.80 - 768p draft first then 2K, and direct 2K\u00a0<strong>One price<\/strong>The official pricing did not leave any discount on the cloud draft. This is exactly where the local deployment will be<\/p>\n<ul>\n<li><strong>An iterative error, running draft local\u3002<\/strong>\u00a0A 5-second 768p walk API wants 2.5, local electricity charges are ignored, and a 50-version tip saves more than $100\u3002<\/li>\n<li><strong>The final draft is clear, two paths\u3002<\/strong>\u00a0ROUTE A:<strong>local overrated 1080p, free<\/strong>I don't know. The community already has a ComfyUI workflow that adjusts the parameters, pulling 768p in pieces with LTX 2.3 or Wan 2.2 5B -- Note that sigmas are selected to lower the sigmas version, and that too high a parameter can change the face, damage the mouth; some of the earlier versions are planted on it. Route B:<strong>API L 2K, \u00a50.30\/S<\/strong>15 seconds into a piece of \u00a54.5. It is not the same as normal overscores \u2014 recreated with the original context, the small words and details are not guessed, the value of the delivery-grade project. The ComfyUI contains a ready-made API template (starting with the three api_in the 3rd step table) and does not have to leave the table\u3002<\/li>\n<li><strong>N CARDS WITHOUT 12GB+, OR LESS THAN A FEW TIMES A YEAR<\/strong>, DO NOT MAKE 40 GB WEIGHTS FOR SEVERAL VIDEOS\u3002<\/li>\n<\/ul>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55652\" title=\"74a99204j00tjhbzw0084d000m800cip\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/74a99204j00tjhbzw0084d000m800cip.jpg\" alt=\"74a99204j00tjhbzw0084d000m800cip\" width=\"800\" height=\"450\" \/><\/p>\n<p>The local savings were mainly the biggest waste of the trial error phase; at the final stage, the 1080p was enough to save even the high-resolution money, and only the official 2K would pay 0.30\/second. Think of this position, a consumer-class card is not just a drafter, it's the main machine that makes a film\u3002<\/p>\n<p>In the end, four categories of people died on the route:<\/p>\n<ul>\n<li><strong>Just trying to play<\/strong>\u00a0\u2192 Step 0 page is enough, don't go down\u3002<\/li>\n<li><strong>PRODUCTION CAPACITY, 2K, DELIVERY<\/strong>\u00a0~ DIRECT API, ONE 5 SECONDS 2K 4 DOLLARS, DON'T MESS AROUND\u3002<\/li>\n<li><strong>N CARD WITH 12GB+, LARGE VOLUME, LOVE<\/strong>\u00a0\u2192 Local deployment + Mixing on top, this lesson is written for you\u3002<\/li>\n<li><strong>I want to fine-tune and study<\/strong>\u00a0\u2192 The full bf16 weight of AdaLN, which is also in 13B, is officially recommended to SGLang Doc, which is beyond the scope of this paper\u3002<\/li>\n<\/ul>\n<p>ATTACH: IF YOU DECIDE TO LEAVE API, THIS SCRIPT WILL BE COPIED<\/p>\n<p>Registered at platform.minimaxi.com (oversea platform.minimax.io), user centre full, create API key in account management. And then three steps: submit the mission \u2192 Rotation status \u2192 Download video:<\/p>\n<p>i'm sorry<\/p>\n<p>API_KEY = os.environ<br \/>\nBASE = \u201chttps:\/\/api.minimaxi.com\u201d # used overseas https:\/\/api.minimax.io<br \/>\nheaders = {\u201cAuthorization\u201d: f\u201d Bearer {API_KEY}}<\/p>\n<p># 1. SUBMISSION OF TASKS<br \/>\npayload ={<br \/>\n\"Model\": \"MiniMax-H3\",<br \/>\n\"content\":<br \/>\n{\u201cype\u201d: \u201ctext\u201d, \u201ctext\u201d: \u201cthe camera shoots an orange cat on the window,\u201d<br \/>\n\"It rained outside the window, and the cat's tail moved slowly, and the image was warm<br \/>\n\"with the sound of rain and the sound of a distant traffic.\" \u266a I'm sorry \u266a<br \/>\n],<br \/>\n\"duration\": integer of 5, # 4-15<br \/>\n\"resolution\": \"768P\", # Draft 768P, final 2K<br \/>\n\"ratio\": \"16:9\" , # plain-text video filled, cannot be written<br \/>\n}<br \/>\nr = reports.post(f) {BASE}\/v2\/video_generation, headers = heads, json = payload)<br \/>\nr.raise_for_status()<br \/>\ntask_id = r.json()[&#8220;task_id&#8221;]<\/p>\n<p># 2. \u8f6e\u8be2\u76f4\u5230\u5b8c\u6210<br \/>\nwhile True:<br \/>\ntime.sleep(10)<br \/>\nq = requests.get(f&#8221;{BASE}\/v2\/query\/video_generation\/{task_id}&#8221;, headers=headers).json()<br \/>\nstatus = q[&#8220;task&#8221;][&#8220;status&#8221;]<br \/>\nif status == &#8220;succeeded&#8221;:<br \/>\nurl = q[&#8220;task&#8221;][&#8220;content&#8221;][&#8220;url&#8221;]<br \/>\nbreak<br \/>\nif status in (&#8220;failed&#8221;, &#8220;canceled&#8221;):<br \/>\nrice SystemExit(q)<\/p>\n<p># 3. DOWNLOAD (LINKS ARE TIME-LIMITED, DO NOT HOARD)<br \/>\nopen(&#8220;out.mp4&#8221;, &#8220;wb&#8221;).write(requests.get(url).content)<\/p>\n<p>All you need to do is have a video\u00a0<strong>listen<\/strong>\u00a0Riga\u00a0<strong>{&#8220;type&#8221;: &#8220;image_url&#8221;, &#8220;image_url&#8221;: {&#8220;url&#8221;: &#8220;https:\/\/\u4f60\u7684\u56fe\u7247\u5730\u5740.png&#8221;}, &#8220;role&#8221;: &#8220;first_frame&#8221;}<\/strong>Add a frame to that\u00a0<strong>i'm sorry, ratio<\/strong>\u00a0Delete (which is automatically proportional to the picture)\u3002<\/p>\n<p>Four new walls. I'll tell you in advance:<\/p>\n<ol>\n<li><strong>duration<\/strong>\u00a0Only\u00a0<strong>Integer number from 4 to 15<\/strong>(note and local 17k+5 frame grid are two sets of rules)\u3002<\/li>\n<li>INDIVIDUAL LIMITATIONS ON MATERIAL: VIDEO 50MB, FIGURE 30MB, AUDIO 15MB<strong>REQUEST TOTAL 64MB<\/strong>Over and pass the URL\u3002<\/li>\n<li><strong>Audio cannot be entered as a separate input<\/strong>, must have a picture or a video\u3002<\/li>\n<li>_Other Organiser\u00a0<strong>7 days<\/strong>, the video link is time-barred and downloads\u3002<\/li>\n<\/ol>\n<p><em>(b) The measured environment: RTX 4090 48GB (magic)\/CommyUI 0.30.0\/torch 2.11 + cu128. Time-consuming, visible and reversible data are measured in the text; power weights are based on the Hugging Face official warehouse; the hint is based on the Mini Max official guide\u3002<\/em><\/p>","protected":false},"excerpt":{"rendered":"<p>The goal of this lesson: a computer with only a system and a graphic card, followed by an AI video with stereo, and learned the official hint -- H3 is good and bad, and most of it is on it. At each step, it was written \u201cwhat to do, how to do, what to see after\u201d, without any code basis. One thing you have to know before you start: the native canvas of the local model is 768p. The model was only trained on the short side of a canvas of 768, and the official 2K came out of another closed-source module. But 768p is not a local ceiling: it's a much bigger resolution<\/p>","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[149,144],"tags":[1490,862,2658],"collection":[],"class_list":["post-55633","post","type-post","status-publish","format-standard","hentry","category-jiaocheng","category-baike","tag-minimax","tag-862","tag-2658"],"acf":[],"_links":{"self":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/55633","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/comments?post=55633"}],"version-history":[{"count":0,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/55633\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/media?parent=55633"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/categories?post=55633"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/tags?post=55633"},{"taxonomy":"collection","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/collection?post=55633"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}