{"id":55801,"date":"2026-08-19T09:36:22","date_gmt":"2026-08-19T01:36:22","guid":{"rendered":"https:\/\/www.1ai.net\/?p=55801"},"modified":"2026-08-11T11:30:11","modified_gmt":"2026-08-11T03:30:11","slug":"comfyui-%e4%bb%8e%e5%ae%89%e8%a3%85%e5%85%a5%e9%97%a8%e5%88%b0%e7%b2%be%e9%80%9a%ef%bc%8ccontrolnet-%e4%b8%89%e5%a4%a7%e5%9c%ba%e6%99%af%ef%bc%88%e7%94%bb%e9%9d%a2%e5%90%ac%e4%bd%a0%e6%8c%87%e6%8c%a5","status":"publish","type":"post","link":"https:\/\/www.1ai.net\/en\/55801.html","title":{"rendered":"ComfyUI from installation to mastery, ControlNet, three scenes"},"content":{"rendered":"<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-55815\" title=\"85b62527j00tjl3hj00mrd000p800e4p\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/08\/85b62527j00tjl3hj00mrd000p800e4p.jpg\" alt=\"85b62527j00tjl3hj00mrd000p800e4p\" width=\"908\" height=\"508\" \/><\/p>\n<p>Title 09 <a href=\"https:\/\/www.1ai.net\/en\/55800.html\/\">ComfyUI from installation entrance to mastery, LoRA uses combat (selection\/down\/discharge\/string weight)<\/a>You've got a LoRA skill pack on the bottom model. But there's a question you'll probably hit sooner or later:<strong>I want it to follow a reference map<\/strong>Or have two people sitting side by side and running around without a hint\u3002<\/p>\n<p>I used to make a \"two handshake\" with three hundred words in the tip, either back-to-back or hand-in-hand, and I had to redraw a dozen. It was only later that ControlNet knew that \u201cthe image was in command\u201d and that it should not have been by the mouth, but by a map. And this hand-handed hand tells you what it is, where the model is, how the nodes are connected, how the three most used scenes fit, how the stength is screwed, and what the red-word error and the 8G mass collapse pits are\u3002<\/p>\n<p><strong>One, ControlNet. What is it<\/strong><\/p>\n<p>One sentence:<strong>ControlNet is a line for the bottom model, and you take a reference map and you use it as a rope, and the image is pressed on that image, and you don't run around\u3002<\/strong>Unlike LoRA -- LoRA is what the foundation model looks like, ControlNet is what it says\u3002<\/p>\n<p>Zenium\u00a0<strong>OpenPose<\/strong>Locked person position (bones), assigned movement, multiple positions<br \/>\nZenium\u00a0<strong>Depth<\/strong>: Lock space deep, perceiving, front- and back-level, building<br \/>\nZenium\u00a0<strong>Canny<\/strong>Locks the contour line, colours on the draft, make-up, rewinding<\/p>\n<p><strong>Remember:<\/strong>The LoRA tube's face, the ControlNet tube's position and structure, is called a genuine control map\u3002<\/p>\n<p><strong>II. Where to place the document (the wrong \/ model is not effective for the whole)<\/strong><\/p>\n<p>The controlNet model (.safetensors) shows this directory:<\/p>\n<p><a href=\"https:\/\/www.1ai.net\/en\/tag\/comfyui\" title=\"_Other Organiser\" target=\"_blank\" >ComfyUI<\/a>\/models\/controlnet\/<\/p>\n<p>Two pits blocked earlier:<br \/>\n1.\u00a0<strong>The model must have a base model<\/strong>\u2014ControlNet for SD 1.5 with SD 1.5 base model, SDXL with SDXL. Combining with LoRA will Blackchart \/ Invalid, page Base Model label will be accepted before\u3002<br \/>\n2.\u00a0<strong>Don't be too big<\/strong>\u2014 WITH 512, WITH 1024 4 TIMES TIME + 3 TIMES VISIBILITY, 8G CARD IS EASY TO OOM\u3002<\/p>\n<p><strong>Trail experience:<\/strong>For the first time, I was thinking of stacking three ControlNets, all of which were original 1024, and then the CUDA OOM, and it was stuck. Later on -- a smaller reference, a smaller CN, less folding, and an 8G card\u3002<\/p>\n<p><strong>III. Hand-to-hand nodes (continuous)<\/strong><\/p>\n<p>For example, OpenPose:<\/p>\n<p>Reference Diagram OpenPose Preprocessor<br \/>\nCheckpoint (MODEL)<br \/>\nPhrasing<\/p>\n<p>1.\u00a0<strong>Add Node<\/strong>: Right-key search for Apple ControlNet, ControlNet Loader, Load Image, and then OpenPose\u3002<br \/>\n2.\u00a0<strong>Select Model<\/strong>:ControlNet Loader contorl_net_name chooses .safetensors\u3002<br \/>\n3.\u00a0<strong>Connect<\/strong>: Reference Diagram \u2192Control for Loader; Model \u2192 Loader for CHeckpoint; CONTROL_NET \u2192 Apply ControlNet; Prostive \u2192Apply ControlNet's commanding, Apply Output KSampler\u3002<br \/>\n4.\u00a0<strong>setstrength<\/strong>: Apply ControlNet default 1.0 (lock to death), adjustable\u3002<\/p>\n<p>One sentence:<strong>Apply ControlNet, then KSampler, Reference Map Preprocessor fed ControlNet Loader\u3002<\/strong><\/p>\n<p><strong>Clean-up ring: 10 minutes, lock character positions with OpenPose (do now)<\/strong><\/p>\n<p>The principles of the previous sections are true. After you've got an extra \"Personal pose and reference map\" in your hand, you'll be able to confirm that ControlNet really works -- not just reading it, but making it\u3002<\/p>\n<p><strong>Preconditions<\/strong>: You've got the Load Checkpoint, two CLIP Text Encode, KSampler, VAE Decode, Save Image. If you lose it, press this five nodes again\u3002<br \/>\n<strong>Target<\/strong>: Take a profile of the person's position, and let the person who's created do the same thing, and don't move with the hints\u3002<\/p>\n<p>Step (sight, don't jump):<br \/>\n1.\u00a0<strong>Down Model<\/strong>OpenPose ControlNet (SD1.5 for control_v11p_sd15_openpose, SDXL for SDXL version), dropped into ComfyUI\/models\/controlnet\/, F5 to refresh\u3002<br \/>\n2.\u00a0<strong>Add Node<\/strong>: Double-click Load Image + OpenPose + ControlNet Loader + Apply ControlNet, one each\u3002<br \/>\n3.\u00a0<strong>Reference Diagram<\/strong>: IMAGE \u2192image for OpenPose preprocessor for Load Image; control for POSE \u2192 ControlNet Loader for preprocessor\u3002<br \/>\n4.\u00a0<strong>Main Chain<\/strong>: MODEL \u2192 ControlNet Loader of CHeckpoint; CONTROL_NET \u2192 Apply ControlNet _net of Loader; Control_net of your positive tipping word \u2192 Apply ControlNet = Conditioning, Apply CONDIIONG Output \u2192 KSampler\u3002<br \/>\n5.\u00a0<strong>Set parameters<\/strong>: Strength fill 1.0 for Apply ControlNet (locked in position), start_percent \/ end_percent for 0\/ 1.0\u3002<br \/>\n6.\u00a0<strong>Fixed Feed + Out<\/strong>Seed for KSampler has fixed numbers (e. g. 12345), point Queue Prompt\u3002<\/p>\n<p><strong>Checkpoint<\/strong>1 Generating a person's position and reference diagrams consistent, with a no-altering gesture = effective; 2 completely disordered = preprocessor failed or reference did not enter Loader (return section III); 3 Report \u201cExpected 1 channels but got 3 channels\u201d = reference diagrams directly linked to Apply ControlNet, without preprocessor (repeated pre-processing, see section VI); 4 Total unchanged = Apply ControlNet was not connected to the patent chain\u3002<\/p>\n<p>the job (chosen): to drop the stength from 1.0 to 0.6 and then run, the character becomes loose, the model becomes more active, and feels the rope is loose\u3002<\/p>\n<p><strong>IV. Three scenario formulations (OpenPose\/Dept\/Canny)<\/strong><\/p>\n<p>A ControlNet ripe, three together is normal. Multi-control method: The last Apply ControlNet CONDITIONG output goes to the next input (positive \u2192 Apply CN#1 \u2192 Apply CN#2 \u2192 KSampler). Here's the beginning of a fine-tuning exercise<\/p>\n<table>\n<tbody>\n<tr>\n<td>\n<section>Group<\/section>\n<\/td>\n<td>\n<section>sength formulation<\/section>\n<\/td>\n<td>\n<section>use<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>OpenPose + Decth<\/section>\n<\/td>\n<td>\n<section>0.85 \/ 0.5 ~ 0.65<\/section>\n<\/td>\n<td>\n<section>Lock position + Background level<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Canny+Deptth<\/section>\n<\/td>\n<td>\n<section>0.65 \/ 0.55<\/section>\n<\/td>\n<td>\n<section>contour +space, architecture \/ still<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Lineart + OpenPose<\/section>\n<\/td>\n<td>\n<section>0.7 \/ 0.8<\/section>\n<\/td>\n<td>\n<section>Binary + Action<\/section>\n<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p>end_percent techniques: 1.0 whole-step control; 0.7-0.85 front-end control diagrams, post-stage free-colouring of models (canny rewind permanent 0.7, OpenPose set 1.0)\u3002<\/p>\n<p><strong>Trail experience:<\/strong>I wanted the picture to be \"both in position and in depth\" and set it directly at 0.0 for Apply ControlNet, and the result was to create a completely ignored reference -- and later I knew that 0.0 was \" hung and uncontrollable\" and to control it was to give value greater than zero\u3002<\/p>\n<p><strong>V. Common pits \/ mined areas<\/strong><\/p>\n<p>1.\u00a0<strong>Canny will insert pre-treatment<\/strong>: Reference diagrams directly linked to ControlNet \"Expected 1 channels but got 3 channels\". The Canny Edge Test Node (low_threshold 100\/ high_threshold 200 applies 95% scenes) must be inserted between Load Image and Apply. OpenPose\/Depth has a preprocessor, too, except for\u3002<br \/>\n2.\u00a0<strong>settings 0.0<\/strong>\u00a0= Generates but completely ignores reference maps, often with errors\u3002<br \/>\n3.\u00a0<strong>Reference scale mismatch<\/strong>: USING 512 INSTEAD OF 1024, OR 4 TIMES TIME + 3 TIMES VISIBILITY, EASY OOM\u3002<br \/>\n4.\u00a0<strong>Base model does not match<\/strong>: CN with SDXL base model for SD1.5, black map for light and invalid weight. Allows the Base Model tag before the next\u3002<br \/>\n5.\u00a0<strong>The constraints are contradictory<\/strong>: Too many ControlNets and asked to fight, the scenes are falling apart. First, double, then single, then group\u3002<\/p>\n<p><strong>VI. QUIREMENTS OF PARLIAMENTS<\/strong><\/p>\n<table>\n<tbody>\n<tr>\n<td>\n<section>take<\/section>\n<\/td>\n<td>\n<section>Preprocessor<\/section>\n<\/td>\n<td>\n<section>startstrength<\/section>\n<\/td>\n<td>\n<section>Remark<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Lock position<\/section>\n<\/td>\n<td>\n<section>OpenPose<\/section>\n<\/td>\n<td>\n<section>1.0<\/section>\n<\/td>\n<td>\n<section>Lock the whole way<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Lock Space<\/section>\n<\/td>\n<td>\n<section>Depth<\/section>\n<\/td>\n<td>\n<section>0.5~0.65<\/section>\n<\/td>\n<td>\n<section>With OpenPose 0.85\/0.5<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Lock Outline<\/section>\n<\/td>\n<td>\n<section>Canny<\/section>\n<\/td>\n<td>\n<section>0.65<\/section>\n<\/td>\n<td>\n<section>change wind end set 0.7<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Binary wire<\/section>\n<\/td>\n<td>\n<section>Lineart<\/section>\n<\/td>\n<td>\n<section>0.7<\/section>\n<\/td>\n<td>\n<section>With OpenPose<\/section>\n<\/td>\n<\/tr>\n<tr>\n<td>\n<section>Double-control superimpose<\/section>\n<\/td>\n<td>\n<section>See above<\/section>\n<\/td>\n<td>\n<section>Come on<\/section>\n<\/td>\n<td>\n<section>One after two<\/section>\n<\/td>\n<\/tr>\n<\/tbody>\n<\/table>\n<p><strong>VII. 4060 \/ 8G SPECIALIZED &amp; I STEPPED ON PITS<\/strong><\/p>\n<p>Skills:<strong>Union SDXL is the 8G User Gospel<\/strong>- A 2.5-GB model from xinsir contains 10+CN types, which, in deduction, inserts the SetUnionControlNetType node type (canny\/dispth\/pose\/tile)<strong>Toggle type without reloading model, save memory + save time<\/strong>, the production environment is preferred. In the early days of SDXL, only Canny\/Dept, Union completely changed the situation\u3002<\/p>\n<p>More CN superclaves visible warning: each CN is an additional U-Net level reasoning, much more visible\u3002<strong>3 CN SUPERIMPOSED 16GB+ VRAM; 2 ARE DESSERT<\/strong>- We have up to two 8G users and suggest \u2013lowvram at startup. The more restraints, the easier they are to contradict each other, the more the picture collapses, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more they are, the more, the more they are the more they are\u3002<\/p>\n<p>Five pits say all at once:\u00a0<strong>Canny, no pre-treatment<\/strong>(No. 1 red-word pit) 2\u00a0<strong>settings 0.0<\/strong>\u00a0Equal to uncontrolled;3\u00a0<strong>Reference figure 1024<\/strong>\u00a0Visible explosive; 4\u00a0<strong>Base model does not match<\/strong>\u00a0Black Chart 5\u00a0<strong>3 CNS STACKED<\/strong>\u00a08G CARP. OOM\u3002<\/p>\n<p>ControlNet will work, and you've been able to put the picture in command -- position, space, contours. But there is a need for both:<strong>To lock a specific face and make it look the same in every episode\u3002<\/strong>It's gonna have to be a high-resolution zoom in on IPAdapter and back\u3002<strong>Page 11 Deep + High Zoom<\/strong>, add clarity and detail\u3002<\/p>","protected":false},"excerpt":{"rendered":"<p>Book 09 ComfyUI, from installation to mastery, LoRA uses the field (selection\/down\/display\/string weight) you have placed the LoRA skill pack on the bottom model, and the style, role and hand are steady. But there's one question that you're probably going to hit sooner or later: I want the picture to follow a reference, or I want two people standing one by the other and not running, and I don't know what the hint is. I used to make a \"two handshake\" with three hundred words in the tip, either back-to-back or hand-in-hand, and I had to redraw a dozen. It was only later that ControlNet knew that \u201cthe image was in command\u201d and that it should not have been by the mouth, but by a map. This hand-to-hand lesson<\/p>","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[149,144],"tags":[1989,6276,4749],"collection":[],"class_list":["post-55801","post","type-post","status-publish","format-standard","hentry","category-jiaocheng","category-baike","tag-comfyui"],"acf":[],"_links":{"self":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/55801","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/comments?post=55801"}],"version-history":[{"count":0,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/55801\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/media?parent=55801"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/categories?post=55801"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/tags?post=55801"},{"taxonomy":"collection","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/collection?post=55801"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}