{"id":57916,"date":"2026-10-11T10:22:11","date_gmt":"2026-10-11T02:22:11","guid":{"rendered":"https:\/\/www.1ai.net\/?p=57916"},"modified":"2026-09-20T14:58:14","modified_gmt":"2026-09-20T06:58:14","slug":"ai%e6%8f%90%e7%a4%ba%e8%af%8d%e5%88%9b%e4%bd%9c%e7%ac%ac%e4%ba%94%e5%8d%81%e4%b8%89%e8%8a%82%ef%bc%9a%e8%ae%a9ai%e4%ba%ba%e7%89%a9%e6%9b%b4%e5%83%8f%e7%9c%9f%e4%ba%ba%e7%9a%84","status":"publish","type":"post","link":"https:\/\/www.1ai.net\/en\/57916.html","title":{"rendered":"SECTION 53 OF THE A.I.D.: THE KEY TO MAKING A.I. MORE LIKE A REAL PERSON IS NOT A FACE, BUT SOMETHING TO DO"},"content":{"rendered":"<p>A LOT OF PEOPLE WHO DO AID VIDEOS FOCUS ON THEIR FACES: FIVE OFFICIALS ARE REAL, THEIR SKIN IS DELICATE, THEIR EYES ARE NATURAL, AND THEIR SHADOWS ARE CINEMATIC. BUT WHEN THE VIDEO CAME OUT, IT WAS EASY TO SEE IT WAS AI\u3002<\/p>\n<p>The reason is simple: video is not static. It's enough to look good for a moment, but the video needs the character to stay real for a few seconds. As long as people stand too steady, their faces remain unchanged and their movements are not justified, the picture will appear rigid\u3002<\/p>\n<p>Real people don't always park there like statues. People look at their mobile phones while waiting for a car, they turn cups at a coffee shop, they look at the floors in an elevator and they get windy when they walk. It's these subtle moves and life reactions that make people seem to be going through a real moment\u3002<\/p>\n<p>Let's talk about this<a href=\"https:\/\/www.1ai.net\/en\/tag\/ai%e8%a7%86%e9%a2%91\" title=\"[View articles tagged with [AI Video]]\" target=\"_blank\" >AI Video<\/a>The core skill of character tips: not just what the person looks like, but what the person is doing, why, and how the movement continues\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-57919\" title=\"ea0dd484jlzab00sjd000v9000d9p\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/09\/ea0dd484j00tlhzab00sjd000v900d9p.jpg\" alt=\"ea0dd484jlzab00sjd000v9000d9p\" width=\"1125\" height=\"477\" \/><\/p>\n<p><strong>I. WHY IS IT EASY TO FAKE AN AI VIDEO CHARACTER<\/strong><\/p>\n<p>Look at this\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-57917\" title=\"0344d20dj00tlzhzar00md000v9000d9p\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/09\/0344d20dj00tlhzar00mdd000v900d9p.jpg\" alt=\"0344d20dj00tlzhzar00md000v9000d9p\" width=\"1125\" height=\"477\" \/><\/p>\n<p>The characters, costumes, scenes are fine, but if you make a video of it, just write:<\/p>\n<p>An ancient woman stood in her room with a letter in her hand, slowly looking at the camera, and the sleeves moved gently\u3002<\/p>\n<p>IT'S DYNAMIC, BUT IT'S LIKE AI. 'CAUSE WHY DOES A PERSON HAVE A LETTER? WHY DOES HE LOOK AT THE CAMERA? IS SHE WAITING FOR THE NEWS OR JUST READING SOMETHING IMPORTANT? IF THERE'S NO REASON, IT'S JUST A SHOW\u3002<\/p>\n<p>IT'S NOT JUST HOW PEOPLE MOVE, IT'S WHY PEOPLE MOVE\u3002<\/p>\n<p>For example:<\/p>\n<p>The ancient woman stood in a quiet room with her hands and a letter she had just received. She used to look down at the letter, like reading half of a sudden stop, with her sub-finger feeling tight on the paper, and then slowly looking forward, with her expression turning from calm to mild anxiety. The cuffs move slowly, the camera moves, the picture is quiet and restrained, and there is an ancient film feeling\u3002<\/p>\n<p>In this way, the person is not simply \u201cstanding with the letter\u201d, but is going through a moment when she reads something, so she stops, squeezes the paper, heads up, changes in emotions\u3002<\/p>\n<p>SO, THE FOCUS OF THE AI VIDEO ALERT IS NOT JUST A SIMPLE DESCRIPTION OF THE DYNAMICS, BUT A CLEAR STATEMENT:<\/p>\n<p>Where the characters are, what they're doing, why they move like this, and how they react\u3002<\/p>\n<p>In the same picture, static looks at the beauty of the picture; when it's made on video, it's about whether the person is actually in a situation or not\u3002<\/p>\n<p><strong>II. The heart of the video message: action is motivated by behaviour<\/strong><\/p>\n<p>A lot of people write video alerts that directly describe the action:<\/p>\n<p>Women up\u3002<br \/>\nGirl back\u3002<br \/>\nBoys drink coffee\u3002<br \/>\nPeople walk slowly\u3002<br \/>\nHair floats with the wind\u3002<\/p>\n<p>They're not wrong, but they're just \"face of action.\" If there's no reason to move, it's easy for the video to become a setup\u3002<\/p>\n<p>Case<\/p>\n<p>And look at this\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-57918\" title=\"e77a4d77j00tlzbe00nxd000v9000d9p\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/09\/e77a4d77j00tlhzbe00nxd000v900d9p.jpg\" alt=\"e77a4d77j00tlzbe00nxd000v9000d9p\" width=\"1125\" height=\"477\" \/><\/p>\n<p>The person in the picture was standing in the ancient wind room with a fan in his hand, and the clothes, lights and scenes were complete. The static is already pretty good, but if it's a video, just write:<\/p>\n<p>The woman stands in the room, holds a fan, smiles slowly, shakes softly, and the camera moves slowly\u3002<\/p>\n<p>This hint is not without movement. People laugh, fans move, cameras push. But the question is: Why did she shake the fan? Why did she look forward? Was she waiting for someone, or did she just hear something outside the door<\/p>\n<p>If there's no reason, it's easy to act\u3002<\/p>\n<p>IT'S NOT JUST WHAT SHE DID, BUT WHY IT HAPPENED\u3002<\/p>\n<p>For example:<\/p>\n<p>The ancient woman stood in a quiet library with a fan in her hand, like someone waiting outside the door. She would have looked down at the scrolls on the table and slowly looked forward after hearing the light footsteps from outside the door. The fan in his hand had moved softly and then slowly stopped, and his expression had moved from calm to an expectation of restraint. It's a little twirl, it's a little twirl, it's a slow move, it's an ancient film feeling, it's quiet and natural\u3002<\/p>\n<p>In this way, the person is not simply \u201cstand with a fan\u201d, but is in a clear position: she is waiting for someone\u3002<br \/>\nSo she looks up, she stops the fan, and her face changes\u3002<\/p>\n<p>It's also a \u201chand-in-hand fan\u201d, which is used in a normal way to make a person move; a better way to write is to allow a person to move naturally because he hears a voice, waits for someone, sees a change\u3002<\/p>\n<p>When writing video alerts, you can ask more:<\/p>\n<p>What triggered this move<\/p>\n<p>People shake fan because they wait in peace\u3002<br \/>\nThe character stopped because he heard voices outside the door\u3002<br \/>\nThe character looked up because someone was close\u3002<br \/>\nThe person's face changed because of emotional disruptions\u3002<\/p>\n<p>When the action is motivated, the character in the video is no longer a pose, but a real moment\u3002<\/p>\n<p><strong>III. Single-person video alert formula<\/strong><\/p>\n<p>When a single person video is made, many people write only one action, such as \"head down\" \"eye up\" \"turning\" and \"doing\", but the person that is written in this way tends to be like a swing\u3002<\/p>\n<p>A better way to do that would be to break the phrase down into a complete process\u3002<br \/>\nThe single-person video can be applied to this formula:<\/p>\n<p>Personal identity + specific scene + current state + action cause + continuous action + microresponse + lens language<\/p>\n<p>We'll use this as a case\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-57920\" title=\"cb38ab0ej00tlzhzby00sdd000v9000hfp\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/09\/cb38ab0ej00tlhzby00sdd000v900hfp.jpg\" alt=\"cb38ab0ej00tlzhzby00sdd000v9000hfp\" width=\"1125\" height=\"627\" \/><\/p>\n<p>The image is of an ancient woman, a night garden, a lantern on which she sits alone in front of the scene of the crime on the porch, with a slight forwarding of her body and her eyes, and an entire person has a quiet and low feeling. It's a good picture for a video like \"Persons are Disgusting\"\u3002<\/p>\n<p>If only:<\/p>\n<p>An ancient woman sits in the courtyard, and her head is down, the camera is moving slowly, and she feels like an ancient film\u3002<\/p>\n<p>This is a picture, but it's still space. Because it only wrote the results, not the writing process. Why do people think? Did she just read something or wait for something? Did she react with her head down? None of this was explained\u3002<\/p>\n<p>Open it according to the formula, and it becomes clearer:<\/p>\n<p>Personal status: an ancient woman<\/p>\n<p>Specific scene: Lights bright under the courtyard at night<\/p>\n<p>Current state: sitting alone in front of a crime<\/p>\n<p>Reason for movement: It's like I'm thinking of something, and I'm feeling a little low<\/p>\n<p>Sequencing: lower hair, soft fingers to the table, then slow eyes to the side<\/p>\n<p>Slight reaction: paused eyes, light breathing, thin hair, quiet and restraining<\/p>\n<p>Scene language: slow advance, light view deep, night wind<\/p>\n<p>, which can be written in:<\/p>\n<p>An ancient woman sits alone under the porch of the night, with a light on her forehead. It was like she was thinking about something in her head, quietly down, with her fingers down at the table, and her body leaning forward. After a moment of pause, her eyes slowly moved to the side, with a slight loss in her mind. The character has a natural blinking and a mild breath, a twirl and a little restraint. The camera is moving slowly, it's shallow, it's ancient, it's quiet\u3002<\/p>\n<p>The advantage of writing in this way is that the person does not have one \u201clow head\u201d action, but a complete process of change:<\/p>\n<p>Let's go<br \/>\nAnd stop<br \/>\nI'm not sure what I'm talking about<br \/>\nFinally bring out emotions\u3002<\/p>\n<p>This is the most important part of the personal video alert:<\/p>\n<p>Do not write a single action, write about the person's current state, movement changes and micro-responses\u3002<\/p>\n<p>The same picture<br \/>\nIt'll be empty if it's just a \"women's heads down.\"<br \/>\nThe video would be more natural if it were clear: \"Why does she go out of her mind, how she moves and how she moves?\"\u3002<\/p>\n<p><strong>IV. Making people look like real people: add a weak life response<\/strong><\/p>\n<p>This is an ancient woman standing in the palace, dressed in fine clothes, sanctified scenes, who is not fit to do too much\u3002<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-57921\" title=\"67a1f597jlzcf00wed000v9000hfp\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/09\/67a1f597j00tlhzcf00wed000v900hfp.jpg\" alt=\"67a1f597jlzcf00wed000v9000hfp\" width=\"1125\" height=\"627\" \/><\/p>\n<p>If only:<\/p>\n<p>The ancient women stood in the palace, slowly turning their heads, their sleeves drifted, the lens pushed, and the ancient film sense\u3002<\/p>\n<p>This paragraph, although dynamic, remains easy to fake. Because the character is just following the hint and there's no real physical reaction\u3002<\/p>\n<p>In this sacrificial scene, the person needs no big action, but rather a light life response to make it real\u3002<\/p>\n<p>for example:<\/p>\n<p>She stood still in the palace, waiting for a call. Her eyes would have looked to the side, slid down after a moment of pause, and her fingers were slightly tightened in her sleeve. The candlelights were smooth, the hair was swayed, and the sleeves were softly up and down as they were breathing. She's not acting clearly, she's restraining, she's holding her heart down. The camera is moving slowly, it's quiet, it's ancient\u3002<\/p>\n<p>There's not much action in this hint, but it's more like real people\u3002<\/p>\n<p>Because real people don't even stand still. She'll breathe, her eyes will stop, her fingers will be fine, and she'll react lightly to emotional changes\u3002<\/p>\n<p>This chapter highlights:<\/p>\n<p>Realism in the video does not necessarily come from big moves, but from small reactions\u3002<\/p>\n<p>In particular, the more restrained the figures, the more detailed they will be:<\/p>\n<p>The eyes are paused<\/p>\n<p>Tighten your fingers<\/p>\n<p>The respiratory cuffs rise and fall<\/p>\n<p>Slight shake of hair<\/p>\n<p>Keep your lips down<\/p>\n<p>There's a slight change in the weight of the body<\/p>\n<p>The expression goes from calm to restraint<\/p>\n<p>These details do not destroy the image's sense of dignity, but they make people look like statues\u3002<\/p>\n<p>SO WHEN YOU WRITE AN AI VIDEO MESSAGE, YOU CAN ADD THE LAST SENTENCE:<\/p>\n<p>The character is not completely static, with natural blinking, slight breathing, short pauses in the eyes, small movements of the finger, and overall restraint in the movement of nature and not exaggerating performances\u3002<\/p>\n<p>Even if he was just standing there, he would be more like a real person in a situation\u3002<br \/>\n<strong><br \/>\nV. Multi-person video: one person moves, another reacts<\/strong><\/p>\n<p>Multi-person video is the easiest to fake because many people allow everyone to act at the same time\u3002<\/p>\n<p>For example, this map<\/p>\n<p><img loading=\"lazy\" decoding=\"async\" class=\"alignnone size-full wp-image-57922\" title=\"cfb423a5j00tlzzc00vd000v9000hfp\" src=\"https:\/\/www.1ai.net\/wp-content\/uploads\/2026\/09\/cfb423a5j00tlhzct00vvd000v900hfp.jpg\" alt=\"cfb423a5j00tlzzc00vd000v9000hfp\" width=\"1125\" height=\"627\" \/><\/p>\n<p>If only:<\/p>\n<p>The ancient women and men stood in the room and looked up slowly, the men approached her, the camera moved slowly, and the ancient film sense\u3002<\/p>\n<p>This paragraph, although interactive, is more superficial. Because why are two people close? Why are they looking? Are women nervous, hesitant, or are they holding back? If it's not clear, it's easy to frame\u3002<\/p>\n<p>The key to multi-person video is not to move as much as possible, but to act on one side and react on the other\u3002<\/p>\n<p>May be replaced by:<\/p>\n<p>In the glare of the wind, women were in red dress and stood in front of men. The man whispered a word, and his body was slightly near her. The woman did not reply immediately, but looked at him, with a brief pause in her eyes, and her lips squeezed, as if she were squeezing. The man stopped in front of her and did not continue to approach. The camera is slowly advancing, the light is deep, the picture is quiet and tense, the ancient film sense\u3002<\/p>\n<p>It's like this<\/p>\n<p>Men come in and talk<br \/>\nThe women heard a pause, looked up and reacted with restraint<br \/>\nMen stop again, and a quiet pull\u3002<\/p>\n<p>It's more like a real video than \"two people against each other.\"\u3002<\/p>\n<p>A multi-person tip can remember one thing:<\/p>\n<p>Don't let everybody do it, let one trigger an event, another react\u3002<\/p>\n<p>for example:<\/p>\n<p>One man comes near, the other one steps back\u3002<\/p>\n<p>One person speaks, the other silent for a moment\u3002<\/p>\n<p>One man reached out, the other one looked away\u3002<\/p>\n<p>One stopped and the other noticed emotional change\u3002<\/p>\n<p>One looked at the other, the other looked away and slowly looked back\u3002<\/p>\n<p>This type of interaction creates a real relationship between the characters rather than a stand-in\u3002<\/p>\n<p>So, instead of just saying, \"What are they doing\" when you write a multi-person video, you have to be clear:<\/p>\n<p>Whoever acts first, who reacts, how light it is and what changes have taken place in the atmosphere\u3002<\/p>\n<p><strong>wind up<\/strong><\/p>\n<p>AS AN AI VIDEO CHARACTER, DON'T JUST THINK ABOUT GETTING PEOPLE MOVING\u3002<\/p>\n<p>The fact that people look back, blink, walk doesn't mean the picture is real. What's really important is that there's no reason, no change, no micro-responses\u3002<\/p>\n<p>A person with a letter can just stand and swing; he or she can read something important, pause, squeeze paper, and look slowly\u3002<\/p>\n<p>The difference is:<br \/>\nThe former is acting, the latter is going through things\u3002<\/p>\n<p>SO WHEN YOU WRITE AN AI VIDEO MESSAGE, REMEMBER ONE SENTENCE:<\/p>\n<p>It's not a simple description of dynamics, it's a way for people to move because of something\u3002<\/p>","protected":false},"excerpt":{"rendered":"<p>A LOT OF PEOPLE WHO DO AID VIDEOS FOCUS ON THEIR FACES: FIVE OFFICIALS ARE REAL, THEIR SKIN IS DELICATE, THEIR EYES ARE NATURAL, AND THEIR SHADOWS ARE CINEMATIC. BUT WHEN THE VIDEO CAME OUT, IT WAS EASY TO SEE IT WAS AI. THE REASON IS SIMPLE: VIDEO IS NOT STATIC. IT'S ENOUGH TO LOOK GOOD FOR A MOMENT, BUT THE VIDEO NEEDS THE CHARACTER TO STAY REAL FOR A FEW SECONDS. AS LONG AS PEOPLE STAND TOO STEADY, THEIR FACES REMAIN UNCHANGED AND THEIR MOVEMENTS ARE NOT JUSTIFIED, THE PICTURE WILL APPEAR RIGID. REAL PEOPLE DON'T ALWAYS PARK THERE LIKE STATUES. PEOPLE LOOK AT THEIR MOBILE PHONES WHILE WAITING FOR A CAR, THEY TURN CUPS AT A COFFEE SHOP, THEY LOOK AT THE FLOORS IN AN ELEVATOR AND THEY GET WINDY WHEN THEY WALK. IT'S THESE SUBTLE MOVES AND LIFE REACTIONS THAT MAKE PEOPLE SEEM TO BE GOING THROUGH A REAL MOMENT. IN THIS SECTION, LET'S TALK ABOUT THE CORE OF THE AVP<\/p>","protected":false},"author":1,"featured_media":0,"comment_status":"closed","ping_status":"closed","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[149,144],"tags":[802,3149,956,5321,192,6481],"collection":[],"class_list":{"0":"post-57916","1":"post","2":"type-post","3":"status-publish","4":"format-standard","5":"hentry","6":"category-jiaocheng","7":"category-baike","8":"tag-ai","12":"tag-prompt","13":"tag-6481"},"acf":[],"_links":{"self":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/57916","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/comments?post=57916"}],"version-history":[{"count":0,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/posts\/57916\/revisions"}],"wp:attachment":[{"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/media?parent=57916"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/categories?post=57916"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/tags?post=57916"},{"taxonomy":"collection","embeddable":true,"href":"https:\/\/www.1ai.net\/en\/wp-json\/wp\/v2\/collection?post=57916"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}