MACHINE ASR ACCESSIBILITY AID

Making_Videos_with_AI_Audiobook.mp3

Not an editorially verified transcript. This text was generated automatically from the preserved recording and may contain recognition, language-detection, spelling, segmentation or name errors. Consult the source recording for authoritative content.
Collection
Part 9 · VIDEHA MITHILA MAITHILI DISCUSSION CRITICISM SERIES PART 9
Status
asr-draft
Human verified
No
Editorial review
not-reviewed
ASR model
small
Detected language
en (0.997887)
Duration
55:05
Source
Open preserved recording

Timestamped machine output

  1. 0:00–0:01Making videos with AI.
  2. 0:01–0:06Teach yourself series, written by Gage and Rathakar.
  3. 0:06–0:08Preface.
  4. 0:08–0:13This book is for everyone who wants to say something through video but has no camera, no studio,
  5. 0:13–0:15and no training in editing.
  6. 0:15–0:20Artificial intelligence AI has now made all three possible with an ordinary mobile phone or computer.
  7. 0:20–0:24What once required equipment worth hundreds of thousands and a team of dozens can now
  8. 0:24–0:28be done with a written instruction, called a prompt.
  9. 0:28–0:30but let one thing be clear at the very outset.
  10. 0:30–0:32AI is no magic wand.
  11. 0:32–0:34It is a tool, like an axe, like a pen.
  12. 0:34–0:38The axe cuts the wood, but which tree to fell and what house to build.
  13. 0:38–0:40The carpenter decides, in the same way.
  14. 0:40–0:45AI will make the visuals and generate the voice, but what to say, why to say it,
  15. 0:45–0:49and for whom, the answers to these three questions remain with you.
  16. 0:49–0:53This book will teach you to handle the tool, the craftsmanship will remain your own.
  17. 0:53–0:59This book is written in the teacher-self style, that is, no teacher or training institute is
  18. 0:59–1:04needed. Each chapter is one step, begin with the first chapter, move forward in order,
  19. 1:04–1:09and be sure to do the exercise give at the end of every chapter. Reading and doing must walk
  20. 1:09–1:14together, only then does learning happen. One more thing, all the illustrations in this
  21. 1:14–1:20book are original explanatory figures, not copies of any company's actual screens. There are
  22. 1:20–1:26For two reasons. First, the screens of AI tools change every few months, so a real screenshot
  23. 1:26–1:28would be outdated before the book left the press.
  24. 1:28–1:33Second, once you understand the principle, any new screen will feel familiar. If you
  25. 1:33–1:37memorize a screen, every change will leave you stranded. The figures in this book show
  26. 1:37–1:42the structure that applies to nearly every tool alike. Where the prompt box sits, what
  27. 1:42–1:46the settings contain, what a timeline looks like.
  28. 1:46–1:50Once you grasp this framework, you will open any new tool and work it out on your
  29. 1:50–1:55own, and that is the true meaning of teach yourself. The absence of technical books in
  30. 1:55–1:59Mathalie has long been a sore point, and this book was first written in Mathalie. This English
  31. 1:59–2:05edition carries the same conviction. Whatever you learn here, use it to put your own language,
  32. 2:05–2:11your own region, and your own stories on screen. The tales, psalms, paintings, history and
  33. 2:11–2:17festivals of Mathila, and of every homeland like it, are all waiting for their videos.
  34. 2:17–2:21Number 1. What is AI video?
  35. 2:21–2:26Artificial Intelligence. AI is the capacity of a computer to do human-like work. Understanding
  36. 2:26–2:33language, recognizing images, and now, creating images and video. When we write to an AI,
  37. 2:33–2:37show two children playing by a pond, and it produces a moving picture of that very
  38. 2:37–2:42scene on its own. This is called generative AI. To generate means to bring forth. This
  39. 2:42–2:45This AI does not fetch a video from somewhere.
  40. 2:45–2:49It composes a new video that never existed anywhere before.
  41. 2:49–2:51How is this possible?
  42. 2:51–2:52Understand it in simple terms.
  43. 2:52–2:58These AI models have been shown crores of images and videos, having seen so much.
  44. 2:58–3:02They have learned what an egret looks like, how water ripples, what color the morning
  45. 3:02–3:05light takes, and how a walking person's feet rise and fall.
  46. 3:05–3:09When you write an egret beside a pond at dawn, the model builds such a scene from
  47. 3:09–3:13its learned knowledge, just as a painter, who has seen thousands of ponds can now draw a
  48. 3:13–3:17pond from memory without one before their eyes.
  49. 3:17–3:20There are chiefly four routes to making video with AI.
  50. 3:20–3:25First, text to video, you only write, and the AI creates the entire scene.
  51. 3:25–3:30This is the most astonishing route, but it offers somewhat less control, the scene that
  52. 3:30–3:33arrives may not match your imagination exactly.
  53. 3:33–3:39Second, image of video, you supply one still image, a photograph or an AI generated picture.
  54. 3:39–3:44and the AI sets it in motion, here the control is greater because you yourself chose the
  55. 3:44–3:46opening frame.
  56. 3:46–3:48Third, avatar video.
  57. 3:48–3:53A digital speaker avatar sits on screen and speaks your written text, like a news anchor,
  58. 3:53–3:57for lessons, announcements and lectures this is extremely useful.
  59. 3:57–3:59Fourth, AI assisted editing.
  60. 3:59–4:04Here the footage is your own recording, but the cutting and joining, the subtitles,
  61. 4:04–4:06the background removal.
  62. 4:06–4:08All this the AI does.
  63. 4:08–4:12This book covers all four routes, because in practice a good video is usually a blend of
  64. 4:12–4:14them.
  65. 4:14–4:19Now understand what AI cannot do, AI does not yet produce long videos in one go, it typically
  66. 4:19–4:24makes short clips of 5 to 20 seconds, which are joined to build a longer video, AI may
  67. 4:24–4:30see sometimes contain errors, a hand grows too many or too few fingers, written letters
  68. 4:30–4:34come out garbled, a character's face changes between one clip and the next, and the
  69. 4:34–4:40biggest point of all, AI knows nothing of Mathila, of Mepheli, or of what lies in your heart,
  70. 4:40–4:45it will give only as much as you know how to ask, that is why half of the spoke is devoted
  71. 4:45–4:48to how to ask, that is, planning and prompts.
  72. 4:48–4:54Exercise, on paper, write down three subjects for videos you would like to make, for each,
  73. 4:54–5:00write one line, who will watch this video and what will they gain from watching it.
  74. 5:00–5:03Chapter 2, Getting Ready, What Do You Need?
  75. 5:03–5:08No costly machinery is needed to make video with AI, all that is required is this.
  76. 5:08–5:15First, a device, a smartphone is sufficient, a computer or laptop adds convenience, writing
  77. 5:15–5:20prompts, managing files and editing are easier on a large screen, no specially powerful computer
  78. 5:20–5:25is needed because the video is made not on your device but on the company servers, your
  79. 5:25–5:29device merely sends the instruction and downloads the finished video.
  80. 5:29–5:35Second, the Internet, the better the speed, the shorter the weight, video files are large,
  81. 5:35–5:37so downloading will consume data.
  82. 5:37–5:39Budget a few gigabytes a month.
  83. 5:39–5:44Third, an email account, nearly every AI tool requires an account, and most allow direct
  84. 5:44–5:48sign-in with a Google email Gmail, one suggestion.
  85. 5:48–5:51Keep a separate email for this work, so that the newsletters of AI tools do not
  86. 5:51–5:53flood your main inbox.
  87. 5:53–5:56Fourth, an understanding of credits.
  88. 5:56–6:01First AI video tools run on a credit system, a credit is a kind of coupon.
  89. 6:01–6:06Making one video deducts some credits, a free account receives a small allowance, daily
  90. 6:06–6:09and some tools monthly and others, once only in a few.
  91. 6:09–6:16Paid plans bring more credits, higher quality such as 1080p or 4K, videos without watermarks
  92. 6:16–6:18and longer clips.
  93. 6:18–6:22Adapt a practical policy here, which I call cheap first, then deep, while learning
  94. 6:22–6:27a new tool, run small experiments on the free credits, learn to write prompts, learn the
  95. 6:27–6:33tools temperament, when it feels time for serious work, say, running a YouTube channel, then
  96. 6:33–6:39take a paid plan on one tool, buying plans on every tool is wasteful, one or two suffice.
  97. 6:39–6:445th, a system for files, it sounds a small matter, but the experienced know how work
  98. 6:44–6:48drowns without order, make a separate folder for each video project on your computer
  99. 6:48–6:55or phone, inside it keep four subfolders, script scripts and prompts, clips raw AI made video,
  100. 6:55–7:00voice music voiceover and music, and final the finished video, give every file a meaningful
  101. 7:00–7:05name, a file called clip a one-pond on will still be recognizable six months later, video
  102. 7:05–7:08three final new two will not.
  103. 7:08–7:14And finally, patience, AI tools are sometimes busy, sometimes give strange results, sometimes
  104. 7:14–7:18fail entirely to understand your perfectly good prompt, all this is natural, those
  105. 7:18–7:23Those who persist through three or four attempts learn, those who quit at the first odd result
  106. 7:23–7:25are left behind.
  107. 7:25–7:30Exercise, create the folder system described above on your device, make a new email account
  108. 7:30–7:35if you wish, and open the website of any one AI video tool simply to see what the free
  109. 7:35–7:40plan offers, do not make anything yet, only look.
  110. 7:40–7:44Chapter 3 – Planning the video, from idea to script
  111. 7:44–7:47The biggest mistake beginners make is to open the tool straight away and start
  112. 7:47–7:53pressing buttons, a video made without a plan looks exactly like a house built without a drawing.
  113. 7:53–7:58That is why this chapter comes before the tools. Before every video, write the answers to three
  114. 7:58–8:05questions. 1. What is the purpose? To teach say how Matubati painting is made, to tell a story,
  115. 8:05–8:11to show scenes of a village or to an else. 1. Video, 1 purpose. Hold to this rule. A video
  116. 8:11–8:17that contains everything contains nothing. 2. Who is the audience? Children, students,
  117. 8:17–8:23expatriates far from home, curious outsiders, or the general viewer, the audience decides
  118. 8:23–8:27how simple the language should be, how long the video should run, and what the visuals
  119. 8:27–8:32should look like, bright colors and quick movement for children, stillness and gravity
  120. 8:32–8:38for adults. Free, how long? In the beginning, make videos of 30 seconds to 2 minutes.
  121. 8:38–8:43A short video is easier to make, easier to fix and better liked by today's viewer.
  122. 8:43–8:50Remember, AI produces clips of 5-10 seconds, so a 1-minute video means 6-12 clips.
  123. 8:50–8:51Now the script.
  124. 8:51–8:54A script is nothing complicated or literary.
  125. 8:54–8:55It is simply a two-column table.
  126. 8:55–8:58In the left column, what is seen the visual.
  127. 8:58–9:01In the right column, what is heard the voice or subtitle.
  128. 9:01–9:05For a 30-second video, 5 or 6 rows are enough.
  129. 9:05–9:11I can help in a second role as script assistant give your subject to a chat bot such as Claude chat
  130. 9:11–9:17GPT or Gemini and ask it to draft the script but the asking needs skill merely saying write a video
  131. 9:17–9:24script on Madhubani will fetch a generic lifeless text ask like this I am making a 60 second
  132. 9:24–9:29YouTube video on Madhubani painting audience young Indian viewers who have heard the name but no
  133. 9:29–9:36a little more. Voice, simple and warm, I need a script in two columns, on the left, a description
  134. 9:36–9:41of each visual which I will give to an AI video tool, on the right, the voice of her
  135. 9:41–9:47text. Six scenes, 10 seconds each. Notice, subject, duration, audience, language,
  136. 9:47–9:52format, and use are all stated. The clearer the ask, the better the result. This same
  137. 9:52–9:57principle will apply later to video prompts. To not accept the chatbot's script with
  138. 9:57–10:00with your eyes closed, read it, test it against your own knowledge.
  139. 10:00–10:01Are the facts right?
  140. 10:01–10:04Does the language sound like your own?
  141. 10:04–10:08Change any line that falls flat, or have it rewritten make the third scene more tender?
  142. 10:08–10:13Shorten the last line, the script is your signature, the AI is only a scribe.
  143. 10:13–10:15Exercise.
  144. 10:15–10:19Take one of the three subjects you chose in chapter 1, write the answers to the three
  145. 10:19–10:22questions purpose, audience, duration.
  146. 10:22–10:28Then, using the detailed style of asking shown above, have a chatbot draft a two-column script
  147. 10:28–10:32and make at least two corrections to it with your own hand.
  148. 10:32–10:33Chapter 4.
  149. 10:33–10:34The Storyboard.
  150. 10:34–10:36A Map of Scenes.
  151. 10:36–10:41The script is a map of words, the storyboard is a map of scenes, a storyboard lays out
  152. 10:41–10:46one sketch, a rough drawing or description, for every scene of the video in order.
  153. 10:46–10:50In film making this method is a century old, and in the AI age its importance has
  154. 10:50–10:51not shrunk but grown.
  155. 10:51–10:52Why?
  156. 10:52–10:57Because AI needs a separate, clear instruction for every clip, and the storyboard is precisely
  157. 10:57–10:59that list of instructions.
  158. 10:59–11:05Do not worry, no drawing skill is required, round faces, stick figures, and arrows are
  159. 11:05–11:09enough, if you would rather not sketch at all, write three lines for each scene, what
  160. 11:09–11:12is seen, what the camera does, and how many seconds.
  161. 11:12–11:16Learn a little of the camera's language for these very words will serve you later
  162. 11:16–11:18in prompts.
  163. 11:18–11:23Wide shot, the whole scene from a distance, for establishing the place, like a full view
  164. 11:23–11:24of the village.
  165. 11:24–11:31Close up, from near, for feeling and fine detail, like the steam over a cup of tea on the hearth.
  166. 11:31–11:35Tracking shot, the carer moves along with the subject, like following a child on the
  167. 11:35–11:37way to school.
  168. 11:37–11:42Drone view, from above as a bird sees, like a sweeping view of pond and fields.
  169. 11:42–11:47slash out, the carousel slowly draws near or pulls away, for emotional weight.
  170. 11:48–11:52Static shot, the camera stays in one place, for calm, settled scenes.
  171. 11:53–11:58Keep one simple formula for the order of scenes, establish, develop, resolve.
  172. 11:58–12:02The first scene tells where we'd are established, usually a wide shot.
  173. 12:02–12:08The middle scenes bring the subject close develop, close ups, tracking, the last scene gathers at
  174. 12:08–12:12at all, a feeling, a message or title card resolve.
  175. 12:12–12:16Figure 2 shows the storyboard of a 32nd video called My Village.
  176. 12:16–12:19See how six scenes travel from dawn to dusk.
  177. 12:19–12:24Morning mist establish, tea, school and pond develop, lamps and title resolve.
  178. 12:24–12:28The side each scene its duration and camera note are written, each of these scenes will,
  179. 12:28–12:31further on, become one prompt.
  180. 12:31–12:33While making the storyboard, attend to continuity.
  181. 12:33–12:37If the first scene is morning, the second must not suddenly be night.
  182. 12:37–12:42a character wears a red kurta. The red kurta must appear in every scene, and this must
  183. 12:42–12:47be written into every prompt, because AI does not remember the previous clip.
  184. 12:47–12:51Write the character's description on a separate sheet, a character card, and paste it word
  185. 12:51–12:55for word into every prompt. This small trick is the simplest way to keep
  186. 12:55–12:57the clips consistent.
  187. 12:57–13:03Exercise. Turn your script into a 16 storyboard. For each scene, write the visual description,
  188. 13:03–13:08The camera, the duration, if there is a character, make a character card dress, age, appearance,
  189. 13:08–13:11in three lines.
  190. 13:11–13:14Chapter 5 The Art of Prompt Writing
  191. 13:14–13:17A prompt is the written instruction you give to the AI.
  192. 13:17–13:22It is the most valuable skill of the AI age, and happy news is that it demands no technical
  193. 13:22–13:23knowledge.
  194. 13:23–13:27Only clarity of language and language is our home ground.
  195. 13:27–13:30See the difference between a poor prompt and a good one.
  196. 13:30–13:31Poor.
  197. 13:31–13:36This tells that AI nothing, a village of which country, which season, day or night, what is
  198. 13:36–13:42happening, it will invent something from its own mind, most likely some European or placeless
  199. 13:42–13:48village, now the good one, a village in North India, early morning, light mist over the fields,
  200. 13:48–13:53thatched and tiled houses, a banyan tree in the distance, camera panning slowly to
  201. 13:53–14:00the right, soft golden light, realistic documentary style, now the AI holds the complete picture.
  202. 14:00–14:03Keep in mind the six-part formula shown in Figure 3.
  203. 14:03–14:07Subdecked, who or what is central, two egrets.
  204. 14:07–14:09Action, what are they doing?
  205. 14:09–14:11Catching fish.
  206. 14:11–14:12Seen.
  207. 14:12–14:15Where and when, a pond full of lotus, et don.
  208. 14:15–14:18Camera, what does the camera do?
  209. 14:18–14:19Slowly moving in.
  210. 14:19–14:21Light, what kind of light?
  211. 14:21–14:23Golden morning light.
  212. 14:23–14:24Style, what look?
  213. 14:24–14:27Documentary, realistic.
  214. 14:27–14:31Let every prompt carry all six parts and keep roughly this order.
  215. 14:31–14:33Subject and action first.
  216. 14:33–14:34Style last.
  217. 14:34–14:35Some useful words of style.
  218. 14:35–14:36Realistic.
  219. 14:36–14:37Cinematic.
  220. 14:37–14:38Animation.
  221. 14:38–14:39Like a watercolor painting.
  222. 14:39–14:40Like old film.
  223. 14:40–14:41Documentary.
  224. 14:41–14:44The question of language.
  225. 14:44–14:47Most AI video tools understand English prompts best.
  226. 14:47–14:50Some also follow Hindi and other languages.
  227. 14:50–14:51The practical path is this.
  228. 14:51–14:56They can script in your own language and shape the final prompt in English.
  229. 14:56–15:00a chatbot helps, give it your scene description in your language and say, turn this into an
  230. 15:00–15:07English video prompt with six-part subject, action, scene, camera, light, style, thus the
  231. 15:07–15:12imagination stays yours, only the translation is mechanical.
  232. 15:12–15:16There is also the practice of the negative prompt, where you state what you do not want,
  233. 15:16–15:22for example, no blur, no distorted hands, no text on screen, no watermark, some
  234. 15:22–15:26Some tools give it a separate box, in others it is added to the main prompt.
  235. 15:26–15:31Now the most important principle of all, the improvements cycle figure 4.
  236. 15:31–15:35The first prompt rarely yields the video of your wishes, and this is no failure, it is
  237. 15:35–15:37the method itself.
  238. 15:37–15:41Look at the result, name the fault, change just that much in the prompt, and generate
  239. 15:41–15:42again.
  240. 15:42–15:44Did the egret come out too large?
  241. 15:44–15:46At a small egret, does the scene look garish?
  242. 15:46–15:52Change to soft, gentle light, usually within 3-5 cycles a usable clip arrives.
  243. 15:52–15:54Change only one or two things per cycle.
  244. 15:54–15:58Change everything at once and you will never know which change did the work.
  245. 15:58–16:01Keep saving your good prompts in one file.
  246. 16:01–16:04The you of six months hence will thank the you of today.
  247. 16:04–16:05Exercise.
  248. 16:05–16:09For the first scene of your storyboard, write one complete prompt using the six-part formula
  249. 16:09–16:14in your own language, then make its English form yourself or through a chatbot.
  250. 16:14–16:17Keep both in your script folder.
  251. 16:17–16:18Chapter 6.
  252. 16:18–16:19Text video tools.
  253. 16:19–16:21The first clip.
  254. 16:21–16:23The moment has come to take the tool in hand.
  255. 16:23–16:28In this chapter we understand the common structure of text-to-video tools and make the first clip.
  256. 16:28–16:34First, an introduction to the tools, this field changes at great speed, every few months a
  257. 16:34–16:37new model arrives and the old ones grow stronger.
  258. 16:37–16:43At the time of writing 2026 the leading names are, Google's BO, OpenAI's Sora, Kling,
  259. 16:43–16:48One Way, Pixverse, Seedance, Haleuo, Luma, and Pika.
  260. 16:48–16:53Each has its own temperament, one excels at realistic scenes, another at stylized or artistic
  261. 16:53–16:59ones, another is faster and cheaper, the names will keep changing, but the structure, described
  262. 16:59–17:05below, remains nearly the same in all of them, so learn the structure, not the names.
  263. 17:05–17:09Look at figure 5, in almost every tool you will find these 5 things.
  264. 17:09–17:101.
  265. 17:10–17:11The prompt box.
  266. 17:11–17:15A large empty field where you write your instruction, this is the heart of the
  267. 17:15–17:16tool.
  268. 17:16–17:172.
  269. 17:17–17:24selection. A single company offers several models, new and powerful costly, older or fast cheap,
  270. 17:24–17:29choose the cheap model while learning. The good model for final, publishable clips.
  271. 17:30–17:383. Aspect Ratio, 16,9 for YouTube and television-wide, 9,16 for Reels, Shorts, and Status Upright,
  272. 17:38–17:431,1 square, the side at the outset where the video will go, because changing the ratio
  273. 17:43–17:45leader crops the scene.
  274. 17:45–17:464.
  275. 17:46–17:47Duration.
  276. 17:47–17:50Usually options of 5, 8 or 10 seconds.
  277. 17:50–17:54A shorter clip costs fewer credits and carries fewer errors.
  278. 17:54–17:555.
  279. 17:55–17:58The generate button and the results area.
  280. 17:58–18:02Press the button, wait from a few seconds to a few minutes and the clip appears in
  281. 18:02–18:03the results area.
  282. 18:03–18:05Download it from there.
  283. 18:05–18:08Now the method for the first clip, step by step.
  284. 18:08–18:091.
  285. 18:09–18:12Open the tools website and create an account with your email.
  286. 18:12–18:132.
  287. 18:13–18:16free credits, how many, and when they renew.
  288. 18:16–18:173.
  289. 18:17–18:22Choose the cheap slash fast model, 16 colon 9 ratio, and the shortest duration.
  290. 18:22–18:234.
  291. 18:23–18:26Pace the prompt you built in Chapter 5.
  292. 18:26–18:275.
  293. 18:27–18:28Press Generate and Wait.
  294. 18:28–18:296.
  295. 18:29–18:34Watch the finished clip in full, not once, but two or three times, watch the hands,
  296. 18:34–18:37the faces, any lettering, the way things move.
  297. 18:37–18:387.
  298. 18:38–18:42Download it into your clip's folder under a meaningful name, even if the clip is
  299. 18:42–18:48perfect, it will serve for comparison. 8. Run the improvement cycle. Name the fault,
  300. 18:48–18:54refine the prompt, generate again. A few practical tricks. Generate two or three clips from the
  301. 18:54–19:00same prompt. AI gives a somewhat different result each time, and you pick the best. A clip whose
  302. 19:00–19:05main subject is right but whose edges carry faults can often be saved by cropping in the edit,
  303. 19:05–19:10do not discard it at once, and set yourself a daily usage limit before the credits run dry,
  304. 19:10–19:16or weeks credits will vanish in one enthusiastic evening. Exercise, make the clip for the first
  305. 19:16–19:21scene of your storyboard, running at least three improvement cycles, save the final clip together
  306. 19:21–19:28with the prompt that produced it. Chapter 7, From Enage to Video, The Road of Greater Control
  307. 19:29–19:34Text to video has won in convenience, you have no hold over the opening scene. Whatever the AI
  308. 19:34–19:40makes, it makes, the remedy is image of video, first prepare a still image that is exactly to
  309. 19:40–19:44to your mind, then tell the AI to set this image in motion.
  310. 19:44–19:49Hear the look of the scene, the character's face, the clothing, all are fixed in advance.
  311. 19:49–19:51The AI only adds the movement.
  312. 19:51–19:53Where will the image come from?
  313. 19:53–19:54Free sources.
  314. 19:54–20:00First, your own photographs, your village, your festivals, nature, art, animating a
  315. 20:00–20:04photograph you took yourself is the most authentic or out.
  316. 20:04–20:07Remember the photograph must be your own or used with the owner's permission, and
  317. 20:07–20:12before animating a photograph of a living person, be sure to take their consent. This is both
  318. 20:12–20:17courtesy and part of the ethics, described in chapter 15.
  319. 20:17–20:22Second. AI generated images. There are separate tools for image generation and most video
  320. 20:22–20:28tools include an image making feature. The image prompt follows the same six-part formula,
  321. 20:28–20:33only in place of camera movement, describe the composition. Images are cheap and quick
  322. 20:33–20:38to make. So run your improvement cycle on the image first. Generating 10 images and choosing
  323. 20:38–20:42the best is far cheaper than generating 10 videos.
  324. 20:42–20:46Third, scanning your own artwork. A work of Mathila painting if it is your own or you
  325. 20:46–20:51hold the rights can be scanned and given gentle motion. A fish stirring, the line ornament
  326. 20:51–20:56shimmering, exercise great restraint here, the dignity of traditional art lies in its
  327. 20:56–21:01stillness, so keep the motion extremely slight, right gentle, slow, subtle motion in the
  328. 21:01–21:07prompt. The method is simple. Choose the image video option in the tool, upload your image,
  329. 21:07–21:12and write the motion instructional on side. This prompt now carries less seen description and
  330. 21:12–21:18more motion description, what moves, in which direction, how fast, and what the camera does.
  331. 21:18–21:23For example, gentle ripples on the water, the egrets wings moving slowly, camera moving forward
  332. 21:23–21:30very slowly, nothing else changes. That last phrase, nothing else changes, matters, without it the
  333. 21:30–21:35AI will sometimes transform the whole scene. This is also the best remedy for the problem of character
  334. 21:35–21:41consistency. If the same character appears again and again in your video, first create or choose one
  335. 21:41–21:46excellent image of that character. And for every scene animate that same image with different
  336. 21:46–21:51motion instructions, some advanced tools offer a feature called character reference or consistent
  337. 21:51–21:56character where the character's image, given once, appears in every clip. Look for this
  338. 21:56–21:58feature in your tool.
  339. 21:58–21:59Exercise.
  340. 21:59–22:04Take any image or own photograph or AI generated and make three different motion versions of
  341. 22:04–22:05it.
  342. 22:05–22:11In one, only the camera moves, in the second, only some element of the scene moves, in
  343. 22:11–22:16the third, both, compare the three, which feels most natural.
  344. 22:16–22:17Chapter 8.
  345. 22:17–22:18Voice.
  346. 22:18–22:20Voice over and text to speech.
  347. 22:20–22:23The soul of a video lives not in the visuals but in the voice.
  348. 22:23–22:28A viewer will forgive a blurry scene, but will close the video at a bad voice, so read this
  349. 22:28–22:34chapter with care, and for speakers of mathily and other less-served languages, it holds some
  350. 22:34–22:35special advice.
  351. 22:35–22:40There are two ways to add voice, record your own, or have AI-generated text-to-speech,
  352. 22:40–22:42TTS for short.
  353. 22:42–22:47First, your own voice, because for mathily and languages like it, this remains the best
  354. 22:47–22:54out the pure pronunciation, the natural cadence, the rise and fall of feeling. No machine yet
  355. 22:54–22:59renders these as well as a native speaker and no studio is needed. A smartphone microphone
  356. 22:59–23:05today is quite good enough, follow a few rules, record in a quiet room fan off, windows shut,
  357. 23:05–23:10night or early morning is best. Hold the phone about a hand span from your mouth, speak
  358. 23:10–23:15standing or sitting upright. The voice stays open. Keep the script before you, but speak
  359. 23:15–23:21as if telling, not reading, and record paragraph by paragraph rather than all in one take,
  360. 23:21–23:24when you slip, you re-speak only that much.
  361. 23:24–23:30AI can polish a recorded voice, many tools, usually named enhanced voice, or studio sound
  362. 23:30–23:35in editing apps strip the background noise and give the voice a studio finish, use it
  363. 23:35–23:39without fail, the difference between a plain recording and an enhanced one will astonish
  364. 23:39–23:40you.
  365. 23:40–23:45Now text to speech figure six, here you type the text, choose the language and the speaker
  366. 23:45–23:52female or male, young or mature, adjust pace and pitch, and download an MP3 file for Hindi,
  367. 23:52–23:57English, and other major languages. This facility is very mature for mathily. The situation is
  368. 23:57–24:03improving. Some tools have begun to offer a mathily voice and tools built for Indian languages are
  369. 24:03–24:08the most likely place to find one. Search in your tool. If mathily appears, first test it
  370. 24:08–24:13with a short passage to hear how pure the pronunciation is. If no mathily voice is available,
  371. 24:13–24:19Two remedies. The first and best. Your own voice by the method above. The second.
  372. 24:19–24:23Making do with the Hindi voice. Write the text phonetically, listen and adjust the spelling
  373. 24:23–24:29until it sounds right. Keep the pace a little slow. Use short sentences. The result will not
  374. 24:29–24:35carry a fully mathal cadence, but it will serve. Remember, this is a compromise, not an ideal.
  375. 24:35–24:40Wherever feeling and purity matter poetry, stories, children's material, give your own voice.
  376. 24:40–24:47A word on a newer facility, voice-cloning, some tools, from a few minutes of your recording,
  377. 24:47–24:52build a digital replica of your voice, which will then read any text in your own tones,
  378. 24:52–24:56for content in a less-served language this is attractive, teach the tool your voice
  379. 24:56–25:01wants, and the voiceovers of many videos can be made, but two iron rules, clone only
  380. 25:01–25:06your own voice, imitating anyone else's voice without written permission is absolutely
  381. 25:06–25:12forbidden, and where a cloned voice is used in a video, disclosing it is good practice.
  382. 25:12–25:16The joining of voice and visuals will happen at the edit chapter 11, so keep the voice
  383. 25:16–25:21file separately in the voice music folder, make the voiceover first and the video clips
  384. 25:21–25:22after.
  385. 25:22–25:26This order is wise, because hearing the length of the voice tells you how many seconds
  386. 25:26–25:27each scene needs.
  387. 25:27–25:28Exercise.
  388. 25:28–25:33Make the voiceover of your script both ways, once recorded in your own voice, once through
  389. 25:33–25:38a TTS tool, listen to both with your eyes closed, which sounds more like you.
  390. 25:38–25:40Why?
  391. 25:40–25:41Chapter 9.
  392. 25:41–25:44Avatar Videos, The Digital Speaker.
  393. 25:44–25:49Imagine, every fortnight you must make an announcement video for a journal's new issue,
  394. 25:49–25:51or fifty lectures for a course.
  395. 25:51–25:55Camera, lighting, dress, recording, every single time.
  396. 25:55–25:56Impossible.
  397. 25:56–25:58The remedy is the Avatar video.
  398. 25:58–26:02A digital human sits on screen and speaks your written, text, lips moving, eyes
  399. 26:02–26:06blinking, hands gesturing, like a news anchor.
  400. 26:06–26:11The well-known tools of this class are Hei-jen, Synthesia, and others, and the class itself
  401. 26:11–26:12is growing fast.
  402. 26:12–26:15The structure is nearly the same in all figure 7.
  403. 26:15–26:161.
  404. 26:16–26:17Choose the avatar.
  405. 26:17–26:22Tools carry hundreds of ready-made avatars of different ages, dress and bearing.
  406. 26:22–26:27Some tools also let you build your own avatar from a photograph or a short video.
  407. 26:27–26:32That is, you remain on screen without recording each time if you make your own avatar.
  408. 26:32–26:37the same consent rule given for voice cloning, only your own likeness, never another's.
  409. 26:38–26:412. Give the script, it can take two forms,
  410. 26:41–26:46written text which the tool will speak through TTS, or your own recorded audio file which the avatar
  411. 26:46–26:52will lip sync, for mathily the second road is usually better, upload your own mathily voiceover,
  412. 26:52–26:56and the avatar speaks it, the lip sync you get is surprisingly good.
  413. 26:56–27:023. Choose the background and layout, a library, an office, a plain color,
  414. 27:02–27:07or an image of your own, choose the aspect ratio 16 colon 9 or 916, and generate.
  415. 27:08–27:13The beauty of the avatar video lies in its practicality, not in spectacle,
  416. 27:13–27:19new style presentation, journal announcements, lesson explanations, introductions of an institution,
  417. 27:19–27:24information videos, in all these it is excellent for the feeling laden delivery of story and
  418. 27:24–27:30poetry, a human is still better. One tip on presentation, do not keep the avatar on screen
  419. 27:30–27:36for the whole video without relief. In between, show related scenes, images or text cards called
  420. 27:36–27:42b-roll while the avatar's voice runs beneath. The video comes alive, this weaving happens at the
  421. 27:42–27:48edit, keep the avatar clip and the b-roll clips as separate files. And yes, when the avatar in a
  422. 27:48–27:53video looks human, the viewer has a right to know it is a digital speaker, write one line in the
  423. 27:53–27:59description, trust stays intact, and trust is a channel's real capital. Exercise, on the free
  424. 27:59–28:06plan of any avatar tool, make a 32nd introduction video, subject, an introduction to my village
  425. 28:06–28:13or an introduction to a favorite book, try both methods, type text and uploaded voice.
  426. 28:13–28:14Chapter 10.
  427. 28:14–28:16Music and sound effects.
  428. 28:16–28:21Visuals for the eye, voice for the ear and music, for the heart, the same scene feels
  429. 28:21–28:25lifeless without music and comes alive with the right score, but with music comes the
  430. 28:25–28:30The greatest danger of all, copyright, put someone's song in your video without permission
  431. 28:30–28:34and YouTube can block the video, others can claim its earnings, and the channel can be
  432. 28:34–28:39penalized, so rule one, never a famous film song or commercial recording, unless you
  433. 28:39–28:41hold written permission.
  434. 28:41–28:43Then where will the music come from?
  435. 28:43–28:45Free lawful sources.
  436. 28:45–28:50First, AI generated music, there are now tools that compose music from a written
  437. 28:50–28:56description. Write slow, tender, flute-led, 60 seconds and the music is ready. Some tools
  438. 28:56–29:01even build a full song, voice included, from your lyrics. Music made this way for your
  439. 29:01–29:06own video is generally safe to use, but read each tool's license once, especially whether
  440. 29:06–29:10commercial use including YouTube monetization is permitted.
  441. 29:10–29:152. Copyright-Free Music Libraries YouTube's own audio library inside YouTube
  442. 29:15–29:22studio is free and safe. Beyond it, many websites offer freely licensed music, some entirely
  443. 29:22–29:26free, some on the condition of attribution. If attribution is required, do not forget to
  444. 29:26–29:29write the musician's name in the video description.
  445. 29:29–29:32Third, your own recorded music.
  446. 29:32–29:37Methila has its own rich musical tradition, and so does every region. If you or someone
  447. 29:37–29:41you know, sings or plays, then a folk tune recorded by yourselves is the most authentic
  448. 29:41–29:45source of all, and it gives your video an identity no AI can.
  449. 29:45–29:50The tune of a folk's song is traditional, but a particular recording or arrangement
  450. 29:50–29:53belongs to its maker, keep this distinction in mind.
  451. 29:53–29:59The craft of laying music, under a voice of or keep the music low, 20 to 30% of the main
  452. 29:59–30:04voice, where there is no voice of or opening, close, scene changes the music may rise,
  453. 30:04–30:09let the mood of the music match the mood of the video, brightness for a morning scene,
  454. 30:09–30:14tenderness for a farewell, and at the end let the music sink away slowly fade out, music
  455. 30:14–30:17cut off abruptly jolts the ear.
  456. 30:17–30:22Sound effects are the small sounds, birdsong, the splash of water, the rustle of wind, they
  457. 30:22–30:26make a scene believable, some newer video models generate sound along with the scene
  458. 30:26–30:31native audio, if your tool has this, keep it on, if not, take sounds from a free
  459. 30:31–30:33library and add them at the edit.
  460. 30:33–30:35Exercise.
  461. 30:35–30:40background music of two different styles for your video, one AI generated, one from a free
  462. 30:40–30:45library. Play each behind the video in your mind's eye at least and consider which mood fits better.
  463. 30:46–30:49Chapter 11. Editing, Turning Clips into a Video
  464. 30:51–30:56Now you hold all the ingredients, video clips, voiceover, music, editing is the
  465. 30:56–31:01kitchen where these ingredients become the dish and the good news, editing skill,
  466. 31:01–31:06Once learned, serves in every video, it does not keep changing the way AI tools do.
  467. 31:06–31:10Choose an editor. Free editors exist for both mobile and computer.
  468. 31:10–31:15CapCite is at present the most popular and the simplest. On the computer,
  469. 31:15–31:19DaVinci Resolve is professional grade even in its free form. Choose either,
  470. 31:19–31:24the structure figure 8 is the same in all. The media area, where you bring and import
  471. 31:24–31:28all your files. The preview, where the video plays as you work. The timeline,
  472. 31:28–31:33The most important of all, the line of time on which the clips are arranged in order,
  473. 31:33–31:39the timeline has several strips tracks, one for video, one for voice, one for music,
  474. 31:39–31:43one for subtitles, stacked one above another, all playing together.
  475. 31:43–31:48Keep the basic order of editing thus. 1. First lay the voice over on the timeline,
  476. 31:48–31:52this is the spine of the video, the visuals will be arranged upon it.
  477. 31:53–31:572. Listening to the voice. Place the video clips in order,
  478. 31:57–32:01Let the scenes show what the words are saying, cut the clips, keep the best portion of each,
  479. 32:01–32:06remove the rest. AI clips are often awkward at the very start and the very end,
  480. 32:06–32:12trim both edges and the clip cleans up. Pre, add transitions, the manner of passing from
  481. 32:12–32:17one clip to the next, the rule, the fewer, the better, the plane cut is the purest,
  482. 32:17–32:21a light fade at a change of mood, thinning, twirling, color transitions are the mark of the
  483. 32:21–32:22the beginner.
  484. 32:22–32:234.
  485. 32:23–32:27Lay the music on its track and to bring its level down Chapter 10.
  486. 32:27–32:285.
  487. 32:28–32:29Color correction.
  488. 32:29–32:31Most editors have a one-click filter or enhance.
  489. 32:31–32:36If your AI clips came from different tools, put the same filter on all, the colors fall
  490. 32:36–32:39into step, and the video feels stitched of one cloth.
  491. 32:39–32:406.
  492. 32:40–32:44Add a title card at the start 3-4 seconds and a closing card at the end the channel's
  493. 32:44–32:47name, a word of thanks.
  494. 32:47–32:48Export settings.
  495. 32:48–32:491080p.
  496. 32:49–32:50MP4 format.
  497. 32:50–32:55format, 30 frames per second, the standard for YouTube, keep the exported file in the final
  498. 32:55–32:57folder.
  499. 32:57–33:02One suggestion after the video is done, watch it once from beginning to end without stopping,
  500. 33:02–33:07as a viewer, wherever your attention drifts, know that a cut is needed there, then show
  501. 33:07–33:11it to someone at home, the reaction of one first viewer teaches more than a hundred
  502. 33:11–33:12critics.
  503. 33:12–33:18Exercise, join all your clips, voice, and music into your first complete video,
  504. 33:18–33:23With title card and closing card, export it, show it to a family member and write down their
  505. 33:23–33:26first reaction.
  506. 33:26–33:27Chapter 12.
  507. 33:27–33:28Subtitles.
  508. 33:28–33:30A video that can be read.
  509. 33:30–33:35Most viewers today watch video without sound, on the bus, in the office, in bed at night.
  510. 33:35–33:40No subtitles, no viewers, and for content in a language like Mathalie the importance
  511. 33:40–33:42of subtitles is doubled.
  512. 33:42–33:46Subtitles in their original language build the habit of reading it, while Hindi or
  513. 33:46–33:50English subtitles bring in viewers who do not know the language at all, that is, your
  514. 33:50–33:53story travels the whole world.
  515. 33:53–33:58The standard format of subtitles is SRT, a plain text file in which three things repeat
  516. 33:58–34:04over and over figure nine, a serial number, a timeline from which second to which second,
  517. 34:04–34:08and the text, this file can be made and corrected even in an ordinary text editor
  518. 34:08–34:09notepad.
  519. 34:09–34:15But matching the timing by hand is laborious, and here AI helps again, two ways.
  520. 34:15–34:20First, Automatic Transcription, the auto captions feature in an editor such as CapCut listens
  521. 34:20–34:24to the video's voice and writes the subtitles itself.
  522. 34:24–34:25Timing included.
  523. 34:25–34:27In Hindi and English this is very accurate.
  524. 34:27–34:32A mathily voice it will usually hear as Hindi and write accordingly, then you correct the
  525. 34:32–34:37text it made, even so, correcting is far faster than writing from scratch, because the timing
  526. 34:37–34:39arrives ready made.
  527. 34:39–34:44Second, from the script, you already have the script written, give a chatbot your script
  528. 34:44–34:49and the video's total duration and say, divide this into SRT format, each subtitle at most
  529. 34:49–34:54two lines, at a comfortable reading pace, load the file in the editor and nudge the timings
  530. 34:54–35:01forward or back. The craft rules of subtitling, at most two lines at a time, roughly 32 to
  531. 35:01–35:0740 characters per line, each text stays on screen at least one second, a sentence breaks
  532. 35:07–35:12where the meaning allows I went slash to the market, no, I went to the market together,
  533. 35:12–35:16centers in white with a light dark shadow or strip behind, so they can be read even over
  534. 35:16–35:21a bright scene, place them at the lower middle of the screen, but not so low that the real
  535. 35:21–35:23format cuts them off.
  536. 35:23–35:28On YouTube, subtitles can be given in two ways, burned into the video joint at export
  537. 35:28–35:33from the editor or upload it as a separate SRT file which the viewer can switch on and
  538. 35:33–35:38off, the best method, give the original language subtitles as a separate file, and
  539. 35:38–35:43at Hindi and English as separate SRT files too, a chatbot will help with the translation
  540. 35:43–35:49the duty of checking it remains yours, thus one video reaches the viewers of three languages.
  541. 35:49–35:54Exercise, make the SRT of your video in its own language by either method, then make its
  542. 35:54–35:59Hindi or English translation file, run both with the video and check, is the timing
  543. 35:59–36:02right, is any line too long?
  544. 36:02–36:05Chapter 13, Publishing on YouTube.
  545. 36:05–36:09The video is made, now carry it to the viewer.
  546. 36:09–36:13YouTube remains the largest and the most lasting platform, a video placed here keeps
  547. 36:13–36:18being watched year upon year, while on real format platforms a video's life is a few
  548. 36:18–36:23days, so make YouTube the main house, let reels and short speed its windows.
  549. 36:23–36:28Making a channel is simple, sign into YouTube with your Google account and create one,
  550. 36:28–36:32choose the channel's name with thought, short, easy to say, and suggestive of the
  551. 36:32–36:38subject, the channel picture logo and banner too can be made with an AI image tool.
  552. 36:38–36:42At upload time keep the checklist of figure 10 before you, a few points in detail.
  553. 36:42–36:48Title, the main matter in the first 3 or 4 words, with a title in your own language,
  554. 36:48–36:52adding Hindi or English in brackets helps the video surface in search, because seekers
  555. 36:52–36:54search in every language.
  556. 36:54–36:58Description, the first 2 lines are the most valuable.
  557. 36:58–37:03These appear in search results, write the video's essence here, key words included, below them,
  558. 37:03–37:09the chapter list what comes at which minute, the list of sources, and the channel introduction.
  559. 37:09–37:12Thumbnail The viewer sees the thumbnail before the title,
  560. 37:12–37:17the rule, one image, one feeling, at most three words, and words large enough to be
  561. 37:17–37:22read on a small mobile screen, make an attractive thumbnail with an AI image tool, but never
  562. 37:22–37:24a misleading one.
  563. 37:24–37:28That is in the thumbnail must be in the video where the viewer feels cheated and trust in
  564. 37:28–37:30the channel is gone.
  565. 37:30–37:34The AI disclosure, YouTube now expects that realistic looking AI generated or AI altered
  566. 37:34–37:39content be declared at upload in answer to the altered content question.
  567. 37:39–37:43This is not mere rule keeping, it is honesty with the viewer, for plainly imaginary styles
  568. 37:43–37:48such as animation the duty usually does not arise, but when in doubt, declaring is
  569. 37:48–37:50always the better course.
  570. 37:50–37:51Language setting.
  571. 37:51–37:53Choose the video's language.
  572. 37:53–37:56Mathalie is in YouTube's list, as are many others.
  573. 37:56–38:01This helps the video reach those searching in that language, and it strengthens the statistics
  574. 38:01–38:04of the language's content besides.
  575. 38:04–38:05After the upload, what then?
  576. 38:05–38:10Watch the response of the first hours and days, reply to comments, the early conversation
  577. 38:10–38:13waters the channel's roots, and keep regularity.
  578. 38:13–38:18One video a fortnight makes 24 in a year, and this bears more fruit than a hundred
  579. 38:18–38:19videos at random.
  580. 38:19–38:24attach themselves to a program, not to scattered surprises.
  581. 38:24–38:28For reels and shorts, cut the most engaging 30 to 60 seconds of your main video into the
  582. 38:28–38:349-16 ratio with the editor's reframe or crop feature, and right at the end, full video
  583. 38:34–38:38on the channel, this is the window that leads new viewers to the house.
  584. 38:38–38:40Exercise.
  585. 38:40–38:43Upload your video, completing all 7 points of the checklist.
  586. 38:43–38:46Also cut a shorts version and upload it separately.
  587. 38:46–38:51After one week, compare the figures of the two views, watch time.
  588. 38:51–38:52Chapter 14.
  589. 38:52–38:55Special Considerations for Mathalie Content
  590. 38:55–38:57This chapter is the heart of this book.
  591. 38:57–39:01The tools are universal, but our purpose is particular, immathalie, for Mithila, and
  592. 39:01–39:05readers working in any less served language will find the same principles applied to
  593. 39:05–39:06their own.
  594. 39:06–39:12First, purity of language, AI tools, when writing Mathalie, usually let the shadow
  595. 39:12–39:16of Hindi fall across it because they have learned far more Hindi.
  596. 39:16–39:19Trips, translations, subtitles made by a chatbot.
  597. 39:19–39:22Check every one with your own eyes, the plain rule.
  598. 39:22–39:27The A.I.s Mathili is a draft, not an authority, where in doubt, trust your ear, speak the
  599. 39:27–39:32line aloud, whatever grates on the ear is the shadow of Hindi.
  600. 39:32–39:34Second, the question of script.
  601. 39:34–39:39Mathili is written in Devanagari and it also has its own ancient script, Turhuta Mythalikshara.
  602. 39:39–39:45the video subtitles and text cards in Devanagari, the most people will be able to read them,
  603. 39:45–39:50use Turhuta for beauty and identity, entitle cards, in the logo, in a watermark,
  604. 39:50–39:56thus the script stays before the eye, and curiosity awakens too, remember,
  605. 39:56–40:01AI image tools cannot yet write Devanagari or Turhuta letters correctly, always add written,
  606. 40:01–40:04text yourself at the edit, never have the AI write it.
  607. 40:04–40:11Third, authenticity of the visuals. Tell an AI and Indian village and it will produce a generalized
  608. 40:11–40:16North Indian scene, which is not Mithila. Give the prompt Mithila's particular signs. The pond,
  609. 40:16–40:22the banyan, the mango orchard, the patty field, fish, pond, the thatched house, the
  610. 40:22–40:27Tulsi platform in the courtyard, wall ornament in the manner of Arapan. Even then, what comes
  611. 40:27–40:33will be Mithila like, not Mithila, so wherever possible, blend in real photographs and footage
  612. 40:33–40:38by the method of Chapter 7, the mixture of AI scenes and real scenes gives the most authentic result
  613. 40:38–40:44of all. Fourth, the honor of Madhubani's slash Mathila painting, the AI can be told to generate in
  614. 40:44–40:49Madhubani style, and it will imitate the colors and the line, but pause here and think. Mathila
  615. 40:49–40:54painting is a living tradition, the livelihood of thousands of artists rests on it, and each
  616. 40:54–41:00of its manners Barney, Kachni, Godna, Gober carries its own lineage, and AI made Madhubana
  617. 41:00–41:04like image lifts the tradition's appearance without its labor and its knowledge, my counsel,
  618. 41:04–41:09in your videos show real works by real artists with permission and credit.
  619. 41:09–41:12This honors the artist and strengthens your video at once.
  620. 41:12–41:17If you do use AI ornament in the Madhubani manner, say plainly that it is an AI mediation,
  621. 41:17–41:19not authentic Madhubani.
  622. 41:19–41:205.
  623. 41:20–41:24An inexhaustible store of subjects The field of mathily video is still nearly
  624. 41:24–41:25empty.
  625. 41:25–41:31Whatever makes makes first, some directions, children's material songs, tales, letters,
  626. 41:31–41:37the child audiences, the fastest growing of all, festival explainer C. H. H. I., Sama
  627. 41:37–41:44Chaikpa, Jersital, Huat, Why, How, recipes, folk tales and the verses of Vidya Patti presented
  628. 41:44–41:48with images, the vocabulary of village and home a visual dictionary of the words now
  629. 41:48–41:53slipping away, introductions to the places of Mathila, each direction could be a channel
  630. 41:53–41:55in itself.
  631. 41:55–42:00the last word, the patience of quality. In Methili the audience will be smaller than in Hindi.
  632. 42:00–42:05This is natural, but the loyalty of the Methili viewer is greater. They will comment,
  633. 42:05–42:10they will share, they will return again and again, look not at numbers but at relationships.
  634. 42:10–42:13A hundred devoted viewers are worth more than 10,000 indifferent ones.
  635. 42:14–42:19Exercise. Choose one of the six directions above and plan three consecutive videos on its
  636. 42:19–42:25subject plus a one-line summary each. 3. Because one video is an experiment, 3 are a direction.
  637. 42:27–42:31Chapter 15. Copyright, ethics, and responsibility.
  638. 42:32–42:37With a powerful tool comes responsibility. This chapter is short, but bring its every line into
  639. 42:37–42:43practice. Copyright, the root principle, what you did not make, you do not use without permission,
  640. 42:43–42:49film songs, portions of others' videos, the text of books, other people's photographs,
  641. 42:49–42:52The rule covers them all. Everyone does it is no argument.
  642. 42:52–42:57YouTube's automatic system content ID catches it, and the channel bears the penalty.
  643. 42:57–43:02In freely licensed material too, read the conditions, one says a tribution required,
  644. 43:02–43:08another no commercial use. No also the question of rights over your own AI made material.
  645. 43:08–43:14In most tools terms, permission for commercial use of generated video and images comes with
  646. 43:14–43:19the paid plan, and is sometimes limited on the free one for any tool you work with regularly,
  647. 43:19–43:23read its terms of use once, especially the ownership and commercial use sections.
  648. 43:24–43:30The dignity of persons, three iron prohibitions, no imitation of anyone's face or voice without
  649. 43:30–43:36written consent living or departed, both. Never show a person in a scene that lowers their honor,
  650. 43:36–43:41nor put in their mouth words they never spoke the deep ache, and take special care with the
  651. 43:41–43:46images and voices of children before posting a video even of your own child. Consider that
  652. 43:46–43:52it is becoming public. Toward truth. AI can make scenes that look real and therefore it
  653. 43:52–43:56can deceive. The rule. Let fiction be called fiction and news be news.
  654. 43:56–44:01When you make a video on a historical or cultural subject, check the facts against your
  655. 44:01–44:06own trusted sources. An AI chatbot speaks its errors with full confidence this is called
  656. 44:06–44:11hallucination. Wherever an AI scene looks real, declare it chapter 13.
  657. 44:11–44:15The viewers trust takes years to earn and one video to lose.
  658. 44:15–44:16And toward yourself.
  659. 44:16–44:19Let AI's convenience never turn into laziness.
  660. 44:19–44:21Take work from AI, not thought.
  661. 44:21–44:26The day you hand the script, the choices, and the final judgment over to AI, that day
  662. 44:26–44:30the video ceases to be yours, and viewers can smell the difference.
  663. 44:30–44:32Keep the tool in your hand.
  664. 44:32–44:34Let not the hand become the tool.
  665. 44:34–44:35Exercise.
  666. 44:35–44:40Run a responsibility check on the first video you made, the music's license, permission
  667. 44:40–44:45for the image sources, the facts verified, the day I disclosure, keep the four answers
  668. 44:45–44:46in writing.
  669. 44:46–44:52Repeat this exercise with every video, within days it will become second nature.
  670. 44:52–44:58Chapter 16 The Practice Project, a first video from start to finish.
  671. 44:58–45:03Now all the threads in one place, this chapter is the model of one complete project, a
  672. 45:03–45:0790 second video called CHathai, the great festival of sun worship.
  673. 45:07–45:11it, then repeat the same sequence on a subject of your own.
  674. 45:11–45:12Day 1.
  675. 45:12–45:17Planning Chapters 3, 4, Purpose, to explain simply the spirit and the observance of Siddhi
  676. 45:17–45:23Chathai, audience, Mathal families living away from home, the new generation especially,
  677. 45:23–45:29duration, 90 seconds, that is, 9 or 10 clips, a two column script drafted by a chat pot,
  678. 45:29–45:34corrected by hand in three places, local detail added to the description of Nahikai, the
  679. 45:34–45:36feeling of the final line deepened.
  680. 45:36–45:41Ten scenes on the storyboard, the river ghat at dawn establish, the courtyard where thequeua
  681. 45:41–45:46is being made, the supandora of the offering, the evening argia, the morning argia, the
  682. 45:46–45:49parent, and the closing card resolve.
  683. 45:49–45:55Day 2 Voice Chapter 8, the voiceover recorded in my own voice, at night, in a quiet room,
  684. 45:55–46:01paragraph by paragraph, polished with the editor's enhanced voice, total 82 seconds,
  685. 46:01–46:03which means the scene plan sits right.
  686. 46:03–46:09Phase 3-4, visuals chapters 5-7, a six-part prompt built for every scene thought out and
  687. 46:09–46:11mathily, translated into English.
  688. 46:11–46:17Methylas signs in each, a river ghat in north Bihar, banana stems, bamboo kania, fruit in
  689. 46:17–46:21the soup and aura, three clips came out well at the first attempt, four clips took two
  690. 46:21–46:24or three improvement cycles.
  691. 46:24–46:28For two scenes the tekua and the soup the image of video method was used, first
  692. 46:28–46:31the image perfected, then subtle motion given.
  693. 46:31–46:37And in one scene a real photograph of my own was used, the crowd at the GAT, which no AI
  694. 46:37–46:38could ever have made.
  695. 46:38–46:41This mixture is the essence of Chapter 14.
  696. 46:41–46:42Day 5.
  697. 46:42–46:45Music and Assembly Chapters 1011.
  698. 46:45–46:46AI Music.
  699. 46:46–46:52Slow, devotional, flute and emridang, 90 seconds, the license checked, the editing order,
  700. 46:52–46:57voice first, then clips, plain cuts, a fade in two places, one and the same light warm
  701. 46:57–47:02filter on every clip, on the title card CHathai in Terhuda and the full title in Devanagari.
  702. 47:03–47:09Day 6, subtitles and publication chapters 12-13, the Mathili SRT from the script,
  703. 47:09–47:14then Hindi and English translation files translated by Chatbot, checked by hand.
  704. 47:14–47:18The thumbnail, the scene of the evening argh, three words,
  705. 47:18–47:25hui chathai, ad upload, language Mathili, the AI disclosure switched on, music credit and sources
  706. 47:25–47:29in the description. Alongside, a 35-second shorts version.
  707. 47:30–47:32Day 7 Review Chapter 15,
  708. 47:32–47:36the responsibility check on all four points shown to the family,
  709. 47:36–47:41the suggestion came that the morning argia scene feels short, noted down for the next video.
  710. 47:41–47:46See, seven days, an hour or two a day, and from nothing to a published video,
  711. 47:46–47:51the first project will take longer than this, let it, by the third or fourth video the
  712. 47:51–47:55sequence becomes habit and the time falls by half.
  713. 47:55–47:56A last word.
  714. 47:56–47:58This book ends here, but the learning begins here.
  715. 47:58–48:03The tools will change, the models will change, the screens will change, but the framework
  716. 48:03–48:09you have learned plan, prompt, improvement cycle, assembly, responsibility will stand.
  717. 48:09–48:15Now go and tell the story of your Mathila, in your own voice, in your own language.
  718. 48:15–48:18The Pentaxe, Lausary.
  719. 48:18–48:23AI artificial intelligence, the capacity of a computer to learn and understand in a human-like
  720. 48:23–48:24way.
  721. 48:24–48:31Generative AI, AI that creates new material, text, images, video, voice.
  722. 48:31–48:32Prompt.
  723. 48:32–48:34The written instruction given to an AI.
  724. 48:34–48:38Negative prompt, the list of what is not wanted.
  725. 48:38–48:42Text to video, the method of making video from a written description.
  726. 48:42–48:46Image of video, the method of setting a still image in motion.
  727. 48:46–48:50TTS text-to-speech, the method of producing a spoken voice from written text.
  728. 48:51–48:54Avatar, a digital speaker who delivers text on screen.
  729. 48:55–48:58Lip sync, the matching of voice and lip movement.
  730. 48:59–49:02Voice cloning, making a digital replica of a voice.
  731. 49:03–49:07Credit, the usage currency of AI tools, each generation deducts some.
  732. 49:08–49:11Model, a particular version or engine of an AI.
  733. 49:12–49:15Clip, a short segment of video usually 5-20 seconds.
  734. 49:16–49:23Aspect Ratio, the ratio of a screen's width to its height, 16 colon 9 wide, 9 16 upright.
  735. 49:23–49:29Resolution, the fineness of the image, 1080p standard, 4K extra fine.
  736. 49:29–49:33Watermark, an identifying mark printed on content.
  737. 49:33–49:36Storyboard, the scene by scene sketch plan.
  738. 49:36–49:39B-roll, supporting footage beyond the main speaker.
  739. 49:39–49:44Timeline, the line of time in an editor on which clips are arranged.
  740. 49:44–49:51Track, a layer of the timeline, video, voice, music and so on. Cut, trimming a clip, or passing
  741. 49:51–49:57directly from one clip to the next. Transition, the manner of joining two clips fade and the rest.
  742. 49:58–50:03Fade in slash out, slow emerging slash, slow vanishing. Render slash export,
  743. 50:03–50:10turning the edited video into the final file. SRT, the standard file format of subtitles.
  744. 50:10–50:15Burned in, subtitles fixed permanently into the video.
  745. 50:15–50:18Thumbnail, the face image of a video.
  746. 50:18–50:21Hallucination, an AI's confident error.
  747. 50:21–50:25Content ID, YouTube's copyright checking system.
  748. 50:25–50:29Deepfake, a deceptive imitation of someone's face or voice.
  749. 50:29–50:32Monetization, earning from videos.
  750. 50:32–50:36Appendix B, a collection of prompts.
  751. 50:36–50:40The lower 10 ready prompt frames, fill the brackets with your own details and
  752. 50:40–50:43shape the final prompt by the method of Chapter 5.
  753. 50:43–50:441.
  754. 50:44–50:45Village Dawn
  755. 50:45–50:50A village in North Bihar, Dawn, light mist over the fields, thatched houses, a bayon tree
  756. 50:50–50:56in the distance, your detail, camera panning slowly to the right, soft golden light, realistic
  757. 50:56–50:57documentary style.
  758. 50:57–50:582.
  759. 50:58–50:59Pond Scene
  760. 50:59–51:04A pond covered with lotus and water lilies, seasoned, on the bank tree slash gat.
  761. 51:04–51:09Gentle ripples on the water, subtle movement of bird slash fish, camera moving forward
  762. 51:09–51:15very slowly. Time of day light, semantic. Free, courtyard scene, the courtyard of a
  763. 51:15–51:21mythillahome, a tolcy platform, action, say, food cooking on the hearth, folk ornament on
  764. 51:21–51:28the wall, warm intimate light, close up, realistic. For, festival scene, preparations for name
  765. 51:28–51:34of festival, central object, say, soup and dora, lamps, courtyard or get, close ups
  766. 51:34–51:39of busy hands, festive joy, warm light, documentary style.
  767. 51:39–51:405.
  768. 51:40–51:46Nature scene, green paddy fields waving in the wind, time, the far horizon, bird in flight,
  769. 51:46–51:50drone view rising slowly, natural light, ultra-fine detail.
  770. 51:50–51:526.
  771. 51:52–51:57Object study dish slash craft, an extreme close-up of object, background, light steam
  772. 51:57–52:02or sheen, camera circling the object slowly, soft studio-like light, appetizing slash
  773. 52:02–52:04appealing style.
  774. 52:04–52:107. Doreen illustration animation. Character description doing action. Place. Hand drawn
  775. 52:10–52:16animation style, soft watercolor, like a children's book illustration. Slow pleasant motion.
  776. 52:16–52:228. Historical imagination. An imagined scene of period methela. Action. Colors like an
  777. 52:22–52:28old palm leaf painting, slow solemn camera, museum documentary style. The AI disclosure
  778. 52:28–52:31is obligatory here. This is imagination.
  779. 52:31–52:379. Title card background, an abstract ornamental pattern drifting slowly, color scheme.
  780. 52:37–52:43Say, for Millie and Red and Yellow, no letters, no figures, come even motion, 10 second loop,
  781. 52:43–52:49add the lettering yourself at the edit. 10. Motion instruction for image video,
  782. 52:49–52:54in this image, gentle natural motion in element, say, water slash leaves slash cloth,
  783. 52:54–52:59camera static slash moving forward very slowly, nothing else changes, the image's
  784. 52:59–53:03original color and line remain intact.
  785. 53:03–53:08The Pentax C, a guide to the tools as of 2026.
  786. 53:08–53:13This list reflects the state of things at the time of writing 2026, names and features
  787. 53:13–53:19keep changing, learn the categories, do not memorize the names, before adopting any tool,
  788. 53:19–53:24check three things, what the free plan gives, whether commercial use of the output is permitted,
  789. 53:24–53:27and where the watermark stands.
  790. 53:27–53:34Text-to-video slash image video, Google's Vivo, OpenAI's Sora, Kling, Runway, Pixverse, Seedance,
  791. 53:34–53:42Haleuo, Luma, Pika, Realistic Scenes, Cinematic Clips, Avatar video, Hey-Gen, Synthesia, Digital
  792. 53:42–53:45Speakers, LipSync, Many Languages.
  793. 53:45–53:51Text-to-speech and voice cloning, 11 Labs and others, services centered on Indian languages
  794. 53:51–53:56are also growing, a mathily option will appear there first, keep watching.
  795. 53:56–54:03AI music, Suno, Udio and others, music and song from a description or from lyrics.
  796. 54:03–54:08Image generation, included in nearly every video tool, separately, Mijirni, and the image
  797. 54:08–54:11features attached to the chatbots.
  798. 54:11–54:17Editing, Cap Cut, Mobile and Computer, simple auto-cactions included, DaVinci Resolve,
  799. 54:17–54:24Computer, Free, Professional Grade, Canva, Thumbnails, Banners, Simple Video.
  800. 54:24–54:30Call in one plant a video, tools of the in-video kind, which, given a topic, join script, visuals
  801. 54:30–54:35and voice on their own, good for quick work, but your own stamp is faint in them, so for
  802. 54:35–54:39learning, the step-by-step method of this book is better.
  803. 54:39–54:46Chatpot script, translation, SRT, brainstorming, Claude, chat GPT, Gemini.
  804. 54:46–54:51And last of all, this book too will age with time, but you will not, the teach yourself
  805. 54:51–54:56temper this book has built in you will apply to every new tool when you seal a new tool
  806. 54:56–55:01ask three questions where is the prompt written what is in the settings where and how does
  807. 55:01–55:04the result arrive and within five minutes you will be working it.

Plain text

Making videos with AI. Teach yourself series, written by Gage and Rathakar. Preface. This book is for everyone who wants to say something through video but has no camera, no studio, and no training in editing. Artificial intelligence AI has now made all three possible with an ordinary mobile phone or computer. What once required equipment worth hundreds of thousands and a team of dozens can now be done with a written instruction, called a prompt. but let one thing be clear at the very outset. AI is no magic wand. It is a tool, like an axe, like a pen. The axe cuts the wood, but which tree to fell and what house to build. The carpenter decides, in the same way. AI will make the visuals and generate the voice, but what to say, why to say it, and for whom, the answers to these three questions remain with you. This book will teach you to handle the tool, the craftsmanship will remain your own. This book is written in the teacher-self style, that is, no teacher or training institute is needed. Each chapter is one step, begin with the first chapter, move forward in order, and be sure to do the exercise give at the end of every chapter. Reading and doing must walk together, only then does learning happen. One more thing, all the illustrations in this book are original explanatory figures, not copies of any company's actual screens. There are For two reasons. First, the screens of AI tools change every few months, so a real screenshot would be outdated before the book left the press. Second, once you understand the principle, any new screen will feel familiar. If you memorize a screen, every change will leave you stranded. The figures in this book show the structure that applies to nearly every tool alike. Where the prompt box sits, what the settings contain, what a timeline looks like. Once you grasp this framework, you will open any new tool and work it out on your own, and that is the true meaning of teach yourself. The absence of technical books in Mathalie has long been a sore point, and this book was first written in Mathalie. This English edition carries the same conviction. Whatever you learn here, use it to put your own language, your own region, and your own stories on screen. The tales, psalms, paintings, history and festivals of Mathila, and of every homeland like it, are all waiting for their videos. Number 1. What is AI video? Artificial Intelligence. AI is the capacity of a computer to do human-like work. Understanding language, recognizing images, and now, creating images and video. When we write to an AI, show two children playing by a pond, and it produces a moving picture of that very scene on its own. This is called generative AI. To generate means to bring forth. This This AI does not fetch a video from somewhere. It composes a new video that never existed anywhere before. How is this possible? Understand it in simple terms. These AI models have been shown crores of images and videos, having seen so much. They have learned what an egret looks like, how water ripples, what color the morning light takes, and how a walking person's feet rise and fall. When you write an egret beside a pond at dawn, the model builds such a scene from its learned knowledge, just as a painter, who has seen thousands of ponds can now draw a pond from memory without one before their eyes. There are chiefly four routes to making video with AI. First, text to video, you only write, and the AI creates the entire scene. This is the most astonishing route, but it offers somewhat less control, the scene that arrives may not match your imagination exactly. Second, image of video, you supply one still image, a photograph or an AI generated picture. and the AI sets it in motion, here the control is greater because you yourself chose the opening frame. Third, avatar video. A digital speaker avatar sits on screen and speaks your written text, like a news anchor, for lessons, announcements and lectures this is extremely useful. Fourth, AI assisted editing. Here the footage is your own recording, but the cutting and joining, the subtitles, the background removal. All this the AI does. This book covers all four routes, because in practice a good video is usually a blend of them. Now understand what AI cannot do, AI does not yet produce long videos in one go, it typically makes short clips of 5 to 20 seconds, which are joined to build a longer video, AI may see sometimes contain errors, a hand grows too many or too few fingers, written letters come out garbled, a character's face changes between one clip and the next, and the biggest point of all, AI knows nothing of Mathila, of Mepheli, or of what lies in your heart, it will give only as much as you know how to ask, that is why half of the spoke is devoted to how to ask, that is, planning and prompts. Exercise, on paper, write down three subjects for videos you would like to make, for each, write one line, who will watch this video and what will they gain from watching it. Chapter 2, Getting Ready, What Do You Need? No costly machinery is needed to make video with AI, all that is required is this. First, a device, a smartphone is sufficient, a computer or laptop adds convenience, writing prompts, managing files and editing are easier on a large screen, no specially powerful computer is needed because the video is made not on your device but on the company servers, your device merely sends the instruction and downloads the finished video. Second, the Internet, the better the speed, the shorter the weight, video files are large, so downloading will consume data. Budget a few gigabytes a month. Third, an email account, nearly every AI tool requires an account, and most allow direct sign-in with a Google email Gmail, one suggestion. Keep a separate email for this work, so that the newsletters of AI tools do not flood your main inbox. Fourth, an understanding of credits. First AI video tools run on a credit system, a credit is a kind of coupon. Making one video deducts some credits, a free account receives a small allowance, daily and some tools monthly and others, once only in a few. Paid plans bring more credits, higher quality such as 1080p or 4K, videos without watermarks and longer clips. Adapt a practical policy here, which I call cheap first, then deep, while learning a new tool, run small experiments on the free credits, learn to write prompts, learn the tools temperament, when it feels time for serious work, say, running a YouTube channel, then take a paid plan on one tool, buying plans on every tool is wasteful, one or two suffice. 5th, a system for files, it sounds a small matter, but the experienced know how work drowns without order, make a separate folder for each video project on your computer or phone, inside it keep four subfolders, script scripts and prompts, clips raw AI made video, voice music voiceover and music, and final the finished video, give every file a meaningful name, a file called clip a one-pond on will still be recognizable six months later, video three final new two will not. And finally, patience, AI tools are sometimes busy, sometimes give strange results, sometimes fail entirely to understand your perfectly good prompt, all this is natural, those Those who persist through three or four attempts learn, those who quit at the first odd result are left behind. Exercise, create the folder system described above on your device, make a new email account if you wish, and open the website of any one AI video tool simply to see what the free plan offers, do not make anything yet, only look. Chapter 3 – Planning the video, from idea to script The biggest mistake beginners make is to open the tool straight away and start pressing buttons, a video made without a plan looks exactly like a house built without a drawing. That is why this chapter comes before the tools. Before every video, write the answers to three questions. 1. What is the purpose? To teach say how Matubati painting is made, to tell a story, to show scenes of a village or to an else. 1. Video, 1 purpose. Hold to this rule. A video that contains everything contains nothing. 2. Who is the audience? Children, students, expatriates far from home, curious outsiders, or the general viewer, the audience decides how simple the language should be, how long the video should run, and what the visuals should look like, bright colors and quick movement for children, stillness and gravity for adults. Free, how long? In the beginning, make videos of 30 seconds to 2 minutes. A short video is easier to make, easier to fix and better liked by today's viewer. Remember, AI produces clips of 5-10 seconds, so a 1-minute video means 6-12 clips. Now the script. A script is nothing complicated or literary. It is simply a two-column table. In the left column, what is seen the visual. In the right column, what is heard the voice or subtitle. For a 30-second video, 5 or 6 rows are enough. I can help in a second role as script assistant give your subject to a chat bot such as Claude chat GPT or Gemini and ask it to draft the script but the asking needs skill merely saying write a video script on Madhubani will fetch a generic lifeless text ask like this I am making a 60 second YouTube video on Madhubani painting audience young Indian viewers who have heard the name but no a little more. Voice, simple and warm, I need a script in two columns, on the left, a description of each visual which I will give to an AI video tool, on the right, the voice of her text. Six scenes, 10 seconds each. Notice, subject, duration, audience, language, format, and use are all stated. The clearer the ask, the better the result. This same principle will apply later to video prompts. To not accept the chatbot's script with with your eyes closed, read it, test it against your own knowledge. Are the facts right? Does the language sound like your own? Change any line that falls flat, or have it rewritten make the third scene more tender? Shorten the last line, the script is your signature, the AI is only a scribe. Exercise. Take one of the three subjects you chose in chapter 1, write the answers to the three questions purpose, audience, duration. Then, using the detailed style of asking shown above, have a chatbot draft a two-column script and make at least two corrections to it with your own hand. Chapter 4. The Storyboard. A Map of Scenes. The script is a map of words, the storyboard is a map of scenes, a storyboard lays out one sketch, a rough drawing or description, for every scene of the video in order. In film making this method is a century old, and in the AI age its importance has not shrunk but grown. Why? Because AI needs a separate, clear instruction for every clip, and the storyboard is precisely that list of instructions. Do not worry, no drawing skill is required, round faces, stick figures, and arrows are enough, if you would rather not sketch at all, write three lines for each scene, what is seen, what the camera does, and how many seconds. Learn a little of the camera's language for these very words will serve you later in prompts. Wide shot, the whole scene from a distance, for establishing the place, like a full view of the village. Close up, from near, for feeling and fine detail, like the steam over a cup of tea on the hearth. Tracking shot, the carer moves along with the subject, like following a child on the way to school. Drone view, from above as a bird sees, like a sweeping view of pond and fields. slash out, the carousel slowly draws near or pulls away, for emotional weight. Static shot, the camera stays in one place, for calm, settled scenes. Keep one simple formula for the order of scenes, establish, develop, resolve. The first scene tells where we'd are established, usually a wide shot. The middle scenes bring the subject close develop, close ups, tracking, the last scene gathers at at all, a feeling, a message or title card resolve. Figure 2 shows the storyboard of a 32nd video called My Village. See how six scenes travel from dawn to dusk. Morning mist establish, tea, school and pond develop, lamps and title resolve. The side each scene its duration and camera note are written, each of these scenes will, further on, become one prompt. While making the storyboard, attend to continuity. If the first scene is morning, the second must not suddenly be night. a character wears a red kurta. The red kurta must appear in every scene, and this must be written into every prompt, because AI does not remember the previous clip. Write the character's description on a separate sheet, a character card, and paste it word for word into every prompt. This small trick is the simplest way to keep the clips consistent. Exercise. Turn your script into a 16 storyboard. For each scene, write the visual description, The camera, the duration, if there is a character, make a character card dress, age, appearance, in three lines. Chapter 5 The Art of Prompt Writing A prompt is the written instruction you give to the AI. It is the most valuable skill of the AI age, and happy news is that it demands no technical knowledge. Only clarity of language and language is our home ground. See the difference between a poor prompt and a good one. Poor. This tells that AI nothing, a village of which country, which season, day or night, what is happening, it will invent something from its own mind, most likely some European or placeless village, now the good one, a village in North India, early morning, light mist over the fields, thatched and tiled houses, a banyan tree in the distance, camera panning slowly to the right, soft golden light, realistic documentary style, now the AI holds the complete picture. Keep in mind the six-part formula shown in Figure 3. Subdecked, who or what is central, two egrets. Action, what are they doing? Catching fish. Seen. Where and when, a pond full of lotus, et don. Camera, what does the camera do? Slowly moving in. Light, what kind of light? Golden morning light. Style, what look? Documentary, realistic. Let every prompt carry all six parts and keep roughly this order. Subject and action first. Style last. Some useful words of style. Realistic. Cinematic. Animation. Like a watercolor painting. Like old film. Documentary. The question of language. Most AI video tools understand English prompts best. Some also follow Hindi and other languages. The practical path is this. They can script in your own language and shape the final prompt in English. a chatbot helps, give it your scene description in your language and say, turn this into an English video prompt with six-part subject, action, scene, camera, light, style, thus the imagination stays yours, only the translation is mechanical. There is also the practice of the negative prompt, where you state what you do not want, for example, no blur, no distorted hands, no text on screen, no watermark, some Some tools give it a separate box, in others it is added to the main prompt. Now the most important principle of all, the improvements cycle figure 4. The first prompt rarely yields the video of your wishes, and this is no failure, it is the method itself. Look at the result, name the fault, change just that much in the prompt, and generate again. Did the egret come out too large? At a small egret, does the scene look garish? Change to soft, gentle light, usually within 3-5 cycles a usable clip arrives. Change only one or two things per cycle. Change everything at once and you will never know which change did the work. Keep saving your good prompts in one file. The you of six months hence will thank the you of today. Exercise. For the first scene of your storyboard, write one complete prompt using the six-part formula in your own language, then make its English form yourself or through a chatbot. Keep both in your script folder. Chapter 6. Text video tools. The first clip. The moment has come to take the tool in hand. In this chapter we understand the common structure of text-to-video tools and make the first clip. First, an introduction to the tools, this field changes at great speed, every few months a new model arrives and the old ones grow stronger. At the time of writing 2026 the leading names are, Google's BO, OpenAI's Sora, Kling, One Way, Pixverse, Seedance, Haleuo, Luma, and Pika. Each has its own temperament, one excels at realistic scenes, another at stylized or artistic ones, another is faster and cheaper, the names will keep changing, but the structure, described below, remains nearly the same in all of them, so learn the structure, not the names. Look at figure 5, in almost every tool you will find these 5 things. 1. The prompt box. A large empty field where you write your instruction, this is the heart of the tool. 2. selection. A single company offers several models, new and powerful costly, older or fast cheap, choose the cheap model while learning. The good model for final, publishable clips. 3. Aspect Ratio, 16,9 for YouTube and television-wide, 9,16 for Reels, Shorts, and Status Upright, 1,1 square, the side at the outset where the video will go, because changing the ratio leader crops the scene. 4. Duration. Usually options of 5, 8 or 10 seconds. A shorter clip costs fewer credits and carries fewer errors. 5. The generate button and the results area. Press the button, wait from a few seconds to a few minutes and the clip appears in the results area. Download it from there. Now the method for the first clip, step by step. 1. Open the tools website and create an account with your email. 2. free credits, how many, and when they renew. 3. Choose the cheap slash fast model, 16 colon 9 ratio, and the shortest duration. 4. Pace the prompt you built in Chapter 5. 5. Press Generate and Wait. 6. Watch the finished clip in full, not once, but two or three times, watch the hands, the faces, any lettering, the way things move. 7. Download it into your clip's folder under a meaningful name, even if the clip is perfect, it will serve for comparison. 8. Run the improvement cycle. Name the fault, refine the prompt, generate again. A few practical tricks. Generate two or three clips from the same prompt. AI gives a somewhat different result each time, and you pick the best. A clip whose main subject is right but whose edges carry faults can often be saved by cropping in the edit, do not discard it at once, and set yourself a daily usage limit before the credits run dry, or weeks credits will vanish in one enthusiastic evening. Exercise, make the clip for the first scene of your storyboard, running at least three improvement cycles, save the final clip together with the prompt that produced it. Chapter 7, From Enage to Video, The Road of Greater Control Text to video has won in convenience, you have no hold over the opening scene. Whatever the AI makes, it makes, the remedy is image of video, first prepare a still image that is exactly to to your mind, then tell the AI to set this image in motion. Hear the look of the scene, the character's face, the clothing, all are fixed in advance. The AI only adds the movement. Where will the image come from? Free sources. First, your own photographs, your village, your festivals, nature, art, animating a photograph you took yourself is the most authentic or out. Remember the photograph must be your own or used with the owner's permission, and before animating a photograph of a living person, be sure to take their consent. This is both courtesy and part of the ethics, described in chapter 15. Second. AI generated images. There are separate tools for image generation and most video tools include an image making feature. The image prompt follows the same six-part formula, only in place of camera movement, describe the composition. Images are cheap and quick to make. So run your improvement cycle on the image first. Generating 10 images and choosing the best is far cheaper than generating 10 videos. Third, scanning your own artwork. A work of Mathila painting if it is your own or you hold the rights can be scanned and given gentle motion. A fish stirring, the line ornament shimmering, exercise great restraint here, the dignity of traditional art lies in its stillness, so keep the motion extremely slight, right gentle, slow, subtle motion in the prompt. The method is simple. Choose the image video option in the tool, upload your image, and write the motion instructional on side. This prompt now carries less seen description and more motion description, what moves, in which direction, how fast, and what the camera does. For example, gentle ripples on the water, the egrets wings moving slowly, camera moving forward very slowly, nothing else changes. That last phrase, nothing else changes, matters, without it the AI will sometimes transform the whole scene. This is also the best remedy for the problem of character consistency. If the same character appears again and again in your video, first create or choose one excellent image of that character. And for every scene animate that same image with different motion instructions, some advanced tools offer a feature called character reference or consistent character where the character's image, given once, appears in every clip. Look for this feature in your tool. Exercise. Take any image or own photograph or AI generated and make three different motion versions of it. In one, only the camera moves, in the second, only some element of the scene moves, in the third, both, compare the three, which feels most natural. Chapter 8. Voice. Voice over and text to speech. The soul of a video lives not in the visuals but in the voice. A viewer will forgive a blurry scene, but will close the video at a bad voice, so read this chapter with care, and for speakers of mathily and other less-served languages, it holds some special advice. There are two ways to add voice, record your own, or have AI-generated text-to-speech, TTS for short. First, your own voice, because for mathily and languages like it, this remains the best out the pure pronunciation, the natural cadence, the rise and fall of feeling. No machine yet renders these as well as a native speaker and no studio is needed. A smartphone microphone today is quite good enough, follow a few rules, record in a quiet room fan off, windows shut, night or early morning is best. Hold the phone about a hand span from your mouth, speak standing or sitting upright. The voice stays open. Keep the script before you, but speak as if telling, not reading, and record paragraph by paragraph rather than all in one take, when you slip, you re-speak only that much. AI can polish a recorded voice, many tools, usually named enhanced voice, or studio sound in editing apps strip the background noise and give the voice a studio finish, use it without fail, the difference between a plain recording and an enhanced one will astonish you. Now text to speech figure six, here you type the text, choose the language and the speaker female or male, young or mature, adjust pace and pitch, and download an MP3 file for Hindi, English, and other major languages. This facility is very mature for mathily. The situation is improving. Some tools have begun to offer a mathily voice and tools built for Indian languages are the most likely place to find one. Search in your tool. If mathily appears, first test it with a short passage to hear how pure the pronunciation is. If no mathily voice is available, Two remedies. The first and best. Your own voice by the method above. The second. Making do with the Hindi voice. Write the text phonetically, listen and adjust the spelling until it sounds right. Keep the pace a little slow. Use short sentences. The result will not carry a fully mathal cadence, but it will serve. Remember, this is a compromise, not an ideal. Wherever feeling and purity matter poetry, stories, children's material, give your own voice. A word on a newer facility, voice-cloning, some tools, from a few minutes of your recording, build a digital replica of your voice, which will then read any text in your own tones, for content in a less-served language this is attractive, teach the tool your voice wants, and the voiceovers of many videos can be made, but two iron rules, clone only your own voice, imitating anyone else's voice without written permission is absolutely forbidden, and where a cloned voice is used in a video, disclosing it is good practice. The joining of voice and visuals will happen at the edit chapter 11, so keep the voice file separately in the voice music folder, make the voiceover first and the video clips after. This order is wise, because hearing the length of the voice tells you how many seconds each scene needs. Exercise. Make the voiceover of your script both ways, once recorded in your own voice, once through a TTS tool, listen to both with your eyes closed, which sounds more like you. Why? Chapter 9. Avatar Videos, The Digital Speaker. Imagine, every fortnight you must make an announcement video for a journal's new issue, or fifty lectures for a course. Camera, lighting, dress, recording, every single time. Impossible. The remedy is the Avatar video. A digital human sits on screen and speaks your written, text, lips moving, eyes blinking, hands gesturing, like a news anchor. The well-known tools of this class are Hei-jen, Synthesia, and others, and the class itself is growing fast. The structure is nearly the same in all figure 7. 1. Choose the avatar. Tools carry hundreds of ready-made avatars of different ages, dress and bearing. Some tools also let you build your own avatar from a photograph or a short video. That is, you remain on screen without recording each time if you make your own avatar. the same consent rule given for voice cloning, only your own likeness, never another's. 2. Give the script, it can take two forms, written text which the tool will speak through TTS, or your own recorded audio file which the avatar will lip sync, for mathily the second road is usually better, upload your own mathily voiceover, and the avatar speaks it, the lip sync you get is surprisingly good. 3. Choose the background and layout, a library, an office, a plain color, or an image of your own, choose the aspect ratio 16 colon 9 or 916, and generate. The beauty of the avatar video lies in its practicality, not in spectacle, new style presentation, journal announcements, lesson explanations, introductions of an institution, information videos, in all these it is excellent for the feeling laden delivery of story and poetry, a human is still better. One tip on presentation, do not keep the avatar on screen for the whole video without relief. In between, show related scenes, images or text cards called b-roll while the avatar's voice runs beneath. The video comes alive, this weaving happens at the edit, keep the avatar clip and the b-roll clips as separate files. And yes, when the avatar in a video looks human, the viewer has a right to know it is a digital speaker, write one line in the description, trust stays intact, and trust is a channel's real capital. Exercise, on the free plan of any avatar tool, make a 32nd introduction video, subject, an introduction to my village or an introduction to a favorite book, try both methods, type text and uploaded voice. Chapter 10. Music and sound effects. Visuals for the eye, voice for the ear and music, for the heart, the same scene feels lifeless without music and comes alive with the right score, but with music comes the The greatest danger of all, copyright, put someone's song in your video without permission and YouTube can block the video, others can claim its earnings, and the channel can be penalized, so rule one, never a famous film song or commercial recording, unless you hold written permission. Then where will the music come from? Free lawful sources. First, AI generated music, there are now tools that compose music from a written description. Write slow, tender, flute-led, 60 seconds and the music is ready. Some tools even build a full song, voice included, from your lyrics. Music made this way for your own video is generally safe to use, but read each tool's license once, especially whether commercial use including YouTube monetization is permitted. 2. Copyright-Free Music Libraries YouTube's own audio library inside YouTube studio is free and safe. Beyond it, many websites offer freely licensed music, some entirely free, some on the condition of attribution. If attribution is required, do not forget to write the musician's name in the video description. Third, your own recorded music. Methila has its own rich musical tradition, and so does every region. If you or someone you know, sings or plays, then a folk tune recorded by yourselves is the most authentic source of all, and it gives your video an identity no AI can. The tune of a folk's song is traditional, but a particular recording or arrangement belongs to its maker, keep this distinction in mind. The craft of laying music, under a voice of or keep the music low, 20 to 30% of the main voice, where there is no voice of or opening, close, scene changes the music may rise, let the mood of the music match the mood of the video, brightness for a morning scene, tenderness for a farewell, and at the end let the music sink away slowly fade out, music cut off abruptly jolts the ear. Sound effects are the small sounds, birdsong, the splash of water, the rustle of wind, they make a scene believable, some newer video models generate sound along with the scene native audio, if your tool has this, keep it on, if not, take sounds from a free library and add them at the edit. Exercise. background music of two different styles for your video, one AI generated, one from a free library. Play each behind the video in your mind's eye at least and consider which mood fits better. Chapter 11. Editing, Turning Clips into a Video Now you hold all the ingredients, video clips, voiceover, music, editing is the kitchen where these ingredients become the dish and the good news, editing skill, Once learned, serves in every video, it does not keep changing the way AI tools do. Choose an editor. Free editors exist for both mobile and computer. CapCite is at present the most popular and the simplest. On the computer, DaVinci Resolve is professional grade even in its free form. Choose either, the structure figure 8 is the same in all. The media area, where you bring and import all your files. The preview, where the video plays as you work. The timeline, The most important of all, the line of time on which the clips are arranged in order, the timeline has several strips tracks, one for video, one for voice, one for music, one for subtitles, stacked one above another, all playing together. Keep the basic order of editing thus. 1. First lay the voice over on the timeline, this is the spine of the video, the visuals will be arranged upon it. 2. Listening to the voice. Place the video clips in order, Let the scenes show what the words are saying, cut the clips, keep the best portion of each, remove the rest. AI clips are often awkward at the very start and the very end, trim both edges and the clip cleans up. Pre, add transitions, the manner of passing from one clip to the next, the rule, the fewer, the better, the plane cut is the purest, a light fade at a change of mood, thinning, twirling, color transitions are the mark of the the beginner. 4. Lay the music on its track and to bring its level down Chapter 10. 5. Color correction. Most editors have a one-click filter or enhance. If your AI clips came from different tools, put the same filter on all, the colors fall into step, and the video feels stitched of one cloth. 6. Add a title card at the start 3-4 seconds and a closing card at the end the channel's name, a word of thanks. Export settings. 1080p. MP4 format. format, 30 frames per second, the standard for YouTube, keep the exported file in the final folder. One suggestion after the video is done, watch it once from beginning to end without stopping, as a viewer, wherever your attention drifts, know that a cut is needed there, then show it to someone at home, the reaction of one first viewer teaches more than a hundred critics. Exercise, join all your clips, voice, and music into your first complete video, With title card and closing card, export it, show it to a family member and write down their first reaction. Chapter 12. Subtitles. A video that can be read. Most viewers today watch video without sound, on the bus, in the office, in bed at night. No subtitles, no viewers, and for content in a language like Mathalie the importance of subtitles is doubled. Subtitles in their original language build the habit of reading it, while Hindi or English subtitles bring in viewers who do not know the language at all, that is, your story travels the whole world. The standard format of subtitles is SRT, a plain text file in which three things repeat over and over figure nine, a serial number, a timeline from which second to which second, and the text, this file can be made and corrected even in an ordinary text editor notepad. But matching the timing by hand is laborious, and here AI helps again, two ways. First, Automatic Transcription, the auto captions feature in an editor such as CapCut listens to the video's voice and writes the subtitles itself. Timing included. In Hindi and English this is very accurate. A mathily voice it will usually hear as Hindi and write accordingly, then you correct the text it made, even so, correcting is far faster than writing from scratch, because the timing arrives ready made. Second, from the script, you already have the script written, give a chatbot your script and the video's total duration and say, divide this into SRT format, each subtitle at most two lines, at a comfortable reading pace, load the file in the editor and nudge the timings forward or back. The craft rules of subtitling, at most two lines at a time, roughly 32 to 40 characters per line, each text stays on screen at least one second, a sentence breaks where the meaning allows I went slash to the market, no, I went to the market together, centers in white with a light dark shadow or strip behind, so they can be read even over a bright scene, place them at the lower middle of the screen, but not so low that the real format cuts them off. On YouTube, subtitles can be given in two ways, burned into the video joint at export from the editor or upload it as a separate SRT file which the viewer can switch on and off, the best method, give the original language subtitles as a separate file, and at Hindi and English as separate SRT files too, a chatbot will help with the translation the duty of checking it remains yours, thus one video reaches the viewers of three languages. Exercise, make the SRT of your video in its own language by either method, then make its Hindi or English translation file, run both with the video and check, is the timing right, is any line too long? Chapter 13, Publishing on YouTube. The video is made, now carry it to the viewer. YouTube remains the largest and the most lasting platform, a video placed here keeps being watched year upon year, while on real format platforms a video's life is a few days, so make YouTube the main house, let reels and short speed its windows. Making a channel is simple, sign into YouTube with your Google account and create one, choose the channel's name with thought, short, easy to say, and suggestive of the subject, the channel picture logo and banner too can be made with an AI image tool. At upload time keep the checklist of figure 10 before you, a few points in detail. Title, the main matter in the first 3 or 4 words, with a title in your own language, adding Hindi or English in brackets helps the video surface in search, because seekers search in every language. Description, the first 2 lines are the most valuable. These appear in search results, write the video's essence here, key words included, below them, the chapter list what comes at which minute, the list of sources, and the channel introduction. Thumbnail The viewer sees the thumbnail before the title, the rule, one image, one feeling, at most three words, and words large enough to be read on a small mobile screen, make an attractive thumbnail with an AI image tool, but never a misleading one. That is in the thumbnail must be in the video where the viewer feels cheated and trust in the channel is gone. The AI disclosure, YouTube now expects that realistic looking AI generated or AI altered content be declared at upload in answer to the altered content question. This is not mere rule keeping, it is honesty with the viewer, for plainly imaginary styles such as animation the duty usually does not arise, but when in doubt, declaring is always the better course. Language setting. Choose the video's language. Mathalie is in YouTube's list, as are many others. This helps the video reach those searching in that language, and it strengthens the statistics of the language's content besides. After the upload, what then? Watch the response of the first hours and days, reply to comments, the early conversation waters the channel's roots, and keep regularity. One video a fortnight makes 24 in a year, and this bears more fruit than a hundred videos at random. attach themselves to a program, not to scattered surprises. For reels and shorts, cut the most engaging 30 to 60 seconds of your main video into the 9-16 ratio with the editor's reframe or crop feature, and right at the end, full video on the channel, this is the window that leads new viewers to the house. Exercise. Upload your video, completing all 7 points of the checklist. Also cut a shorts version and upload it separately. After one week, compare the figures of the two views, watch time. Chapter 14. Special Considerations for Mathalie Content This chapter is the heart of this book. The tools are universal, but our purpose is particular, immathalie, for Mithila, and readers working in any less served language will find the same principles applied to their own. First, purity of language, AI tools, when writing Mathalie, usually let the shadow of Hindi fall across it because they have learned far more Hindi. Trips, translations, subtitles made by a chatbot. Check every one with your own eyes, the plain rule. The A.I.s Mathili is a draft, not an authority, where in doubt, trust your ear, speak the line aloud, whatever grates on the ear is the shadow of Hindi. Second, the question of script. Mathili is written in Devanagari and it also has its own ancient script, Turhuta Mythalikshara. the video subtitles and text cards in Devanagari, the most people will be able to read them, use Turhuta for beauty and identity, entitle cards, in the logo, in a watermark, thus the script stays before the eye, and curiosity awakens too, remember, AI image tools cannot yet write Devanagari or Turhuta letters correctly, always add written, text yourself at the edit, never have the AI write it. Third, authenticity of the visuals. Tell an AI and Indian village and it will produce a generalized North Indian scene, which is not Mithila. Give the prompt Mithila's particular signs. The pond, the banyan, the mango orchard, the patty field, fish, pond, the thatched house, the Tulsi platform in the courtyard, wall ornament in the manner of Arapan. Even then, what comes will be Mithila like, not Mithila, so wherever possible, blend in real photographs and footage by the method of Chapter 7, the mixture of AI scenes and real scenes gives the most authentic result of all. Fourth, the honor of Madhubani's slash Mathila painting, the AI can be told to generate in Madhubani style, and it will imitate the colors and the line, but pause here and think. Mathila painting is a living tradition, the livelihood of thousands of artists rests on it, and each of its manners Barney, Kachni, Godna, Gober carries its own lineage, and AI made Madhubana like image lifts the tradition's appearance without its labor and its knowledge, my counsel, in your videos show real works by real artists with permission and credit. This honors the artist and strengthens your video at once. If you do use AI ornament in the Madhubani manner, say plainly that it is an AI mediation, not authentic Madhubani. 5. An inexhaustible store of subjects The field of mathily video is still nearly empty. Whatever makes makes first, some directions, children's material songs, tales, letters, the child audiences, the fastest growing of all, festival explainer C. H. H. I., Sama Chaikpa, Jersital, Huat, Why, How, recipes, folk tales and the verses of Vidya Patti presented with images, the vocabulary of village and home a visual dictionary of the words now slipping away, introductions to the places of Mathila, each direction could be a channel in itself. the last word, the patience of quality. In Methili the audience will be smaller than in Hindi. This is natural, but the loyalty of the Methili viewer is greater. They will comment, they will share, they will return again and again, look not at numbers but at relationships. A hundred devoted viewers are worth more than 10,000 indifferent ones. Exercise. Choose one of the six directions above and plan three consecutive videos on its subject plus a one-line summary each. 3. Because one video is an experiment, 3 are a direction. Chapter 15. Copyright, ethics, and responsibility. With a powerful tool comes responsibility. This chapter is short, but bring its every line into practice. Copyright, the root principle, what you did not make, you do not use without permission, film songs, portions of others' videos, the text of books, other people's photographs, The rule covers them all. Everyone does it is no argument. YouTube's automatic system content ID catches it, and the channel bears the penalty. In freely licensed material too, read the conditions, one says a tribution required, another no commercial use. No also the question of rights over your own AI made material. In most tools terms, permission for commercial use of generated video and images comes with the paid plan, and is sometimes limited on the free one for any tool you work with regularly, read its terms of use once, especially the ownership and commercial use sections. The dignity of persons, three iron prohibitions, no imitation of anyone's face or voice without written consent living or departed, both. Never show a person in a scene that lowers their honor, nor put in their mouth words they never spoke the deep ache, and take special care with the images and voices of children before posting a video even of your own child. Consider that it is becoming public. Toward truth. AI can make scenes that look real and therefore it can deceive. The rule. Let fiction be called fiction and news be news. When you make a video on a historical or cultural subject, check the facts against your own trusted sources. An AI chatbot speaks its errors with full confidence this is called hallucination. Wherever an AI scene looks real, declare it chapter 13. The viewers trust takes years to earn and one video to lose. And toward yourself. Let AI's convenience never turn into laziness. Take work from AI, not thought. The day you hand the script, the choices, and the final judgment over to AI, that day the video ceases to be yours, and viewers can smell the difference. Keep the tool in your hand. Let not the hand become the tool. Exercise. Run a responsibility check on the first video you made, the music's license, permission for the image sources, the facts verified, the day I disclosure, keep the four answers in writing. Repeat this exercise with every video, within days it will become second nature. Chapter 16 The Practice Project, a first video from start to finish. Now all the threads in one place, this chapter is the model of one complete project, a 90 second video called CHathai, the great festival of sun worship. it, then repeat the same sequence on a subject of your own. Day 1. Planning Chapters 3, 4, Purpose, to explain simply the spirit and the observance of Siddhi Chathai, audience, Mathal families living away from home, the new generation especially, duration, 90 seconds, that is, 9 or 10 clips, a two column script drafted by a chat pot, corrected by hand in three places, local detail added to the description of Nahikai, the feeling of the final line deepened. Ten scenes on the storyboard, the river ghat at dawn establish, the courtyard where thequeua is being made, the supandora of the offering, the evening argia, the morning argia, the parent, and the closing card resolve. Day 2 Voice Chapter 8, the voiceover recorded in my own voice, at night, in a quiet room, paragraph by paragraph, polished with the editor's enhanced voice, total 82 seconds, which means the scene plan sits right. Phase 3-4, visuals chapters 5-7, a six-part prompt built for every scene thought out and mathily, translated into English. Methylas signs in each, a river ghat in north Bihar, banana stems, bamboo kania, fruit in the soup and aura, three clips came out well at the first attempt, four clips took two or three improvement cycles. For two scenes the tekua and the soup the image of video method was used, first the image perfected, then subtle motion given. And in one scene a real photograph of my own was used, the crowd at the GAT, which no AI could ever have made. This mixture is the essence of Chapter 14. Day 5. Music and Assembly Chapters 1011. AI Music. Slow, devotional, flute and emridang, 90 seconds, the license checked, the editing order, voice first, then clips, plain cuts, a fade in two places, one and the same light warm filter on every clip, on the title card CHathai in Terhuda and the full title in Devanagari. Day 6, subtitles and publication chapters 12-13, the Mathili SRT from the script, then Hindi and English translation files translated by Chatbot, checked by hand. The thumbnail, the scene of the evening argh, three words, hui chathai, ad upload, language Mathili, the AI disclosure switched on, music credit and sources in the description. Alongside, a 35-second shorts version. Day 7 Review Chapter 15, the responsibility check on all four points shown to the family, the suggestion came that the morning argia scene feels short, noted down for the next video. See, seven days, an hour or two a day, and from nothing to a published video, the first project will take longer than this, let it, by the third or fourth video the sequence becomes habit and the time falls by half. A last word. This book ends here, but the learning begins here. The tools will change, the models will change, the screens will change, but the framework you have learned plan, prompt, improvement cycle, assembly, responsibility will stand. Now go and tell the story of your Mathila, in your own voice, in your own language. The Pentaxe, Lausary. AI artificial intelligence, the capacity of a computer to learn and understand in a human-like way. Generative AI, AI that creates new material, text, images, video, voice. Prompt. The written instruction given to an AI. Negative prompt, the list of what is not wanted. Text to video, the method of making video from a written description. Image of video, the method of setting a still image in motion. TTS text-to-speech, the method of producing a spoken voice from written text. Avatar, a digital speaker who delivers text on screen. Lip sync, the matching of voice and lip movement. Voice cloning, making a digital replica of a voice. Credit, the usage currency of AI tools, each generation deducts some. Model, a particular version or engine of an AI. Clip, a short segment of video usually 5-20 seconds. Aspect Ratio, the ratio of a screen's width to its height, 16 colon 9 wide, 9 16 upright. Resolution, the fineness of the image, 1080p standard, 4K extra fine. Watermark, an identifying mark printed on content. Storyboard, the scene by scene sketch plan. B-roll, supporting footage beyond the main speaker. Timeline, the line of time in an editor on which clips are arranged. Track, a layer of the timeline, video, voice, music and so on. Cut, trimming a clip, or passing directly from one clip to the next. Transition, the manner of joining two clips fade and the rest. Fade in slash out, slow emerging slash, slow vanishing. Render slash export, turning the edited video into the final file. SRT, the standard file format of subtitles. Burned in, subtitles fixed permanently into the video. Thumbnail, the face image of a video. Hallucination, an AI's confident error. Content ID, YouTube's copyright checking system. Deepfake, a deceptive imitation of someone's face or voice. Monetization, earning from videos. Appendix B, a collection of prompts. The lower 10 ready prompt frames, fill the brackets with your own details and shape the final prompt by the method of Chapter 5. 1. Village Dawn A village in North Bihar, Dawn, light mist over the fields, thatched houses, a bayon tree in the distance, your detail, camera panning slowly to the right, soft golden light, realistic documentary style. 2. Pond Scene A pond covered with lotus and water lilies, seasoned, on the bank tree slash gat. Gentle ripples on the water, subtle movement of bird slash fish, camera moving forward very slowly. Time of day light, semantic. Free, courtyard scene, the courtyard of a mythillahome, a tolcy platform, action, say, food cooking on the hearth, folk ornament on the wall, warm intimate light, close up, realistic. For, festival scene, preparations for name of festival, central object, say, soup and dora, lamps, courtyard or get, close ups of busy hands, festive joy, warm light, documentary style. 5. Nature scene, green paddy fields waving in the wind, time, the far horizon, bird in flight, drone view rising slowly, natural light, ultra-fine detail. 6. Object study dish slash craft, an extreme close-up of object, background, light steam or sheen, camera circling the object slowly, soft studio-like light, appetizing slash appealing style. 7. Doreen illustration animation. Character description doing action. Place. Hand drawn animation style, soft watercolor, like a children's book illustration. Slow pleasant motion. 8. Historical imagination. An imagined scene of period methela. Action. Colors like an old palm leaf painting, slow solemn camera, museum documentary style. The AI disclosure is obligatory here. This is imagination. 9. Title card background, an abstract ornamental pattern drifting slowly, color scheme. Say, for Millie and Red and Yellow, no letters, no figures, come even motion, 10 second loop, add the lettering yourself at the edit. 10. Motion instruction for image video, in this image, gentle natural motion in element, say, water slash leaves slash cloth, camera static slash moving forward very slowly, nothing else changes, the image's original color and line remain intact. The Pentax C, a guide to the tools as of 2026. This list reflects the state of things at the time of writing 2026, names and features keep changing, learn the categories, do not memorize the names, before adopting any tool, check three things, what the free plan gives, whether commercial use of the output is permitted, and where the watermark stands. Text-to-video slash image video, Google's Vivo, OpenAI's Sora, Kling, Runway, Pixverse, Seedance, Haleuo, Luma, Pika, Realistic Scenes, Cinematic Clips, Avatar video, Hey-Gen, Synthesia, Digital Speakers, LipSync, Many Languages. Text-to-speech and voice cloning, 11 Labs and others, services centered on Indian languages are also growing, a mathily option will appear there first, keep watching. AI music, Suno, Udio and others, music and song from a description or from lyrics. Image generation, included in nearly every video tool, separately, Mijirni, and the image features attached to the chatbots. Editing, Cap Cut, Mobile and Computer, simple auto-cactions included, DaVinci Resolve, Computer, Free, Professional Grade, Canva, Thumbnails, Banners, Simple Video. Call in one plant a video, tools of the in-video kind, which, given a topic, join script, visuals and voice on their own, good for quick work, but your own stamp is faint in them, so for learning, the step-by-step method of this book is better. Chatpot script, translation, SRT, brainstorming, Claude, chat GPT, Gemini. And last of all, this book too will age with time, but you will not, the teach yourself temper this book has built in you will apply to every new tool when you seal a new tool ask three questions where is the prompt written what is in the settings where and how does the result arrive and within five minutes you will be working it.

← Transcript index