There's a TikTok format generating millions of views with minimal technical complexity. It uses still images, not video. It requires no face, no camera, and no advanced editing. The RPM is high—up to $4 per thousand views in some cases. Accounts using this format have launched within the last month and already accumulated tens of thousands of followers with monetization active. This guide breaks down exactly how the niche works, the formula behind it, and the complete system for producing these videos from start to finish using AI.
The niche is educational animation—but not complex animation. It's a slideshow of still images that explain four facts or phenomena from a specific domain. The content is informative, surprising, and designed to stop the scroll. A typical video covers four related subjects, each illustrated with a generated image, with a voiceover explaining the concept.
The psychology behind the format is straightforward. When someone sees a title like "Four Diseases You Can See on Someone's Face" or "The Most Dangerous Jobs in the World," they stop scrolling. The promise of learning something surprising in under a minute is compelling. The four-subject structure gives the video pacing—each new subject resets attention and keeps retention high through the full duration.
Accounts using this format have achieved remarkable results. One account posted its first video on May 15, 2025—roughly one month ago at the time of recording—and has already accumulated 93,000 followers and full monetization. Its videos routinely reach millions of views. Another account using the same format but targeting different subjects reached 35,000 followers with its first video hitting 710,000 views.
The format is deceptively simple. The biggest mistake is copying the exact subjects that successful accounts are already covering. If someone is already making videos about psychological disorders or bodily functions, creating identical content puts you in direct competition with an established account that the algorithm already favors. Copying subject matter is the fastest way to fail.
The correct approach is to extract the format—the structural formula—and apply it to entirely new subject areas. The format is: four related, surprising subjects from a single domain, each explained with a still image and voiceover. The domain can be anything: technology failures, famous betrayals in history, movie production disasters, fast food scandals, automotive engineering breakthroughs. The subject changes; the format stays the same.
A prompt system has been developed that automates the entire process. Given a domain, it generates four high-potential subjects, writes the complete voiceover script for each one, and produces the exact image generation prompts needed to create the visuals. The system handles subject selection, script writing, and visual direction in one workflow.
The prompt system is designed to work with Google Gemini. When you provide a domain—for example, "technology product failures"—Gemini outputs four specific subjects that have viral potential. For each subject, it generates a complete voiceover script and the corresponding image prompts.
The four-subject structure is critical. Each video covers exactly four subjects, giving the content a predictable rhythm that retains viewers. The script for each subject is written to be engaging, informative, and paced for the short-form format.
After generating the four subjects and scripts, the system also provides the image generation prompts. These are designed to create consistent, high-quality still images that match the tone and content of each subject.
Take the script into Google AI Studio. Navigate to Text-to-Speech, select Gemini 2.5 as the model, and choose the Everyday Assistant voice profile. This voice model handles long-form text without the cutoff issues that plague other voice models—scripts of any length render completely.
Select your preferred voice (male or female) and generate the audio. Download the finished voiceover file.
The prompt system outputs image prompts specifically formatted for the subjects in your script. Take these prompts to an image generation platform—Google's ImageFX or any tool running Nano Banana works well. Set the aspect ratio to 16:9. Generate images for all four subjects, maintaining visual consistency across the set. Download all generated images.
Import all images and the voiceover audio into Canva. Create a new mobile video project. Place the images in sequence, matching each to the corresponding section of the voiceover. Apply simple animations—subtle zoom or pan effects prevent the video from feeling static. Time the image transitions to match the script pacing.
Export the completed video and publish to TikTok.
The prompt system can generate unlimited subjects within any domain. When you need new ideas, ask it to propose different niches that work with the same four-subject format. The system can generate subject sets for automotive attraction, film production disasters, fast food scandals, Broadway production stories, and countless other domains. Each new domain is a fresh content category with no direct competition from accounts already using this format.
This format is currently in a high-growth phase. The accounts using it are new. The RPM is high because the content attracts an engaged, English-speaking audience that watches full videos. As more creators discover and copy the format, competition will increase and the algorithm will become saturated. The window for establishing a presence with this specific format is open now. The technical barrier is low—still images and a voiceover require no filming, no editing expertise, and no on-camera presence. The entire production process, from idea to published video, can be completed in under an hour.