Only 15% of 1,274 AI videos are vertical — and short-drama clips rank 16th of 18
For the past few weeks we have been doing something tedious: indexing 1,274 publicly shared H3-generated videos one at a time, tagging each with genre, shot technique, visual style, aspect ratio, duration and generation mode.
One result we did not expect: at a moment when every short-form platform is vertical-first, only 15% of these public examples are vertical — and short-drama clips are among the least vertical of any genre.
This piece pulls out the aspect-ratio numbers on their own, because it is the dimension with the sharpest contrast and the most direct bearing on day-to-day work.
Where the data comes from
The 1,274 cases were collected from public sources, and we kept the original publication link for every one. Aspect ratio was read from the video file itself, not inferred from the poster's description. Genre and technique tags were generated as automatic candidates and then reviewed by a human; anything not yet reviewed keeps an "unreviewed" marker and is excluded from the genre statistics below.
Finding 1: Landscape leads 5.6 to 1
| Aspect | Cases | Share |
|---|---|---|
| Landscape | 1,060 | 83.2% |
| Vertical | 190 | 14.9% |
| Square | 24 | 1.9% |
Vertical plus square comes to 214 cases, or 16.8%. That in itself is unsurprising — AI video tooling and viewing contexts have centred on landscape for a long time. The surprising part is next.
Finding 2: The most vertical genres are not narrative ones
First, where do the vertical cases cluster? (A single video can carry more than one genre tag.)
| Genre | Vertical cases | Share of all vertical |
|---|---|---|
| Product ad | 30 | 15.8% |
| Character performance | 27 | 14.2% |
| Fashion & beauty | 26 | 13.7% |
| Short-drama dialogue | 24 | 12.6% |
| Visual design | 23 | 12.1% |
Read this table alone and short drama sits in the top four, which feels intuitive. But switch to each genre's own vertical rate and the conclusion flips:
| Genre | Vertical rate | Sample |
|---|---|---|
| Visual design | 30.3% | 23 / 76 |
| Fashion & beauty | 27.7% | 26 / 94 |
| Tool comparison demo | 22.7% | 10 / 44 |
| City & architecture | 18.3% | 17 / 93 |
| Character performance | 18.1% | 27 / 149 |
| Product ad | 15.6% | 30 / 192 |
| … | ||
| Short-drama dialogue | 10.7% | 24 / 224 |
| Animation & games | 8.4% | 7 / 83 |
| Action & physics | 7.1% | 5 / 70 |
Short-drama dialogue has 224 cases — more than any other genre — but a vertical rate of 10.7%, ranking 16th out of the 18 genres with enough samples to judge. In other words, the short dramas in this public corpus were overwhelmingly shot landscape.
And the two most vertical genres — visual design and fashion & beauty — are both *display* content, not *narrative* content. In this dataset, vertical framing behaves more like a product-showcase format than a storytelling format.
Finding 3: Clips with speech are almost all landscape
We looked separately at the 251 cases where speech was detected:
| Aspect | Cases | Share |
|---|---|---|
| Landscape | 222 | 88.4% |
| Vertical | 24 | 9.6% |
| Square | 5 | 2.0% |
Among clips with dialogue, landscape is *more* dominant than in the overall set. And the median duration of vertical and landscape dialogue clips is exactly the same — 15.2 seconds both. So the common assumption that "vertical clips are shorter" does not hold here.
Speech-bearing cases concentrate in short-drama dialogue (109), character performance (40) and music & dance (30).
Finding 4: Duration piles into 8–16 seconds
| Duration | Cases | Share |
|---|---|---|
| 8–16s | 1,101 | 86.4% |
| 0–8s | 92 | 7.2% |
| Over 16s | 81 | 6.4% |
86% lands in a single bucket. Short-drama dialogue has a median of 15.2s and a mean of 16.2s — pinned right at the upper edge. That suggests the bucket's upper bound is the model's reliable single-generation limit, not a duration creators actively chose. (We took that "15 seconds" apart in a separate piece; it is more extreme than it looks here.)
What you can do with this
If you make vertical short drama (Douyin, TikTok, Reels): this public corpus is close to useless as direct reference, because 89% of the short-drama samples are landscape. What you need is vertical composition and vertical camera craft, and the closest public material is visual design and fashion & beauty — genres whose vertical rates are three times higher.
If you are evaluating model capability: the narrow 8–16s bucket matters more than the aspect-ratio gap. It means the practical usable unit of these models is roughly a dozen seconds; anything longer has to be built by chaining segments, not by asking for more in one go.
If you make product ads: a 15.6% vertical rate is in line with the overall figure, meaning this genre has no strong framing preference. Decide by placement channel.
Method and limitations
- Sampling bias: only publicly released cases enter the statistics, skewing toward better-looking work. Not representative of all creators.
- Tag provenance: genre and technique tags are automatic candidates plus human review. 1,212 records are still marked "searchable, pending semantic review"; unreviewed markers are excluded from the genre statistics.
- Aspect ratio and duration come from the video files themselves — objective measurements, with error limited to container metadata accuracy.
- Deduplication scope: the ratios above use all 1,274 index records. The cross-host duplication problem is covered in a separate piece; after de-duplication the aspect split (82.5% landscape / 15.3% vertical / 2.3% square) points the same way.
- Correlation is not causation: this piece reports distributions only. "Short drama is landscape because the publishing venue is Twitter" is a hypothesis, not a finding.
Reproduce it
All 1,274 cases — with original links, aspect ratio, duration and tags — can be filtered and checked case by case at https://cases.aishifu.shop/. Every number here can be reproduced with the filters there.