For the past six months the focus has been on photorealistic cinematic production — perfecting character consistency, shot composition, and prop integration. We made significant progress. The Soul ID pipeline is solid. The Shinjitsu Standard gives us provenance. The photorealistic work is at a level I'm confident in.
Today I decided to move into new territory. Animation and stylised CGI character production.
THE PROBLEM WE WALKED INTO
The assumption going in was that our existing prompting logic — the layered structure we built for photorealistic work — would translate across to CGI character generation with some modifications. It doesn't. Not cleanly.
The first test outputs were acceptable at a surface level but flat. The characters read as CGI but lacked the stylised proportion and character design language that makes a CGI character feel intentional rather than generated. We were producing something that looked like a render but not a character.
So we started fine tuning. Facial description, body proportion, render engine variants — Octane, V-Ray, Blender Cycles. We built out the full 10-layer CGI prompt structure and documented it in MIRAI_CGI_Character_Pipeline_v1_0.md. The structure is sound. The problem was something else.
We spent the morning trying to push Ziga — our AI-native virtual human — into a stylised CGI direction. MAPPA anime facial proportions, hyper-elongated fashion-illustration body, small head, long legs. Every iteration fought us back toward a realistic human. The model kept defaulting. At one point she looked like a drug addict. At another point a goldfish. We cracked the facial description eventually — the key was referencing the flat planar geometry of East Asian facial structure without naming ethnicity, letting the geometry descriptors do the work. But the body proportion was a different problem entirely.
The 1-3-6 ratio block helped. Waist width anchored to head width helped. Explicit negative prompts blocking realistic proportions helped. But it was still a fight every generation.
THE BREAKTHROUGH
I stepped back and went to first principles.
The question I asked: what if the proportion problem isn't a prompt problem at all?
I downloaded a 2D character sketch — flat illustration, the kind a concept artist would draw — and fed it to Nano Banana Pro with a single line instruction to render it in Octane style. One pass. The output was exactly the sketch, rendered in full 3D CGI. Proportions intact. Character geometry intact. The model didn't interpret anything. It executed a material pass on geometry that already existed.
I tested it across multiple sketches. Same result every time. One instruction. No proportion drift. No fights.
The insight: AI models anchor to what they can see, not what they are told. When you describe proportions in text, the model has to interpret your description against its training data — which is overwhelmingly real humans. It will always drift toward what it knows. When you show it a sketch, there is nothing to interpret. The geometry is given. The model's only job is to render it.
THE SECOND DISCOVERY
I pushed further. What happens when you combine a sketch with a Soul ID?
The Soul ID — our face-locking pipeline — forces face consistency onto whatever geometry it's given. Feed it a sketch with stylised proportions and it locks Ziga's face onto that geometry. The result is a character that is simultaneously stylised in proportion and consistent in identity. Something that previously required multiple generation passes and constant drift correction collapsed into a single clean output.
Then animation. We tested free models on the rendered sketch output. It works. The stylised character animates.
This is where the real implication becomes clear. AI animation models are largely trained on real-world footage. When you ask them to animate something surreal — hyper-proportion characters, impossible spaces — they have no reference and collapse back to reality. The sketch breaks this. It gives the animation model a constructed reality to work from. The sketch is the anchor point that the training data doesn't have. It doesn't reach for reality because you've given it something else to hold onto.
WHAT THIS ACTUALLY MEANS
We have just identified a production pipeline for generating stylised animation that doesn't exist in the internet's training data. Which means:
- The model cannot copy it — it has to construct it from our inputs
- The illustrator becomes the most important person in the pipeline — not because they're drawing the final output, but because their sketch defines the geometry the AI executes
- Prompt engineering for proportion becomes a fallback, not a primary method
- The pipeline is: sketch → Octane render → Soul ID face lock → animation
This also validates what the industry has documented but not solved — stylised animation and realistic animation are fundamentally different problems for AI models. A model optimised for photorealism collapses on stylised geometry. The sketch-to-render method bypasses this entirely by giving the model a constructed reality to execute rather than asking it to invent one from text.
On the illustrator question — this needs to be stated accurately. AI is already displacing illustrators at the commodity end: stock imagery, generic advertising, template-based work. Entry and mid-level positions in those categories are under real pressure and that is not reversing. But at the creative geometry level — character design, proportion, silhouette, the decisions that make a character feel intentional — illustration is not being replaced. It is becoming the critical input layer that determines everything downstream in a stylised AI production pipeline. Those are two different markets that share a name. The commodity illustration market is shrinking. The character geometry market for AI pipelines is about to grow.