The AI Canteen is where cutting-edge AI meets editorial craft — a place to taste what's possible and leave knowing how it was made.

Chioma Onwumere
Founder & Creative Director
The AI Canteen began with a simple conviction: that generative tools, in the right hands, could produce work with real soul — cinematic, editorial, and unmistakably intentional.
Founded by Chioma Onwumere, the studio blends the instincts of a director with the fluency of a prompt engineer. The result is a body of work — films, photography, and characters — that feels handcrafted, even when a machine held the brush.
But a canteen is meant to be shared. Alongside the work, we publish the prompts, the frameworks, and the hard-won lessons — so you can pull up a chair and make something remarkable of your own.
Come hungry, leave smarter. That isn’t a tagline — it’s the whole idea.
We would rather ship one unforgettable frame than a hundred forgettable ones. Every piece is curated, not churned.
The tools are extraordinary, but taste is still the differentiator. We direct the machine the way a filmmaker directs a set.
Everything we learn becomes a prompt, a tip, or a resource. The canteen is a table, not a gate.
A working toolkit refined across 20+ personal projects — with the exact job each tool does best.
Thinking partners for scripts, research, and prompt refinement — each with the model it runs on today.
The studio’s default writing partner. GPT-5.5 reasons carefully, follows a brief closely, and handles scripts, structure, and prompt refinement — with native image generation, voice mode, and custom GPTs for workflows you run again and again.
Anthropic’s assistant, prized for natural, human writing and its faithfulness to a detailed brief. The Claude 5 family — led by Opus 4.8 — excels at long-form structure, nuanced tone, and careful editing, for copy that reads written, not generated.
Built into X with a live pulse on what’s trending right now. Fast, current, and candid — ideal for catching the cultural moment, drafting scroll-stopping hooks, and pressure-testing ideas, now with native video input and a 1M-token context.
Google’s multimodal flagship with a huge 1M-token context and native understanding of images, video, and audio. Best for deep research, analysing reference material, and Search-grounded fact-checking — and it plugs straight into Google’s creative stack.
Where a still frame becomes a signature look.
OpenAI’s reasoning-powered image model. It plans a shot before it draws it — near-perfect in-image text, reliable composition, and up to 4K output.
The mood-maker. Painterly, art-directed aesthetics with gorgeous light — now faster, with HD 2K renders and sharper small-detail retention.
Black Forest Labs’ photoreal workhorse. Best-in-class realism for people, products, and editorial covers, with dependable multi-reference character consistency.
Motion, camera, and performance — generated.
Kuaishou’s cinematic model with native 4K, up to 15-second clips, and an AI Director that composes multiple shots in one generation while holding continuity.
The director’s toolkit. Top-rated fidelity with consistent characters and locations across scenes, native audio, and precise creative control.
Google DeepMind’s premium model. High-fidelity, coherent cinematic shots with synchronised dialogue and sound, up to 4K.
ByteDance’s expressive motion model. Strong on stylised movement and performance, and able to blend many image, video, and audio references in a single pass.
Voice and sound that carry the story.
The most expressive voice model. Natural voiceover, narration, and character voices in 70+ languages, with emotion, direction, and multi-speaker dialogue via inline audio tags.
Where everything is assembled and finished.
Fast, capable editing where it all comes together — assembly, captions, sound, and export tuned for every platform.