The 2 Big Challenges of Making Manga with AI
The biggest challenges of producing manga or illustration series with AI are character consistency and panel layout (page composition). Unifying all panels with a consistent look by prompts alone is currently hard; you need to combine several techniques.
Character-Consistency Methods
1. Reference Image
- Nano Banana 2 (Google): up to 6 reference images in total, no training needed
- Midjourney V7: Omni Reference, tuned with --ow (the older --cref applies to legacy versions only)
- SD's IP-Adapter / Reference ControlNet
- Video: Runway Gen-4.5 reference features, with comparable controls in Google Veo, Kling and Luma. OpenAI's Sora is discontinued (apps shut down 2026-04-26, Sora 2 API stops 2026-09-24), so it is not an option.
Zero-shot reference models now cover most short works, so try reference images before you train a LoRA.
2. A Dedicated Model via LoRA
Making a LoRA trained on roughly 15-20 images of the same character lets you produce a stable art style every time. Many LoRAs are also published on Civitai.
3. Fixed Seed + Minor Diff
Fix the seed number and finely tweak the prompt. Effective when you want to change only expression/pose.
4. Inpaint
Regenerate just the face separately. Photoshop's generative fill or SD's Inpaint can be used.
How to Make Panel Layout
Tools such as Jenova, Anifusion and COMICPAD now generate a whole page at once from a story: panel layout, art and speech-bubble placement, including right-to-left reading order. For work you want full control over, a division of labor is still the stable approach:
- Generate each panel's art separately with AI
- Page layout in Photoshop / Clip Studio / Figma
- Manually place borders, speech bubbles, sound effects
- Dialogue: AI (plot creation) + humans (revision)
Line Art / Inking
- Stable Diffusion or FLUX.1 + ControlNet (Canny, Lineart) for rough-to-lineart
- Quality improvement with Magnific AI
- Clean lineart tone with FLUX Kontext
- Clip Studio AI features for manga lineart
Screentone/Effects
- Apply tone to AI-generated art with Clip Studio Paint's tone feature
- Monochrome + halftone processing in SD
- Focus lines/sound effects by hand or using assets
Story/Dialogue
LLM use is effective at the plot stage.
- Plot creation (kishotenketsu) with Claude / GPT
- Fix character settings/speech in Project / Custom Instructions
- Have it produce dialogue ideas, humans select
- Translation (multilingual deployment) also uses LLMs
Cautions for Commercial Deployment
- Character imitation (training on existing manga characters) is a copyright risk
- Use commercial-license AI (Adobe Firefly, Getty)
- Tendency for the market to require "AI use" disclosure
- Even in doujin activity, mass-generating others' characters with AI needs care
Example Production Workflow
- Plot + dialogue first draft with Claude
- Train the protagonist's LoRA (about 15-20 images), or skip training and use reference images for a short work
- Generate each panel's art with SD + LoRA + ControlNet
- Page layout/speech bubbles in Clip Studio
- Effect lines/fixes by hand
- Generate translated versions (EN/CN) with AI, native check
Summary
AI manga production's realistic answer is the division "art = AI, layout = human, dialogue = AI+human." Guarantee character consistency with a LoRA/Reference/Inpaint combination and finish panel layout in DTP tools like Clip Studio—an era where even solo creators can efficiently make long manga.



