I've spent the last two weeks generating the same five prompts over and over, in five different tools, until my browser history looked like evidence in some kind of AI hoarding investigation. That's genuinely the only way to answer the question I keep getting asked: is Midjourney still worth it, or has Flux quietly taken over? What about Ideogram? Does Stable Diffusion even matter anymore in 2026?
My name is Artem, and I run the Writingmate blog. I've been testing image models since the first Stable Diffusion checkpoints made every generated hand look like it had survived an accident, so I've watched this category go from "impressive but broken" to genuinely hard to fault. That's exactly why spec sheets don't cut it anymore — every model claims 4K output and "photorealistic" quality. The only way to know which one wins your specific job is to run the same prompt through all of them and look at the actual pixels.
So that's what I did. I picked five prompts that map to real work — a product shot, a text-heavy poster, a consistent character across multiple images, a photorealistic portrait, and a stylized illustration — and ran each one through Midjourney v7, Flux 2 Pro, Ideogram 3, and Stable Diffusion 3.5. Then I ran the identical prompts through Writingmate's built-in text-to-image tool to see how it actually stacks up against going straight to the source. No cherry-picking, no "the tenth attempt was the best one" nonsense — first generation, default settings, same prompt text pasted into each one.
How I Ran This Test
Here's the thing about most "AI image generator comparison" posts: they show you one lucky image per model and call it a day. That tells you almost nothing, because a single generation doesn't reveal how a model behaves when your prompt stacks three or four requirements on top of each other. So I scored every output on four criteria that actually matter when you're shipping something real:
- Prompt adherence — did it follow the specific details (angle, lighting, composition, count of objects) instead of just the vibe?
- Text rendering — when the prompt asked for legible words on the image, did they come out spelled correctly and placed sensibly?
- Character consistency — across three separate generations of the same character, did the face, outfit, and proportions actually hold?
- Cost per image — what you're actually paying once you account for subscriptions, GPU hours, or per-image API pricing, not the marketing number.
One honest note before the results: Midjourney has no public API — it only runs through Discord or its own web app — and Ideogram and classic Stable Diffusion checkpoints aren't in Writingmate's model catalog either. So for those three, I tested through their native apps and compared the output quality against what Writingmate's generator produces with the models it does host, which currently includes FLUX.2 Pro, Nano Banana Pro, GPT Image 2, and Recraft V4.1, among dozens of others. That distinction matters more than most comparison posts admit, and I'll get into exactly where it helps and where it doesn't.
Test 1: The Product Shot
The prompt: "A minimalist glass skincare bottle on a wet dark stone surface, dramatic side lighting, soft shadow, small label reading 'LUME' in a clean sans-serif font, e-commerce product photography style."
This is the bread-and-butter job for anyone running a Shopify store, and it's brutal on models because it demands photorealistic materials, a specific lighting direction, and legible micro-text at the same time. Flux 2 Pro was the clear winner here — the glass refraction and the wet-stone reflections looked like they belonged in a real studio shoot, not a render. Midjourney v7 produced the moodiest, most "campaign-ready" version of the three, but it drifted the light source to the top instead of the side I asked for, which is a real prompt-adherence miss if you're matching a brand's existing shots. Ideogram 3 got the label right on the first try, which is worth something when you're iterating fast. Stable Diffusion 3.5 needed a second pass with a tighter negative prompt before the glass stopped looking slightly plastic — it's capable, but it wants more hand-holding than the closed models to hit commercial polish on the first attempt.
Running the same prompt through Writingmate's FLUX.2 Pro option produced output I couldn't tell apart from the standalone Flux app — same model, same weights, same result. If your work leans product photography, that alone is a reason to keep one subscription instead of two.
Test 2: The Text-Heavy Poster
The prompt: "A bold concert poster for an indie band called 'Night Radio,' halftone texture, warm orange and teal palette, large headline text 'NIGHT RADIO' in a chunky retro font, smaller line below reading 'Live at The Hollow, October 3rd.'"
This is where the gap between models stops being subtle. Ideogram 3 rendered both lines of text correctly, in a font style that actually looked chosen rather than generated, on the very first attempt. Independent benchmarking backs up what I saw: Ideogram sits around 90% text accuracy on this kind of task, while Midjourney manages roughly 30% on short phrases and falls apart on anything longer, according to testing referenced by The Decoder's coverage of the Ideogram 3.0 launch. Midjourney v7 nailed the mood and color grading but mangled "NIGHT RADIO" into a near-legible scribble on two of three attempts, and dropped the second line of text entirely once. Flux 2 Pro's reworked text engine was a real step up from Flux 1 — it got the headline right but occasionally warped the smaller venue-and-date line. Stable Diffusion 3.5 struggled the most, which lines up with the open-source community's own consensus that text remains the weakest spot for diffusion checkpoints without a ControlNet or dedicated text-rendering LoRA layered on top.
"Flux is a new open source image generator that is as good as Midjourney. Midjourney has a better aesthetic and skin texture, Flux has better text and anatomy." — @risphereeditor on X
This is the one test where Writingmate's current lineup shows its limit — Ideogram isn't in the catalog, so for genuinely text-critical work like packaging or signage, I'd still reach for Ideogram directly. GPT Image 2 and Recraft V4.1 inside Writingmate get close enough for social graphics and internal decks, but not close enough to replace a dedicated Ideogram subscription for headline-critical print work.
Test 3: Keeping One Character Consistent
The prompt, run three times with different scene descriptions: "A red-haired woman in a green field jacket, freckles, short curly hair" — first reading a book in a café, then walking through a rainy city street, then standing at a train platform at sunset.
Character consistency across separate generations is still the hardest unsolved problem in this category, and every tool handles it differently. Midjourney v7's --oref parameter is purpose-built for this and it's genuinely convenient — feed it one reference image and the face and hairstyle held up across all three scenes with only minor jacket-color drift. Flux 2 Pro takes a different approach: it accepts up to ten reference images at once, which technically gives it the strongest consistency mechanism of the four once you feed it more than one angle of the character, though with only a single reference it performed about on par with Midjourney. Ideogram's character reference feature kept the face recognizable but changed the jacket style between scenes twice. Stable Diffusion 3.5 was the weakest out of the box — genuine consistency there means training a small LoRA on your character first, which is more setup than a one-off project justifies, but it's also the only option that gives you a permanent, reusable asset once you've done it. If you're building a recurring brand mascot or comic series, that LoRA investment pays for itself; if you need three consistent images by Friday, it won't.
Test 4: The Photorealistic Portrait
The prompt: "Close-up portrait of a middle-aged man with gray stubble, wearing a wool sweater, natural window light, shallow depth of field, shot on an 85mm lens."
Flux 2 Pro won this one clearly, and it matches what most independent testing has converged on this year — if an image needs to pass as an actual photograph, Flux is the model to reach for. Skin texture, the falloff of the window light, and the lens compression all read as camera-real rather than rendered. Midjourney v7 came close and arguably had better emotional presence in the eyes, but it leans toward a slightly polished, editorial look even when you ask for something plainer. Ideogram 3 produced a solid, usable portrait but it's not built or marketed as a photorealism specialist, and it showed — slightly waxy skin under close inspection. Stable Diffusion 3.5's base checkpoint was noticeably behind the three closed models here; the open-source ecosystem's realism specialists (community fine-tunes trained specifically for skin and lighting) close that gap, but the vanilla SD 3.5 weights alone did not.
Test 5: The Stylized Illustration
The prompt: "A whimsical children's book illustration of a fox reading a map by lantern light in a forest, warm watercolor style, soft edges."
Midjourney v7 took this one without much of a contest — it's still the model most tuned toward "make it beautiful," and the watercolor texture, the warmth of the lantern light, and the overall composition looked genuinely art-directed rather than generated. Stable Diffusion 3.5 was the surprise strong second, purely because of its LoRA ecosystem — once you're using a checkpoint fine-tuned for watercolor or storybook illustration, it can match or beat the closed models for that specific niche, which no other tool here can claim. Flux 2 Pro leaned more literal and photographic than the prompt asked for, even with "watercolor style" spelled out twice. Ideogram 3 handled the style reasonably well but the fox's proportions varied more than I'd want for a run of book pages that need to feel like one consistent hand drew them.
Cost Per Image, Compared Honestly
Marketing pages love to hide this math. Here's what each tool actually costs once you convert subscriptions and GPU-hour plans into a per-image number:
Tool | Entry price | Approx. cost per image | Best for |
|---|---|---|---|
Midjourney v7 | $10/mo (Basic, 3.3 GPU hours) | ~$0.01–0.03 in Fast mode; well under $0.01 in Draft Mode | Aesthetic, illustration, mood boards |
Flux 2 Pro | Pay-as-you-go API | ~$0.03–0.045 (BFL's official per-megapixel pricing) | Photorealism, product shots |
Ideogram 3 | Free tier (10 slow credits/day) or $8/mo Basic | Roughly $0.02 on the cheapest paid tier | Text-heavy posters, packaging, logos |
Stable Diffusion 3.5 | Free (open weights) | $0 self-hosted; cents per image on hosted GPU providers | Local control, custom LoRA fine-tuning |
Writingmate (built-in) | $19.99/mo (Pro, 800 credits) | ~$0.05–$0.50, images cost 2–20 credits depending on model | One account for Flux, Nano Banana Pro, GPT Image 2, Recraft, plus chat and video |
Midjourney's own community has been vocal about how much Draft Mode changes this math. As one prolific tester put it:
"In Midjourney --v 7 using draft mode + turbo you can generate 160 images in ~42 seconds and it costs about ~$1.30 pretty good" — Nick St. Pierre on X
That works out to roughly $0.008 per image in Draft Mode, which quietly makes Midjourney one of the cheapest options here for rapid iteration, even though its subscription looks like the most expensive line item on paper.
Where Writingmate's Built-In Generator Actually Fits
I want to be straight about this instead of overselling it: Writingmate doesn't host Midjourney, Ideogram, or Stable Diffusion. Midjourney has no API to integrate — it's Discord and web-app only — and Ideogram and open Stable Diffusion checkpoints simply aren't part of the current model lineup. What Writingmate does host is FLUX.2 Pro, Max, and Flex, plus Nano Banana Pro, GPT Image 2, and Recraft V4.1 — real, current-generation models, not a watered-down alternative. Practically, that means: for the product-shot and photorealistic-portrait jobs, running Flux 2 Pro inside Writingmate gets you the same win as the standalone app, because it's the same model. For text-heavy work, you're better off leaning on GPT Image 2 or Recraft V4.1 inside the platform, or keeping a separate Ideogram account if headline-accurate text is a daily requirement. And since images now draw from the same monthly credit pool as chat and video — a change Writingmate rolled out on its pricing page a couple weeks ago — there's no separate "you're out of images for the day" wall to hit mid-project.
The community discussion around this trade-off — one flexible platform versus juggling specialist accounts — comes up constantly. It's a recurring theme in threads on r/StableDiffusion, where the consensus tends to split along the same line I landed on: if you need the absolute best at one narrow thing (text, in Ideogram's case), go get the specialist tool. If you need "very good" across five different job types without five different logins, a hosted multi-model platform wins on time saved even when it doesn't win every individual test.
Which One Should You Actually Use
Nobody wins all five tests, and that's really the headline finding here. Flux 2 Pro is your best default for anything that needs to look like a real photograph — product shots, portraits, marketing stills. Midjourney v7 still owns aesthetic and illustration work; if a client is judging the output on "does this look beautiful," it's the one to reach for. Ideogram 3 is non-negotiable the moment your image needs to contain real, readable words. And Stable Diffusion 3.5 earns its keep specifically when you want zero recurring cost, full local control, or a custom-trained character LoRA that none of the closed models will ever give you.
If you don't want to manage four separate subscriptions to cover four separate strengths, Writingmate's text-to-image tool gets you most of the way there with Flux 2 Pro, Nano Banana Pro, GPT Image 2, and Recraft in one place, alongside chat and video on the same plan. It won't replace Ideogram if text accuracy is your whole job, but for the other four tests I ran, it held up.
See you in the next one!
Artem
Frequently Asked Questions
Sources
- Black Forest Labs — FLUX.2 official announcement
- Midjourney — V7 is now the default model
- Stability AI — Introducing Stable Diffusion 3.5
- The Decoder's coverage of the Ideogram 3.0 launch
- r/StableDiffusion
- Nick St. Pierre on X
- YouTube — Ultimate AI Image Generator Comparison
- text-to-image tool
- FLUX.2 Pro
- FLUX.2 Pro, Max, and Flex
- @risphereeditor on X
- pricing
Written by
Artem Vysotsky
Ex-Staff Engineer at Meta. Building the technical foundation to make AI accessible to everyone.
Reviewed by
Sergey Vysotsky
Ex-Chief Editor / PM at Mosaic. Passionate about making AI accessible and affordable for everyone.
