ONLINEAGENT_OPS 2026.Q3 HOME ARTICLES CRAFT RECORD BLOG MAP HUBS FAQ SEARCH
HOMETHE CRAFTMODEL GUIDES
THE CRAFT · TOOL SELECTION

Model Guides

Around twenty models and studios with their optimal prompt format — ChatGPT Images 2.5, Seedance, Kling, FLUX, Midjourney, Suno, ElevenLabs and the rest.

READ33 min
WORDS6,553
SECTIONS4
SOURCES15
TYPEGUIDE
CHECKED22 SEP 26
TL;DR — THE SHORT VERSION

Pick a model by task and trust only dated checks: rankings move weekly, prompt formats last longer, and the Sora API shuts down 24 September 2026.

  • Sora is ending. Do not start new work on it; Veo 3.1 and Seedance are the practical migration paths.
  • Check for a negative field. FLUX.2, Runway, Seedance 2.5 and ChatGPT Image do not accept one.
  • Leaderboards are a poor single signal. A model can fall down the rankings while it improves.
  • Seedance 2.5 makes longer clips. It offers native 30-second single-pass generation with co-generated audio.
  • On ChatGPT Image, read revised_prompt. OpenAI rewrites your prompt, so a mismatch may come from the rewrite.
◈ PRICES ON THIS PAGE

Every price here was correct when this site last checked it, between 22 Sep and 23 Sep 2026. Prices change often and this site no longer updates them, so check the vendor’s own page before you rely on one: BytePlus ModelArk (Seedance) pricing · Runway Gen-4.5 credit costs · Black Forest Labs pricing · Pollo AI pricing.

◈ ABOUT THE NEGATIVE: LINES ON THIS PAGE

Each model entry below carries a suggested negative prompt. Four of the models on this page do not accept one. Black Forest Labs states that FLUX.2 "does not support negative prompts"; Runway states negatives are unsupported and that including one "may result in the opposite happening"; Seedance 2.5 has no negative field, so exclusions compete with the positive prompt in the same text; ChatGPT Image has no field either.

Where a field exists — Midjourney's --no, and Veo's negativePrompt on Google's Gemini Enterprise Agent Platform — keep the line short and specific. Kling recommends writing negatives inside the prompt, so on Kling treat the line as prompt text. Where it does not, read the line as a checklist for the positive prompt instead: not "no plastic skin" but "visible pores, uneven natural skin tone." Full per-vendor breakdown: what the vendors actually say.

One more from the vendor docs, for ChatGPT Image: OpenAI’s API automatically revises your prompt before generating, and returns the rewritten version in a revised_prompt field. When output does not match what you asked for, the cause may be the revision rather than your wording — so read revised_prompt before rewriting. OpenAI also states the model struggles with precise text placement and layout-sensitive composition, which prompt tuning does not fix.

Vendor documentation: docs.bfl.ai · help.runwayml.com · Google Cloud Veo guide · Kling API docs. Checked 25 Aug 2026
WHICH MODELS ON THIS PAGE TAKE A NEGATIVE PROMPT
Where there is no field, read each card’s NEGATIVE line as a checklist for the positive prompt.
A NEGATIVE FIELD EXISTS
Yes: Midjourney: --no
Yes: Veo: negativePrompt, on Google’s Gemini Enterprise Agent Platform
Note: Kling: recommends writing negatives inside the prompt, so treat the line as prompt text
NO NEGATIVE FIELD
No: FLUX.2
No: Runway: including one may have the opposite effect
No: Seedance 2.5: exclusions compete with the positive prompt in the same text
No: ChatGPT Image
Vendor documentation: docs.bfl.ai · help.runwayml.com · Google Cloud Veo guide · Kling API docs, checked 25 Aug 2026: FLUX.2 “does not support negative prompts”; on Runway a negative “may result in the opposite happening”. Full per-vendor breakdown on what the vendors actually say.
◈ UPDATE — 22 AUGUST 2026 · ONE MODEL ON THIS PAGE IS BEING SWITCHED OFF

OpenAI is discontinuing Sora. The consumer app and web product ended 26 April 2026; the API shuts down 24 September 2026.OpenAI Help Center, What to know about the Sora discontinuation and the API deprecations page, read at source 9 Sep 2026: developers were notified 24 March 2026, the web and app experiences ended 26 April 2026, and all API video generation including Sora 2 and Sora 2 Pro stops on 24 September 2026. Also reported across Pixo, ChatCut, Kingy and Teamday, Apr–Jul 2026

Do not start a new production workflow on it. The practical migration paths are Veo 3.1 if you were using it for realism and native audio — noting its 8-second native ceiling, or Seedance if you were using it for prompt adherence on commercial work.

◈ WHAT ELSE MOVED SINCE THIS PAGE WAS FIRST WRITTEN

Seedance 2.5 — previewed 23 Jun 2026 at ByteDance's FORCE conference, released 31 July 2026. Native 30-second single-pass generation with co-generated audio, and up to 50 multimodal references: 30 images, 10 video clips, 10 audio files — not 30 images alone.ByteDance’s own Seedance 2.5 page on Dreamina, read at source 11 Sep 2026: “create 4K AI videos up to 30 seconds with up to 50 multimodal references”. Not re-verified: re-read in a browser on 17 Sep 2026, the page no longer carries that sentence; it now says “Generate 30-second videos from up to 50 references.” The split into 30 images, 10 video clips and 10 audio files is not on that page as re-read. Also verified 22 Aug 2026 against Artlist, Morphic, Wiro and VioEvo. Resolution, updated 22 Sep 2026: ByteDance’s own API price list sells Seedance 2.5 at 480p, 720p and 1080p, and gives it no 4K row; Dreamina’s page still says “Create cinematic 4K videos with Seedance 2.5 in Dreamina.”, which this site has not confirmed.ByteDance, BytePlus ModelArk Pricing, page last updated 22 Sep 2026, read in a browser 22 Sep 2026: dreamina-seedance-2-5-260628 “For 1080p outputs: Input without video: 11.7” (USD per million tokens) · Dreamina, Seedance 2.5, read at source 22 Sep 2026. First-hand: until 22 Sep 2026 this line said the resolution was disputed, “reported as 720p at launch, 1080p since”.

The video leaderboard, as read on 22 Sep 2026. Re-query it; do not trust this snapshot. Artificial Analysis, text to video with audio (its default view): Gemini Omni Flash (1233), Wan 3.0 (1229), Minimax H3 Max post-trained by fal (1227), MiniMax H3 Open Weights (1220), Seedance 2.0 720p (1210). The top three share a rank range of 1 to 3, so their order is not settled. Without audio: Wan 3.0 (1336), Gemini Omni Flash (1330), MiniMax H3 Open Weights (1302), HappyHorse-1.0 (1287), HappyHorse-1.1 (1272), Seedance 2.0 720p (1259). Image to video, without audio: Gemini Omni Flash (1369). Seedance 2.5 has no row on the text-to-video board.Artificial Analysis, Text to Video Leaderboard and the image-to-video board beside it, read in a browser 22 Sep 2026, top row with audio: “Gemini Omni Flash 1233”. Elo is a live figure and moves weekly. First-hand: the earlier version of this note, with Elo taken 22 Aug 2026, said the standings sat behind an interactive voting arena that could not be read, and that /video/leaderboard returned 404. The leaderboards can be read in a browser at /video/leaderboard/text-to-video and /video/leaderboard/image-to-video. On 22 Aug the order was Gemini Omni Flash first without audio (1322) and Wan 3.0 first with audio (1244); by 22 Sep the two had swapped. An earlier version still said Seedance 2.0 held #1 at ~1,273 Elo; that was true when written and is no longer.

Runway Gen-4.5 confirmed out of the top 10. It led at launch in Dec 2025 with 1,247 Elo and was displaced by Seedance 2.0, HappyHorse-1.0 and the Kling/Veo cluster by mid-2026.Runway, Introducing Runway Gen-4.5, 1 Dec 2025, read at source 23 Sep 2026: “With 1,247 Elo points, Gen-4.5 currently holds the top position in the Artificial Analysis Text to Video benchmark”. Pinggy, Best Video Generation AI Models in 2026, updated 14 Jul 2026, read at source 23 Sep 2026: “Runway Gen-4.5, which led at launch in late 2025 with 1247 Elo, has dropped out of the top 10.” Also checked against an Artificial Analysis API reading via Banana Flow, 5 Aug 2026 (no link was recorded). On 22 Sep 2026 it sat 19th on the text-to-video board without audio, at 1214 Elo, with a rank range of 15 to 26.Artificial Analysis, Text to Video Leaderboard, no-audio view, read in a browser 22 Sep 2026: “Runway Gen-4.5 1214”, rank range “15-26”. Re-query it; do not trust this snapshot. An earlier version of this line gave an Aug 2026 secondary reading of about 1,215 at rank 15. It also gained native audio generation and editing in May 2026, which closes its longest-standing gap against Veo — kept here because a model falling off a leaderboard while improving is exactly why leaderboards are a poor single signal.

How to read this page

Every entry below carries the date it was last checked, not the date it was written. Model rankings move weekly and several of these will be wrong within a quarter — which is why the date is on the card rather than in a footer.

30s
longest single-pass clip currently available, on Seedance 2.5. Most models still top out at 8–16 seconds and anything longer is stitchedDreamina (ByteDance), Seedance 2.5, read at source 23 Sep 2026: “You can create cinematic videos up to 30 seconds in standard mode or extend them to 180 seconds with the beta long-video mode.”
48kHz
Veo 3.1 speech generation — still the native-audio leader in a July 2026 roundupPinggy, Best Video Generation AI Models in 2026, updated 14 Jul 2026, read at source 23 Sep 2026: “Veo 3.1 still owns this with 48kHz speech generation.”
24 SEP
the date the Sora API shuts down. If you have a pipeline on it, that is your deadline

Each entry carries the prompt format that works for that model and what it is actually good at.

◈ READ THESE AS DATED SNAPSHOTS

These were written in mid-2026. Model capabilities move faster than any guide. The prompt formats are the durable part; the capability claims carry a date and will eventually be wrong.

Check the stamp before trusting a comparison.

Guides for the tools covered here. Each card includes a prompt format and a copyable template; settings and capability claims carry the date on the card.

The best generative AI models right now — with optimized prompt formats for each. Click any model to expand its full prompting guide.

TAKEAWAY

Every card is a dated snapshot. The prompt formats last; the capability claims do not, and several will be wrong within a quarter, so check the stamp.

AI MODELS

IMAGE GENERATION · TEXT-TO-VIDEO · AUDIO · LLM — STANDALONE AI MODELS
🖼️
ChatGPT Images 2.5
OpenAI · Text-to-Image · 2026
CHECKED 22 AUG 2026 · VERSION 16 SEP 2026
🏆 PERSONAL BEST — TEXT-TO-IMAGE
  • Current version: ChatGPT Images 2.5, released 8 September 2026; in the API, gpt-image-2.5-sunburst and gpt-image-2.5-flare.OpenAI, Introducing ChatGPT Images 2.5, read in a browser 16 Sep 2026; Image generation guide, read at source 16 Sep 2026. An earlier version of this card named ChatGPT Images 2.0 as current.
  • Selling Point: Best conversational image editing — refine through natural language in multi-turn dialogue
  • Renders text inside images, with a limit OpenAI states itself: “Although significantly improved, the model can still struggle with precise text placement and clarity.”OpenAI, Image generation guide, read at source 17 Sep 2026. An earlier version of this card said “Excellent text rendering” with no qualification.
  • 4-round iterative workflow produces commercial-grade output
  • Understands complex multi-element scenes
+ VIEW PROMPT TEMPLATE & SAMPLE
OPTIMAL PROMPT FORMAT
Create a photorealistic portrait: 1woman, late 20s, tailored ivory blazer, golden hour backlight, Canon EOS R5 85mm f/1.2 ISO 200, shallow depth of field, Vogue Japan editorial quality, 8K. FOLLOW-UP: "Keep everything identical but change the blazer to deep burgundy"
KEY ADVANTAGE: Use follow-up messages to iterate. "Keep everything the same but..." is the most powerful phrase in ChatGPT Images 2.5.
🍌
Google · Gemini 3.1 Flash Lite Image · Speed · 2026
CHECKED 16 SEP 2026
TEXT-TO-IMAGE · SPEED MODEL
  • What it is: Google's fast, cheap image model — “engineered for velocity and scale where speed and cost are the primary operational constraints”.Google, Nano Banana image generation, read at source 16 Sep 2026: “Our fastest and cheapest Gemini image model, engineered for velocity and scale where speed and cost are the primary operational constraints.” An earlier version of this card called it “Nano Banana Pro 2”, a name Google does not use.
  • Purpose-built for rapid iteration and prompt testing before committing to slower, higher-quality models
  • Not for reference-heavy work. Google: “Not optimized for multiple reference inputs or multi-turn sequential editing.” It also tops out at 1K: “Gemini 3.1 Flash Lite Image only supports 1K resolution.”
  • Ideal workflow: prototype here, then finish in Nano Banana 2 (4K output, stronger text) or another model
  • Best for: rapid concept testing, volume social content, quick client mock-ups
+ VIEW PROMPT TEMPLATE & SAMPLE
NANO BANANA 2 LITE — PROMPT SHAPE
RAW photo, [SUBJECT DESCRIPTION], [LIGHTING], [CAMERA: 85mm f/1.2], [COLOR GRADE], photorealistic, ultra-detailed, 8K Keep prompts under 60 words for fastest generation. Front-load subject and lighting. Add quality tags at the end. For iteration: run 3–5 variations, select the best, upscale with the Universal 8K Upscale prompt. There is no negative prompt field. Google's advice is a “semantic negative prompt”: “Instead of saying "no cars," describe the intended scene positively”.
SPEED WORKFLOW: Nano Banana 2 Lite → select best → Universal 8K Upscale → deliver.
🐴
HappyHorse-1.0
Alibaba-ATH · Video generation · 2026
CHECKED 17 SEP 2026
VIDEO MODEL · NOT AN IMAGE MODEL
  • What it is: a video model. Artificial Analysis lists HappyHorse among its video models, under Alibaba-ATH, and it appears in the text-to-video standings in the leaderboard note at the top of this page.Artificial Analysis, HappyHorse model family page, read in a browser 17 Sep 2026: “Analysis of HappyHorse models and comparison to other video models across key metrics including quality, generation time, and price.” An earlier version of this card presented HappyHorse-1.0 as a text-to-image model, labelled it TOP-2 in April 2026, and gave it an image prompt template. None of that had a source, and it was removed.
  • No prompting guide for it has been read at source for this page, so no prompt format is given here.
🎬
Seedance 2.0 SUPERSEDED BY 2.5
ByteDance · Text-to-Video · 2026
CHECKED 22 AUG 2026
🏆 PERSONAL BEST — TEXT-TO-VIDEO
  • Selling Point: Human motion for fashion, lifestyle and commercial clips (this site’s use, not a benchmark result)
  • This card’s template puts [MOTION] first; ByteDance does not require it
  • Identity lock required: "same consistent face, no morphing" prevents character drift
  • Clips of 4 to 15 seconds through the API (Seedance 2.5 goes to 30), with a cost comparison here.BytePlus, Dreamina Seedance 2.5 tutorial, read at source 22 Sep 2026: the duration row reads “4–30 seconds” for 2.5 and “4–15 seconds” for the 2.0 series. First-hand: until 23 Sep 2026 this line said “Clips up to 10 seconds at 24fps”.
  • Best for: AI influencer video, cinematic human motion, brand lifestyle clips
+ VIEW PROMPT TEMPLATE & SAMPLE
BRACKET FORMAT (ALWAYS USE THIS)
[MOTION]: slow graceful walk, natural heel-to-toe rhythm [SUBJECT]: 1woman, white dress, same face throughout, no morphing [ENV]: Tokyo alley, midnight, neon reflections on wet cobblestones [CAMERA]: tracking dolly, 35mm anamorphic, f/2.0 [LIGHTING]: practical neon only, warm amber and cool blue [STYLE]: motivated low-key light, restrained palette, clean wide framing, 24fps [DURATION]: 8 seconds, no scene cuts
RULE #1: Always start with [MOTION]. This is non-negotiable for Seedance 2.0.
Kuaishou · Text-to-Video · Kling 3.0 · 2026
CHECKED 22 AUG 2026
PHYSICS & ACTION SPECIALIST
  • Selling Point: Best physics-accurate motion — water, fire, fabric, and impact sequences look physically real
  • Kling 3.0 (2026): extended clips, improved facial consistency, strong image-to-video
  • Clips from 3 to 15 seconds. Kling: “The new model generates up to 15 seconds of continuous video, with a flexible duration ranging from 3 to 15 seconds.”Kling AI, Kling VIDEO 3.0 model user guide, read at source 17 Sep 2026. An earlier version of this card said up to 10 seconds.
  • Best for: physics action, product reveal sequences, nature and environment video
+ VIEW PROMPT TEMPLATE & SAMPLE
OPTIMAL PROMPT FORMAT
A crystal champagne glass shatters in extreme slow motion, thousands of fragments suspended mid-air catching studio light, liquid droplets frozen in perfect spheres, black infinity background, overhead lighting creating caustic light patterns, locked camera on tripod, 135mm macro lens, 120fps, hyperrealistic glass physics, 4K
KEY TIP: Describe the physics explicitly. Name the material, the force acting on it, and the expected behavior.
🏃
Runway Gen-4.5 FELL OUT OF TOP 10
Runway · Gen-4.5 · 2026
CHECKED 22 AUG 2026
GEN-4.5 — 2026
  • Selling Point: Best-in-class for brand commercial VFX and style transfer sequences
  • Clips run 2 to 10 seconds, at 12 credits a second. Runway’s Gen-4.5 page lists “Supported durations 2 - 10 seconds”.Runway, Creating with Gen-4.5, read at source 22 Sep 2026: “Cost 12 credits per second”.First-hand: until 22 Sep 2026 this line said “Clips are 5 or 10 seconds”, quoting Runway’s Gen-4 guide: “Gen-4 creates videos in 5 or 10 second durations”. That is Gen-4’s range, not Gen-4.5’s. An earlier version of this card was titled Gen-4 Turbo and said clips run up to 16 seconds.
  • Strong at transformations, morphs, and surrealist visual effects for commercial use
  • Best for: brand commercials, VFX sequences, style transfer, film-grade video production
+ VIEW PROMPT TEMPLATE & SAMPLE
OPTIMAL PROMPT FORMAT
A luxury perfume bottle dissolves into golden light particles that spiral upward, camera slowly orbiting, particles reconverge into a larger perfect version of the bottle against deep black infinity background, cool desaturated grade, controlled low-key light, precise locked framing, 4K HDR, photorealistic particle simulation, 24fps, 10 seconds, no cuts
🌊
FLUX.2 Pro
Black Forest Labs · FLUX.2 [pro] · 2026
CHECKED 22 AUG 2026
FLUX.2 PRO — 2026
  • Selling Point: Highest prompt adherence of any image model — what you write is what you get, precisely
  • Black Forest Labs lists four FLUX.2 tiers: “FLUX.2 [pro], FLUX.2 [flex], FLUX.2 [dev], and FLUX.2 [klein]”. An earlier version of this card described a “FLUX.2 Pro Ultra” tier, which Black Forest Labs does not list.Black Forest Labs, FLUX.2 launch post, 25 Nov 2025, read at source 16 Sep 2026. bfl.ai
  • Settings: this site found no Black Forest Labs source for CFG, steps or sampler on FLUX.2 [pro], so none are given here
  • Black Forest Labs now makes video too. FLUX 3, in preview since 4 Aug 2026, is its first video model: “one model trained across image, video, and audio, generating video with synchronized sound”. Text- or image-to-video costs $0.17 a second at hd, $0.29 at fhd, $0.40 at qhd and $0.80 at uhd, or $0.06 a second for a draft.Black Forest Labs, Release notes, 4 Aug 2026 entry, read at source 22 Sep 2026: “FLUX 3 is our first video model”, “Available now as a preview.” · API Pricing, read at source 22 Sep 2026: “Pay per image, or per second for video”.
  • Best for: technical compositions, detailed architecture, precise product visualization, ComfyUI pipelines
+ VIEW PROMPT TEMPLATE & SAMPLE
OPTIMAL PROMPT FORMAT
A hyperrealistic photograph of 1woman, age 28, wearing a structured ivory blazer, standing in a rain-soaked Tokyo street at blue hour, neon reflections on wet pavement, Canon EOS R5, 85mm f/1.2, ISO 800, depth of field separation, photojournalism aesthetic, ultra-detailed skin texture, 8K resolution, no text, no watermark
🌙
Midjourney · Text-to-Image · V8.2 · 2026
CHECKED 22 AUG 2026 · VERSION 17 SEP 2026
EDITORIAL AESTHETICS — V8.2
  • Selling Point: Strongest editorial aesthetic and artistic composition of any image model — distinctly recognizable visual style
  • Current default: V8.2, which Midjourney says “released as the default version on July 24, 2026.” It describes V8.2 as “an update focused on aesthetics, image quality, and Personalization.”Midjourney, Version, read in a browser 17 Sep 2026. An earlier version of this card presented “Midjourney v7 (April 2026)” as current. Midjourney: “V7 was released on April 3, 2025, and was the default version from June 17, 2025 to June 9, 2026.”
  • Use --ar for aspect ratio, --no instead of negative prompts, --q 2 for quality
  • Best for: editorial fashion, creative direction, concept art, artistic campaigns
+ VIEW PROMPT TEMPLATE & SAMPLE
OPTIMAL PROMPT FORMAT
editorial fashion photography, 1woman ivory blazer, golden hour backlight, Vogue Japan, 85mm f/1.2, ultra-detailed, photorealistic --ar 4:5 --style raw --p [your-persona-id]
TIP: Use --p [persona-id] for a consistent personal style and --style raw for less default styling. With no --v parameter, the prompt runs on the current default version.
🎶
Udio DOWNLOADS DISABLED
AI Music · v1.5 (2024)
CHECKED 22 SEP 2026
AI MUSIC — NO DOWNLOADS
  • The catch: you cannot take a track out. Since its partnership with Universal Music Group, announced 29 Oct 2025, Udio says “downloading of audio, video, and stems has been disabled”.Udio Help Center, Changes associated with the Universal Music Group partnership, read at source 22 Sep 2026: “Note that downloading of audio, video, and stems has been disabled”.
  • Udio v1.5 is a 2024 model. Udio’s post introducing it, with stem downloads as a new feature, was “Published on Jul 23, 2024”.Udio, Introducing v1.5, read at source 22 Sep 2026.First-hand: until 22 Sep 2026 this card gave Udio’s selling point as “Stem separation and export”, dated v1.5 to 2026 and recommended it for “stems for mixing”. Stem downloads had been switched off before the card was checked on 22 Aug 2026.
  • Best for: sketching and remixing songs inside Udio, not for a track you need to use elsewhere
+ VIEW PROMPT TEMPLATE & SAMPLE
OPTIMAL PROMPT FORMAT
lo-fi hip hop, late night study session, warm vinyl crackle, dusty sample chops, muted Rhodes piano, soft brushed snare, 85 BPM, melancholic and focused, Tokyo midnight aesthetic, no vocals, 2-minute seamless loop EXTEND: keep the same groove but introduce a subtle string arrangement in the second minute, bring Rhodes forward in the mix
SUNO vs UDIO: Suno lets you download, within limits set on 3 Sep 2026 (the Suno sheet has them). Udio has switched downloads off, stems included, so a track made there cannot be downloaded.
🎵
AI Music Generation · Suno v6 · 2026
CHECKED 16 SEP 2026
AI MUSIC — SUNO V6 · 2026
  • Suno v6 (9 Sep 2026) — Suno’s flagship, alongside v6-wild, “less predictable and more varied”, and v6-mini, “available to everyone”. “All models prior to v6 have been retired.” An earlier version of this card described v5.5 as current.Suno, v6 release notes, 9 Sep 2026, and v6 FAQ; read at source 16 Sep 2026.
  • Full songs with lyrics, vocals, and production from a single text prompt — any genre, any mood
  • Suno Studio (2026): multi-track editing, stem separation, and remix tools built into the platform
  • Best for: AI film scores, brand content, social media, influencer soundtracks, full song production
+ VIEW PROMPT TEMPLATE & SAMPLE
SUNO V6 PROMPT FORMAT
Genre: [cinematic electronic / indie pop / hip-hop / ambient / etc.] Mood: [tense / uplifting / melancholic / energetic] Instrumentation: [specific instruments + production style] Tempo: [BPM or descriptive: slow / mid-tempo / driving] Vocals: [male/female/none, style: raw/polished/whispered] Energy arc: [builds from X% to Y% over Z seconds] Reference: [Artist A meets Artist B style] Duration: [30s / 60s / full song] Production quality: commercial, broadcast-ready
V6 NOTE: Suno’s v6 FAQ says the Variety slider works by “adjusting and updating your style prompts”. When a style prompt must be followed closely, keep Variety low.
THE PART MOST SUNO GUIDES SKIP — TWO FIELDS, NOT ONE. The block above reads like a single text box. It is not. Suno splits into a style prompt (the sound: genre, mood, instruments, vocals, BPM) and a lyrics field (the words, plus structural metatags like [Verse], [Chorus], [Bridge]). Putting lyrics in the style box, or genre in the lyrics box, is the most common reason output ignores half the prompt.

Suno writes its own style example in sentences: “Style: Contemporary R&B song at 92 BPM in F minor.” Be specific — Suno says to “mention genre, mood, keywords, and instrumentation”.

What you do not want goes in Exclude, in Custom Mode under Advanced Options.Suno, How to make a song and How do I exclude elements of a song?, read at source 16 Sep 2026. An earlier version of this card said the style box wants comma-separated descriptors, not sentences, and gave third-party character limits and a descriptor sweet spot for v4.5–v5.5, models Suno has since retired. Those were removed.
🌊
Seedream 5.0 Pro
ByteDance · Text-to-Image · via Dreamina / CapCut · 2026
CHECKED 16 SEP 2026 · vendor documentation
🖼 TEXT IN IMAGES · MULTI-REFERENCE · PRECISION EDITING
  • Seedream 5.0 arrived in February 2026; 5.0 Pro on 8 July 2026. ByteDance's image model, and the sibling of Seedance — same platform, same documentation style.
  • Text rendering is the headline. Small fonts render more accurately, character repetition is reduced, and the bias toward bold is corrected.
  • Up to six reference images, merged into one result.
  • 5.0 edits by partial selection and pen input; 5.0 Pro adds point and lasso selection and layer separation.ByteDance Seed, Introducing Seedream 5.0 Pro, 8 Jul 2026, read at source 16 Sep 2026: “point selection, lasso selection, sketch rendering, color and material replacement, layer separation, and multi-image fusion”. An earlier version of this card said editing works "rather than layers" and dated the release 10 February 2026; Dreamina's Seedream 5.0 guide is dated 5 February and ByteDance's 5.0 Lite post 13 February.
VENDOR PROMPT ORDER: Purpose → Subject → Visual style & mood → Composition / layout → Must-have details → Reference image — the step list in Dreamina's Seedream 5.0 Pro guide. The guidance is to keep prompts "clear, structured and focused" and to "develop one distinct direction and fill in with supporting visual details."
PUT THE EXACT WORDS IN THE PROMPT. For posters, flyers and UI, ByteDance's instruction is to "include exact text content" rather than describing it. Write the words you want rendered, in quotes, not "a sign that says something about a sale." This is the single highest-value habit on this model.
SIX REFERENCES, EACH WITH A JOB. Uploading six images is not the technique — labelling them is. Their guidance: "clearly describe what each one contributes, such as color schemes, layouts, character poses, or artistic styles." An unlabelled reference stack is a guess; a labelled one is direction. Same principle as Seedance's @Image1 tagging.
WHAT THEY TELL YOU NOT TO DO: do not stuff several styles or conflicting instructions into one prompt; do not leave the main subject ambiguous or implied; and when refining, do not change many elements at once — break edits into small steps so each change stays controlled. The conflict warning is the same failure that breaks long negative prompts elsewhere on this page.
RESOLUTION: 2K is described as sufficient for web and social; 4K for print work such as posters and flyers. The model also decides for itself when to run a real-time search — you do not switch that on.Dreamina (ByteDance) Seedream 5.0 Pro guides, dreamina.capcut.com, checked 25 Aug 2026
🎙️
Text-to-Speech & Voice Cloning · Eleven v3 / v2 · 2026
CHECKED 25 AUG 2026 · vendor documentation
🎧 VOICE — THE ONE WHERE PUNCTUATION IS THE PROMPT
  • There is no prompt box. On text-to-speech, the text you submit is the prompt — punctuation, capitalisation and tags are your only controls over delivery.
  • Eleven v3 and earlier models take different syntax. This is the single most important thing on this entry.
  • Voice cloning holds identity across a content series; the voice you clone constrains what any tag can do to it.
v3 — AUDIO TAGS. Emotion and delivery come from inline square-bracket tags: [whispers], [excited], [sarcastic], [laughing], [sigh]. ElevenLabs: “Eleven v3 does not support SSML break tags.” Use punctuation and audio tags for pauses instead.
v2 — BREAK TAGS. Pauses use <break time="1.5s" />, up to a 3 second maximum. ElevenLabs warns that too many break tags in one generation "can cause instability" — the voice may speed up or introduce artefacts. Use a few deliberately, not one after every line.
PUNCTUATION IS THE UNIVERSAL CONTROL. Ellipses create a pause and add weight. CAPITALISATION increases emphasis. Ordinary punctuation sets the rhythm. Example from the docs: It was a VERY long day [sigh] … nobody listens anymore. These work without any tag syntax and are the safest lever across versions.
THE LIMIT NOBODY MENTIONS: a tag cannot override the voice. ElevenLabs states plainly that a voice will not contradict its training — [shout] on a voice trained on whispering will not produce a shout. Tag effectiveness depends on the voice and its training samples, so a tag that works on one voice may do nothing on another. They also note Professional Voice Clones are not fully optimised for v3. Test your tags on your actual voice before scripting around them.ElevenLabs prompting documentation, elevenlabs.io/docs, checked 25 Aug 2026
ORDER OF OPERATIONS: write the script with punctuation first and generate. Only add tags where the plain reading came out wrong. Tags stacked pre-emptively on every line are the voice equivalent of a bloated negative prompt — they fight each other and destabilise the take.
✂️
OpusClip Pro
AI Video Repurposing · opus.pro · 2026
CHECKED 22 AUG 2026
AI VIDEO CLIPPING
  • Selling Point: Upload any long-form video and it automatically finds the most viral moments, reformats for vertical, and exports short clips with captions
  • AI identifies hooks, engagement peaks, and viral patterns automatically — no manual editing
  • Outputs TikTok, Reels, and Shorts-ready clips with bold captions in one pass
  • Best for: podcasters, YouTubers, educators, and brands turning long content into daily social posts
+ VIEW PROMPT TEMPLATE & SAMPLE
OPUSCLIP OPTIMIZATION GUIDE
For best results with OpusClip Pro: 1. Upload videos 10+ minutes long — more content = better clip variety 2. Set clip length to 45-60 seconds for TikTok/Reels 3. Enable AI captions and choose a bold font style 4. Use the "Viral Score" filter to prioritize highest-engagement clips 5. Batch export 10+ clips per video for a full week of content
WORKFLOW: Record a 30-min podcast → OpusClip Pro extracts 15 viral clips → schedule across TikTok, Reels, Shorts for 2 weeks of content automatically.
🤖
AutoClips
Faceless AI Video Automation · autoclips.app · 2026
CHECKED 22 AUG 2026
FACELESS AI AUTOMATION
  • Selling Point: Full end-to-end faceless video automation — it writes the script, generates voiceover, creates visuals, adds captions, and auto-posts to TikTok, YouTube, Instagram, and Facebook daily.. Set it up once and it runs without you.
  • Auto-posts to all 4 platforms simultaneously — with platform-optimized titles and hashtags for each
  • Generates thumbnails, captions, and optimal posting times automatically — no Canva, no scheduling apps
  • Best for: faceless YouTube channels, TikTok automation, and high-volume content creators
+ VIEW DETAILS & PRICING
AUTOCLIPS — HOW IT WORKS
1. Connect your TikTok, YouTube, Instagram, or Facebook account via secure OAuth 2. Choose your niche and video type (Character Explainer / Faceless / 3D Animation) 3. Set your posting schedule (1–3 videos per day) 4. AutoClips writes a fresh script, generates voiceover, creates visuals, adds captions, and auto-posts on the schedule you set
AUTOMATION STACK: AutoClips for daily volume content → OpusClip Pro for clipping long-form into Shorts/Reels → ElevenLabs for custom voice consistency across all content.
↗ START AUTOMATING FREE
Ready to use these models? Browse production prompts →

AI STUDIOS

CREATIVE PLATFORMS · PRODUCTION SUITES · AI AGGREGATORS

An AI Studio is an all-in-one creative platform that bundles multiple AI capabilities — image generation, video, voice, editing, and automation — into a single interface. Unlike standalone models, studios are designed for full production workflows.

AI Aggregator — platform that gives you access to multiple AI models in one place AI Pipeline — chain of AI tools working in sequence to produce a final output AI Persona — a consistent AI-generated character identity across content

★ There are many AI studios in the market — it comes down to where you feel comfortable and what suits your workflow. Try a few and commit to the one that fits.

🎬
Higgsfield Studio
AI Cinematic Studio · Identity Lock · 2026
CHECKED 22 AUG 2026
🎬 PERSONAL GENERATIVE AI STUDIO
  • Selling Point: Not a model — a full cinematic AI production studio
  • Pioneering platform for Hollywood-grade AI video production
  • AI aggregator: access multiple video models in one interface
  • Upload a face photo → generate that person in any scene
  • Best for: AI influencers, character films, brand content
+ VIEW STUDIO GUIDE
OPTIMAL PROMPT FORMAT
[CHARACTER]: The person in the uploaded reference photo. Maintain exact facial identity throughout every frame. No morphing, no drift. [ACTION]: walks confidently through crowd, natural relaxed stride [CAMERA]: medium tracking shot, 35mm Cooke anamorphic, f/2.0 [ENV]: Tokyo Shibuya crossing, rush hour, evening, wet reflective pavement [STYLE]: atmospheric haze, backlit practical light, wide anamorphic framing, cinematic 24fps
STUDIO TIP: Pair with APOB AI to create your AI persona, then bring them to life here.
🎭
APOB AI
AI Influencer Creation Studio · 2026
CHECKED 22 AUG 2026
AI PERSONA STUDIO
  • Selling Point: Hands-on AI influencer customization — build and control every detail of your AI persona with full creative control, from appearance to identity consistency across unlimited content
  • Virtual try-on, branded outfit swaps, and lifestyle scenario generation — no photography needed
  • Best for: AI influencer creation, brand ambassador content, fashion and beauty campaigns
+ VIEW PROMPT TEMPLATE & SAMPLE
APOB AI PERSONA PROMPT
Create a consistent AI persona: female, late 20s, Southeast Asian, long dark hair, confident expression, athleisure style. Maintain identical face and features across all content. Generate in: outdoor lifestyle setting, golden hour, casual urban environment.
BEST WORKFLOW: Create your persona in APOB AI → generate lifestyle photos → run through ChatGPT Images 2.5 for product placement → Higgsfield for video clips.
Glam AI
Beauty & Fashion AI Studio · 2026
CHECKED 22 AUG 2026
BEAUTY & FASHION AI STUDIO
  • Selling Point: Upload a reference image and Glam AI handles the prompting automatically — template-driven beauty and fashion content without manual prompt writing
  • Sponsored post creation and brand partnership content generation built into the workflow
  • Best for: beauty brands, fashion influencers, product launch campaigns, e-commerce lookbooks
+ VIEW PROMPT TEMPLATE & SAMPLE
GLAM AI CONTENT PROMPT
Beauty editorial shoot: consistent AI model, flawless skin, [MAKEUP LOOK], wearing [OUTFIT/PRODUCT], professional beauty lighting, Vogue editorial aesthetic, sponsored post format, brand-clean background.
🎬
Research-to-Video Agent · opus.pro · 2026
CHECKED 22 AUG 2026
AI VIDEO AGENT
  • Selling Point: Give it a URL, PDF, or prompt — it researches the topic and auto-generates a complete, polished multi-scene explainer or promotional video from scratch
  • No timeline editing, no assets needed — fully autonomous video production from text input
  • Best for: explainer videos, thought leadership content, product demos, educational content
+ VIEW PROMPT TEMPLATE & SAMPLE
AGENT OPUS BRIEF FORMAT
Topic: [SUBJECT/URL/PDF] Goal: Create a 60-90 second explainer video Tone: [professional/conversational/energetic] Key points to cover: [3-5 bullet points] Target audience: [AUDIENCE] Call to action: [CTA] Style: clean, modern, branded
NOTE: Agent Opus creates NEW video from research. Different tool from OpusClip Pro (above) which edits existing video. Both are at opus.pro.
🎞️
Script-to-Video Studio · 2026
CHECKED 22 AUG 2026
SCRIPT-TO-VIDEO STUDIO
  • Selling Point: Converts any text, blog post, or script into a fully edited video with voiceover, music, captions, and stock footage — no editing skills needed
  • 16M+ stock clips, automated scene matching, brand kit, and direct social media publishing
  • Best for: marketers, YouTubers, agencies producing high-volume video content at scale
+ VIEW PROMPT TEMPLATE & SAMPLE
INVIDEO AI SCRIPT PROMPT
Create a 60-second [TOPIC] video for [PLATFORM]. Tone: [professional/casual/energetic] Target: [AUDIENCE DESCRIPTION] Hook: strong question or stat in first 3 seconds Structure: hook → 3 key points → CTA Voice: confident, clear, conversational Music: background, not distracting
BEST FOR: Repurposing blog posts and articles into social videos. Paste a URL and InVideo AI does the rest.
🖼️
Imagine Art
Multi-Model Image Studio · 2026
CHECKED 22 AUG 2026
AI IMAGE AGGREGATOR
  • Selling Point: 20+ AI image models in one platform — run the same prompt through FLUX, SDXL, Stable Diffusion variants simultaneously for comparison
  • ControlNet, LoRA support, and inpainting built-in — advanced workflows without ComfyUI
  • Best for: prompt engineers testing across models, designers exploring multiple visual aesthetics
+ VIEW PROMPT TEMPLATE & SAMPLE
MULTI-MODEL TEST PROMPT
RAW photo, [SUBJECT], [LIGHTING], [CAMERA + LENS], [COLOR GRADE], photorealistic, 8K, ultra-detailed NEGATIVE: cgi, plastic, airbrushed, watermark, blurry Run this prompt across FLUX.2 Pro, SDXL, and Playground for comparison before committing to a final model.
PRO USE: Use Imagine Art to A/B test your prompt across 3+ models before deciding which one to use for your final production run.
OpenArt AI
AI Creative Suite & Workflows · 2026
CHECKED 22 AUG 2026
AI CREATIVE SUITE
  • Selling Point: Visual workflow builder lets you chain multiple AI operations together (generate → edit → upscale → stylize) without any coding
  • Strong AI character consistency and avatar creation — ideal for influencer content pipelines
  • Best for: creators building repeatable AI workflows, AI influencer persona development
+ VIEW PROMPT TEMPLATE & SAMPLE
OPENART CHARACTER CONSISTENCY
Consistent character portrait: [CHARACTER DESCRIPTION], same face across all generations, [STYLE: editorial/commercial/cinematic], [LIGHTING], photorealistic, high detail. Use OpenArt's "consistent character" workflow with a reference image uploaded as the anchor for best identity lock across outputs.
WORKFLOW BUILDER: OpenArt's recipe system is the most accessible way to create repeatable AI image pipelines — no technical knowledge needed.
🎬
AI Video Studio · Free Tier · 2026
CHECKED 22 AUG 2026
AI VIDEO STUDIO
  • Selling Point: Cinematic-quality text-to-video and image-to-video with a free plan — accessible entry point for creators new to AI video
  • Strong realistic human motion and facial consistency, competitive with Kling and Runway
  • Best for: cinematic short clips, AI influencer video, product showcases, social content
+ VIEW PROMPT TEMPLATE & SAMPLE
POLLO AI VIDEO PROMPT
Cinematic close-up shot, [SUBJECT] in natural environment, photorealistic skin texture, natural ambient lighting, shallow depth of field, smooth camera movement, 4K quality, no scene cuts, continuous motion throughout.
FREE TIER: Pollo AI’s pricing page lists a Free plan of 20 credits, which it counts as 4 videos on Pollo 2.5 at 720p, and no daily allowance — enough to prototype a concept before committing to a paid Kling or Runway subscription.Pollo AI, Plans & Pricing, read in a browser 23 Sep 2026: compare table, “Free”, “20 credits”; Pollo 2.5 720P, “5 credits/5s”, “4 videos”. The rendered page does not mention daily credits.First-hand: until 23 Sep 2026 this tip said “Pollo AI offers daily credits”, and the card called the free tier “generous”; no source had been recorded for either.

NEWSLETTER EXCLUSIVEFuture exclusive prompt packs and AI production workflows sent directly to subscribers. Subscribe free ↗

The Google suite, in one place

Google now ships roughly a dozen separate AI tools, and the confusing part is that several overlap. Grouped by what they actually do rather than by launch order:

◈ WHY THIS SECTION EXISTS

One vendor, many products, no obvious map. Three of these do coding, two do video, two do design. Listing them is not endorsing them — it is making the overlap visible so you can pick one and ignore the rest.

Coding and agents

Jules
Google · Async coding agent
CHECKED 22 AUG 2026
  • What it does: handles coding tasks in the background, explains its changes, fixes bugs, prepares code for review
  • The asynchronous pattern — you do not sit and watch it
  • Review still required; see diagnosing a run
Antigravity
Google · Agent-first workspace
CHECKED 22 AUG 2026
  • What it does: a coding workspace where you assign tasks to agents across editor, terminal and browser
  • Broader permission surface than a normal editor — three environments, not one
  • Worth setting tiers before granting it much
Opal
Google · Plain-language mini apps
CHECKED 22 AUG 2026
  • What it does: turns plain-language descriptions into small apps with editable visual workflows
  • The visual workflow is the useful part — you can see what it built
  • Closest thing here to vibe coding with a safety rail

Design and interface

Stitch
Google · Text to UI
CHECKED 22 AUG 2026
  • What it does: text prompts into editable UI designs and front-end code
  • Idea to working interface without opening a design tool
  • Generated front-end code still needs the tier-one checks
Mixboard
Google · AI mood board
CHECKED 22 AUG 2026
  • What it does: collects ideas, generates visuals, remixes them into a direction
  • Direction-finding rather than production
  • Useful before you start prompting seriously, not after

Video and image

Veo 3 STRONGEST NATIVE AUDIO
Google · Text/image to video
CHECKED 22 AUG 2026
  • What it does: realistic video with synchronised sound, dialogue and cinematic motion from text or an image
  • 48kHz synchronised dialogue with sub-120ms lip-sync — as reported by reviewers. One July 2026 roundup calls it the only widely available model with itBuild Fast with AI, Google Veo 3.1 Review (2026), read at source 23 Sep 2026: “Veo 3.1 is the first practical AI video model to generate synchronized audio at 48kHz directly from a text prompt, with lip-sync accuracy within 120ms.” Pinggy, Best Video Generation AI Models in 2026, updated 14 Jul 2026: “Veo is the only widely available model generating 48kHz synchronized dialogue”. Both are third-party reviews, not Google documentation; first checked 22 Aug 2026 against these and oTechWorld. Until 23 Sep 2026 this line said “verified”.
  • But native clips top out at 8 seconds (some sources say 15); longer needs the extend workflow. It also accepts only ~3 reference images against Seedance's 50
  • The migration path if you were using Sora for realism and native audio
Flow
Google · AI filmmaking
CHECKED 22 AUG 2026
  • What it does: builds scenes, manages characters, adjusts camera angles, maintains consistency across a story
  • The consistency claim is the interesting one — see character consistency
  • Camera control assumes you know the vocabulary: camera grammar
Nano Banana 2
Google · Image generation · REPLACED IMAGEN 4
CHECKED 16 SEP 2026
  • What changed: Imagen 4 was replaced by Nano Banana 2. Google lists all three Imagen 4 models with a shutdown date of 17 August 2026, and names gemini-3.1-flash-image, displayed as Nano Banana 2, as the replacement.Read at source 16 Sep 2026: Google, Gemini API deprecations and models pages. An earlier version of this card listed Imagen 4 as current, checked 22 Aug 2026 — five days after its shutdown date.
  • Text rendering differs by model and still needs checking: see text inside images
  • Comparison on the comparison page

Research and knowledge

Gemini Notebook RENAMED FROM NOTEBOOKLM
Google · Grounded workspace
CHECKED 22 AUG 2026
  • What it does: turns uploaded notes, PDFs, papers and transcripts into a workspace for questions and summaries
  • Renamed Gemini Notebook in July 2026. Same product, existing notebooks intact, free tier continuesGoogle, "NotebookLM is now Gemini Notebook", announced 16 July 2026, read at source 11 Sep 2026: “Gemini Notebook and NotebookLM are the same product. You do not need to move your notebooks, recreate your sources, or learn a replacement tool”. Previously cited here to Man of Many, Aug 2026.
  • "Grounded" is the operative word — answers come from your sources rather than from training
  • This is context engineering as a product
Pomelli
Google · Brand campaign generation
CHECKED 22 AUG 2026
  • What it does: analyses a business website to build its "Business DNA", then generates tailored campaigns and branded content
  • Closest match to the creation-as-a-service model
  • Whether a site is enough to infer a brand from is the open question
◈ AVAILABILITY — CHECKED 22 AUG 2026

Several of these are experiments with real limits, and vendor lists rarely say so:

  • Mixboard — US-only, waitlist
  • Opal — US-first at launch, now reported in 160+ countries
  • Stitch — regional beta. Originally Galileo AI, acquired by Google and rebranded
  • Antigravity — free during preview, but free requests reported cut to roughly 20/day. It is a fork of VS Code
  • Pomelli — beta, paid tiers signalled, reported in 170+ countries

Anything under Google Labs is explicitly an experiment. Whisk and ImageFX did not simply close: Google announced on 25 February 2026 that they “are moving directly into Flow”, and both old addresses now redirect to Flow. Free is not a permanent promise, and neither is existence — the Sora shutdown at the top of this page is the same lesson from a different vendor.Man of Many, Aug 2026 · GitHub list-of-free-google-ai-tools · Wikipedia, Google Antigravity. Whisk and ImageFX: Google, Flow updates, 25 Feb 2026, read at source 16 Sep 2026; labs.google/whisk and labs.google/fx/tools/image-fx both redirect to flow.google.com, checked 16 Sep 2026. An earlier version said both “closed in April 2026”, a date no Google source gives.

◈ THE HONEST CAVEAT

This is one vendor's catalogue at one moment. Several of these overlap, and some may merge or be discontinued — as Sora is being at OpenAI: the Sora app closed on 26 April 2026 and the API shuts down on 24 September 2026.OpenAI Help Center, What to know about the Sora discontinuation, read in a browser 17 Sep 2026: “The Sora web and app experiences were discontinued on April 26, 2026.” and “The Sora API will be discontinued on September 24, 2026.”

Pick by what you need rather than by what exists. The decision ladder is on prompt, context, skill or fine-tune, and what it costs to run is on token economics.

TAKEAWAY

Google's AI tools overlap, and several are experiments. Pick one by what you need, check where it is available, and do not treat free as permanent.

ABOUTMETHODVERIFYPRIVACYCONTACTINDEXAI PROMPT GENEER · EVERY ARTICLE CARRIES ITS OWN CHECKED DATE