Qwen Image 3.0 series now available on Atlas — Standard + Pro for image generation and editing

Qwen Image 3.0 series now available on Atlas — Standard + Pro for image generation and editing

The Qwen-Image 3.0 series is now available on Atlas Cloud, with both Qwen-Image 3.0 and Qwen-Image 3.0 Pro.

The series combines image generation and multi-image editing in a single model family, while the Pro tier delivers higher image quality and stronger consistency for more demanding production use cases.

What Qwen-Image 3.0 ships with

  • Support for up to 3 reference images for editing and subject-driven generation
  • Strong instruction following for complex prompts
  • Native 1K / 2K generation
  • Consistency-preserving edits for outfit changes, background replacement, and local repainting
  • Qwen-Image 3.0 Pro further improves detail, color, and consistency in complex scenes

Pricing

Pay-per-image, with failed generations not charged.

Qwen-Image 3.0

  • Generation: $0.03 / image
  • Reference image input: $0.003 / image

Qwen-Image 3.0 Pro

  • 1K generation: $0.04 / image
  • 2K generation: $0.075 / image
  • Reference image input: $0.003 / image

Billing is based on the actual number of generated images returned, along with the corresponding 1K / 2K output tier.

Try it on Atlas Cloud

Qwen-Image 3.0
https://www.atlascloud.ai/models/qwen-image-3.0/text-to-image

Qwen-Image 3.0 Pro
https://www.atlascloud.ai/models/qwen-image-3.0-pro/text-to-image

Both endpoints are live on Atlas Cloud now, with Playground and API access available.

u/atlas-cloud — 1 day ago

Seedance 2.5 now supports up to 4K on Atlas Cloud with ESR resolution upgrades and discounted pricing

We’ve just expanded the Seedance 2.5 resolution options on Atlas Cloud.

Native 1080p generation is now officially supported, so you can generate directly at the model's original 1080p resolution.

On top of that, we've rolled out our in-house ESR super-resolution pipeline. It takes Seedance 2.5 output and enhances it up to 1080p at 60 FPS, 1440p, or 4K, giving you a more flexible and cost-efficient path to higher resolution than generating everything natively at the top tier.

Several of these resolution options are currently discounted

What’s new

Seedance 2.5 on Atlas Cloud now includes:

  • Official 1080p
  • 1080p ESR
  • 1080p ESR at 60 FPS
  • 1440p ESR
  • 4K ESR

Resolution options & pricing

Output Discounted Price/sec List Price/sec Discount
480p $0.14 $0.14
720p $0.30 $0.30
1080p $0.42 $0.53 20% off
720p ESR $0.25 $0.25
1080p ESR $0.45 $0.54 17% off
1080p ESR 60 FPS $0.57 $0.69 17% off
1440p ESR $0.76 $1.32 44% off
4K ESR $1.70 $2.83 40% off

🔥The resolution discounts for official 1080p are for a limited time, until Sept.17th 14:00 (UTC+8)

Try it Now

https://www.atlascloud.ai/models/seedance-2.5

u/atlas-cloud — 4 days ago

Atlas Cloud Weekly Update — August 10, 2026: Kling 3.0, Seedream 5.0 and Free AI Tools

Kling v3.0 Pro Motion Control

Kling v3.0 Pro Motion Control is now available on Atlas Cloud. Upload a character image and a reference motion video, then transfer the dance, action, or gesture performance to your character with smooth, realistic movement.

Key features

  • Transfers motion from a driving video to a character image or source video
  • Supports dance, gestures, action sequences, and performance retargeting
  • Preserves the character’s appearance while following the reference movement
  • character_orientation: image supports videos up to 10 seconds
  • character_orientation: video supports videos up to 30 seconds
  • Optional prompt and negative prompt for scene and style control
  • Option to keep the original audio from the reference video
  • Minimum billable duration: 3 seconds

Pricing

  • Kling v3.0 Std Motion Control: starting at $0.126/sec
  • Kling v3.0 Pro Motion Control: starting at $0.153/sec

Standard vs. Pro

Kling v3.0 Std Motion Control

  • More cost-efficient for rapid iteration and batch generation
  • Suitable for social clips, previews, storyboards, and motion tests

Kling v3.0 Pro Motion Control

  • Higher-fidelity motion transfer and detail preservation
  • Better suited for final renders, cinematic content, and production work

API access

Standard: https://www.atlascloud.ai/models/kwaivgi/kling-v3.0-std/motion-control

Pro: https://www.atlascloud.ai/models/kwaivgi/kling-v3.0-pro/motion-control

Atlas AI Tools

Atlas Cloud’s official website has updated its AI Tools section. The section focuses on authentic emotion and emerging trends, curating high-potential creative themes and visual references to help creators capture timely inspiration and transform sports culture, social conversations, and other trending topics into cinematic visuals.

Tool page
https://www.atlascloud.ai/ai-tools

Pricing
Free: Image Upscaler, Image Background Remover, Object Eraser, etc.

Limited Free: AI Hairstyle Changer, AI Clothing Changer, AI Aging Photo, etc.

Seedream v5.0 Pro Layer Decomposition

Seedream v5.0 Pro Layer Decomposition turns a flat image into independently editable layers. The generated layers can be moved, scaled, recomposed, and reused in downstream design workflows.

Features

  • Decompose one image into 2–20 editable layers
  • Return background and visual elements as transparent PNGs
  • Reconstruct areas hidden behind foreground objects
  • Use prompts to control the number and structure of layers
  • Preserve the original composition while separating design elements
  • Useful for posters, banners, product creatives, UI layouts, and story covers

Pricing
$0.405 per output image

Model page
https://www.atlascloud.ai/models/bytedance/seedream-v5.0-pro/layer-decomposition

All three updates are available through Atlas Cloud’s unified AI infrastructure.

u/atlas-cloud — 10 days ago
▲ 36 r/AtlasCloudAI+1 crossposts

🔥 Limited-Time Lowest Prices: Seedance 2.0 Mini & Fast

Seedance 2.0 Mini is 30% off and Seedance 2.0 Fast is 20% off on AtlasCloud, effective now:

  • 💰 Mini from ≈$0.039/sec (was $0.056), Fast from ≈$0.072/sec (was $0.09)
  • Applied automatically to every call — no coupon, no minimum spend
  • 🎬 Same native audio-video generation, unlimited concurrency, no queuing — now at the lowest price you'll find anywhere

🗓️ Offer runs Aug 7, 2026 6:00 AM → Sep 7, 2026 6:00 AM (UTC)

🔗 Try it now: https://www.atlascloud.ai/models/seedance2

reddit.com
u/atlas-cloud — 13 days ago
▲ 25 r/AtlasCloudAI+1 crossposts

Seedance 2.5 is live on Atlas Cloud: 30s single-take video, 50 reference inputs, 4K

Seedance 2.5 is now live on the Atlas Cloud API. What is new over 2.0:

- A true 30 second single take, no stitching

- Up to 50 multimodal reference inputs across image, video, and audio in one generation

- New 3D camera occlusion control for depth-aware motion

- Native 4K with synced audio, and stronger prompt adherence

Same API key and the same endpoint as the rest of the Seedance line. You point at the new model id and go: bytedance/seedance-2.5/text-to-video (also image-to-video and reference-to-video).

Model page and docs: https://www.atlascloud.ai/models/seedance-2.5

u/atlas-cloud — 14 days ago

FLUX 3 now available on Atlas — $0.187/sec at 720p, native audio, pay-per-use

FLUX 3 is now available on Atlas.

Pricing (pay-per-use, no subscription required):

  • Text-to-Video: $0.187 / sec at 720p, $0.319 / sec at 1080p
  • Image-to-Video: $0.187 / sec at 720p, $0.319 / sec at 1080p (starts from your image, billed by output duration)

What FLUX 3 ships with (developed by Black Forest Labs):

  • Native audio generation — video and audio produced together in a single request, no separate dubbing/SFX pass
  • Explicit duration control — set 5–20 sec directly, keyframe-accurate
  • Flexible aspect ratio — auto (matches input/prompt) or pick from 21:9, 2:1, 16:9, 4:3, 1:1, 3:4, 9:16
  • 720p or 1080p output
  • Adjustable safety tolerance from strict (0) to permissive (4)
  • I2V mode animates any starting image (photo, illustration, render) into motion while preserving its native aspect ratio on auto

API access:

Use cases that map well to FLUX 3 strengths:

  • short-form social clips that need synced audio without a separate audio pass
  • concept visualization / storyboarding straight from a text prompt
  • photo-to-motion for product shots or portraits (I2V)
  • marketing and promo clips generated directly from an existing still asset

Drop questions about duration/aspect-ratio combos or how the native audio syncs to prompts in this thread.

u/atlas-cloud — 15 days ago

Weekly Model Update — New Models Now Live on Atlas

This week’s lineup adds a new multimodal flagship, native-audio video generation, super-resolution tools, controllable reference-to-video, and a new 2K image model.

Qwen3.8 Max

A next-generation flagship model for advanced reasoning, coding, and multimodal AI applications.

Features:

  • Advanced reasoning and long-context workflows
  • Strong coding and agentic task support
  • Multimodal input for text, image, and video understanding
  • Suitable for research, software engineering, and complex automation

Pricing: $2 / $6 per 1M tokens
Model page: https://www.atlascloud.ai/models/qwen/qwen3.8-max

Grok Imagine Video v1.5

Generate short videos directly from text prompts with native synchronized audio.

Features:

  • Text-to-video generation from a single prompt
  • Native synchronized audio
  • Supports clips up to 15 seconds
  • 480p, 720p, and 1080p output
  • Useful for cinematic scenes, social clips, dialogue, and sound-design experiments

Pricing: $0.08/sec
Model page:

Text-to-Video: https://www.atlascloud.ai/models/xai/grok-imagine-video-v1.5/text-to-video

Reference-to-Video: https://www.atlascloud.ai/models/xai/grok-imagine-video-v1.5/reference-to-video

Tencent Image Upscaler

Tencent’s MPS-powered image super-resolution model for enhancing low-resolution stills.

Features:

  • Image super-resolution and detail enhancement
  • Useful for restoring small or compressed images
  • Suitable for product assets, archived images, thumbnails, and creative upscaling

Pricing: $0.01/image
Model page: https://www.atlascloud.ai/models/tencent/image/upscaler

Tencent Video Upscaler

Tencent MPS video quality enhancement for upscaling and restoring footage.

Features:

  • Scene-aware video super-resolution
  • Upscale source footage up to 8K
  • Supports enhancement for common, UGC, short-series, AIGC, and old-film content
  • Optional frame-rate interpolation

Pricing: $0.018/sec
Model page: https://www.atlascloud.ai/models/tencent/video/upscaler

BytePlus Video Upscaler

BytePlus AI MediaKit video quality enhancement for upscaling and restoring source footage.

Features:

  • Upscale and enhance source video up to 8K
  • Scene-aware processing for common, UGC, short-series, AIGC, and old-film footage
  • Optional frame-rate interpolation
  • Suitable for restoration, archive footage, short-form video, and production finishing

Pricing: $0.018/sec
Model page: https://www.atlascloud.ai/models/byteplus/video/upscaler

MiniMax H3

MiniMax H3 is a general-purpose multimodal video model designed to handle text, image, video, and audio context in one workflow.

Features:

  • 5–15 second video generation
  • 24 FPS output
  • Aspect ratios from 21:9 to 9:16
  • Prompt-driven character and background changes
  • Dialogue rewriting and voice-reference workflows
  • Designed for cinematic production, storytelling, and multimodal video pipelines

Pricing: $0.14/sec
Model page:

Text-to-Video: https://www.atlascloud.ai/models/minimax/h3/text-to-video

Image-to-Video: https://www.atlascloud.ai/models/minimax/h3/image-to-video

Reference-to-Video: https://www.atlascloud.ai/models/minimax/h3/reference-to-video

Wan 2.7 Spicy Reference-to-Video

Generate one continuous video from one to four reference images with strong subject binding and prompt control.

Features:

  • Supports 1–4 image references
  • Designed for subject and character consistency
  • Reference images can guide people, objects, or visual identity
  • Image references only; video and audio references are not accepted
  • Useful for character-driven clips, fashion visuals, and creative transformations

Pricing: $0.10/sec
Model page: https://www.atlascloud.ai/models/atlascloud/wan-2.7-spicy/reference-to-video

Youchuan V8.2

A prompt-driven image model that returns multiple variations with native 2K output.

Features:

  • Four image variations per prompt
  • Native 2K HD output
  • Style-reference support
  • Aspect-ratio, stylize, chaos, and weird controls
  • Useful for concept exploration, campaign variations, and visual ideation

Pricing: $0.086/image
Model page:

Text-to-Image: https://www.atlascloud.ai/models/youchuan/v8.2/text-to-image

Image-to-Image: https://www.atlascloud.ai/models/youchuan/v8.2/image-to-image

Image-to-Video: https://www.atlascloud.ai/models/youchuan/v8.2/image-to-video

Try the new models in Playground and share what you build. Feedback on prompt behavior, output quality, and real-world workflows is welcome.

reddit.com
u/atlas-cloud — 16 days ago

Grok Imagine Video v1.5 is now on Atlas, three video routes from $0.08

Grok Imagine Video v1.5 is now available on Atlas across three video routes: Text-to-Video, Reference-to-Video, and Image-to-Video.

The Image-to-Video route starts from a single frame and follows a natural-language motion prompt. It supports clips up to 15 seconds, output at 480p, 720p, or 1080p, plus native synchronized audio for dialogue, lip sync, sound effects, and ambient music.

What ships with Grok Imagine Video v1.5:

- Text-to-Video for prompt-driven short-form video

- Reference-to-Video for directing a generation with visual references

- Image-to-Video for animating a starting frame with a motion prompt

- Native audio generation in the Image-to-Video route

- 1 to 15 second Image-to-Video clips

- 480p, 720p, and 1080p output options

- Standard API access through the same Atlas setup

Pricing:

- Starting at $0.08 per run in the current Image-to-Video Playground

- Check the selected model page for the current rate of each route before generating

API access:

- Text-to-Video:

https://www.atlascloud.ai/models/xai/grok-imagine-video-v1.5/text-to-video

Use cases that fit these routes:

- Prompt-driven product and social clips from Text-to-Video

- Character, object, and style guided scenes from Reference-to-Video, using cleared reference assets

- Turning a hero still into a short scene with motion and synchronized sound through Image-to-Video

- Comparing a starting frame and its animated result with a Playground example video

For the media, attach one real Playground Image-to-Video result with its source frame beside it. The difference in motion, sound, and framing tells the story quickly.

Use the thread for endpoint setup and prompt-specific workflow notes.

u/atlas-cloud — 21 days ago

Youchuan V8.2 (Midjourney) now available on Atlas — $0.086 per 4-image generation, $0.343 per 4-video generation, native 2K HD, text-to-image + image-to-video

Youchuan V8.2 (Midjourney) is now available on Atlas, covering both text-to-image and image-to-video. Youchuan is Midjourney's officially licensed distribution platform in China, operated by Xiaochuan Creative (Shanghai) under the "Midjourney China Lab" branding, not a third-party clone. It's the same underlying Midjourney technology.

Pricing (pay-per-use, no subscription required):

  • Text-to-Image: $0.086 per generation (returns 4 images, native 2K HD optional) — about $0.021 per image if you keep all 4 outputs
  • Image-to-Video: $0.343 per generation (returns 4 × 5-second videos at 480p or 720p) — about $0.086 per 5-second clip

What Midjourney V8.2 ships with (per spec):

  • Native 2K HD output at 2048px with no separate upscaling step, roughly 3x faster and cheaper than V8's upscale path
  • An estimated 4-5x faster generation overall from a GPU-native PyTorch rewrite (Midjourney-stated figure)
  • More reliable in-image text rendering, using quoted strings in the prompt to specify the intended text
  • Stronger prompt-following, needing less prompt-engineering to hit a target composition
  • Restored image conditioning: image prompts and image weights, backward compatible with V7 style references (srefs), moodboards, and personalization profiles
  • Image-to-video mode animates a single input image into four 5-second clips at 480p or 720p with adjustable motion intensity
  • Part of a larger Midjourney V8.2 family on Atlas: image-to-image, blend, style-transfer, and remove-background, each available as its own endpoint

API access:

Use cases that map well to Midjourney V8.2's strengths:

  • batch concept art exploration (4 image variants per generation without rerolling)
  • native 2K assets for print or marketing without a separate upscale pass
  • turning a single hero image into short-form video content via I2V
  • iterative style-reference workflows carried over from V7 (sref, moodboards, personalization)

Drop questions about prompt portability from official Midjourney to Youchuan V8.2 on Atlas, or specific aesthetic comparisons, in this thread.

u/atlas-cloud — 22 days ago

Doubao Seed Character now available on Atlas — $0.2/$0.8 per M tokens in/out, 131K context, multimodal input

Doubao Seed Character is now available on Atlas as a flagship LLM.

Pricing (token-based, pay-per-use):

- Input: $0.2/M tokens

- Output: $0.8/M tokens

- Context window: 131.07K tokens

- Max output: 32.77K tokens

- Cache-Based and Gradient-Based pricing modes supported

What Doubao Seed Character ships with (per ByteDance spec):

- Premium reasoning and coding performance

- Multimodal input (text and image)

- Text output

- Enterprise-grade performance at scale

- 131K context window with 32.77K max output

API access:

https://www.atlascloud.ai/models/bytedance/doubao-seed-character-260628

Use cases that map well to Doubao Seed Character's strengths:

- long-context reasoning and coding tasks

- multimodal workflows that need image understanding alongside text input

- high-volume agent pipelines where cache-based pricing cuts repeated-context cost

- enterprise workloads needing consistent performance at scale

Drop questions about benchmark comparisons or prompt/context portability from other Doubao variants in this thread.

reddit.com
u/atlas-cloud — 22 days ago

Half this prompt just tells Seedream 5.0 Pro to keep the skin imperfect on purpose

The giveaway on an AI portrait is skin that is too perfect, so the interesting move in this one is spending half the prompt telling Seedream 5.0 Pro to keep the imperfections: real pores, fine peach fuzz, natural tone variation, soft highlights, and explicitly no plastic airbrushing.

It is a plain, bright scene. A young woman leaning on a sunny kitchen table, white off-shoulder top, a fruit bowl, plants blurred through the window, that high-key everyday warmth. Original synthetic character, no real-person likeness. Nothing dramatic is happening, which is exactly why the skin has to carry it.

Seedream 5.0 Pro held it. Natural texture under soft window light, an 85mm shallow-focus falloff, the face sharp while the kitchen dissolves behind, no waxy influencer sheen. The long Avoid list at the end of the prompt is doing as much work as the description. That is where you kill the CG look.

Prompt, Avoid list included, below.

u/atlas-cloud — 25 days ago
▲ 6 r/AtlasCloudAI+1 crossposts

111 Seedance prompts and a skill for the community, built to switch to 2.5 on day one

For everyone in this community waiting on Seedance 2.5, we put together something to use in the meantime and to make launch day painless: awesome-seedance-2.5-prompts-skills, a curated library of 111 cinematic Seedance prompts plus an installable agent skill.

Seedance 2.5 is expected in August, and Atlas Cloud is one of the first official API launch partners. Rather than have everyone scramble to relearn it on day one, we wrote the whole thing to the 2.5 paradigm now, up to 50 multimodal references, long single takes, synced audio, local region editing. The prompts and the skill already speak that language, but the skill runs on Seedance 2.0 today and checks which model a provider actually exposes instead of assuming 2.5 is live. When 2.5 lands on Atlas, the same prompts and skill route to it automatically, so nothing you build now goes to waste.

The skill turns a brief or a rough prompt into a production-ready Seedance prompt, builds and reviews a Seedream 5.0 Pro storyboard when a shot needs multi-shot planning, then generates the video and polls it to completion.

Install:

npx skills add AtlasCloudAI/awesome-seedance-2.5-prompts-skills --skill seedance-2-5-skill

CC BY 4.0 and PRs are open. If you have a Seedance prompt this community would use, add it.

Repo: https://github.com/AtlasCloudAI/awesome-seedance-2.5-prompts-skills

Browse the prompts: https://www.atlascloud.ai/prompts-hub/seedance-2-5-prompt

u/atlas-cloud — 28 days ago

Heads up: Seedream 5.0 Pro is 20% off on Atlas for two weeks, starting today

Quick one for the community. Starting 00:00 UTC today, July 23, Seedream 5.0 Pro is 20% off on Atlas Cloud for a two-week run.

If you already generate stills here for your Seedance pipelines, this is the window to stock up on character sheets, style references, and first frames. It runs on the same one-key OpenAI-compatible API as Seedance 2.0, so the stills flow straight into animation without switching anything.

It is the strongest model we have for holding an art style across reposes and keeping a character consistent shot to shot. Two weeks, then pricing goes back to normal.

https://www.atlascloud.ai/models/seedream-5.0-pro

u/atlas-cloud — 29 days ago

Seedream 5.0 Pro is 20% off on Atlas Cloud for two weeks, starting 00:00 UTC July 23.

It is our strongest image model for holding an art style across reposes and keeping a character consistent shot to shot, and it runs on the same one-key API as Seedance 2.0, so you go from stills to video without switching providers.

Good window to build a look and lock it in.

Try it 👉 https://www.atlascloud.ai/models/seedream-5.0-pro

u/atlas-cloud — 29 days ago

Kimi K3 just landed at #4 on Agent Arena, and it is live on Atlas Cloud today

Arena.ai just published their latest Agent Arena update, and Kimi K3 landed at #4 on the agentic leaderboard, in the same band as Claude Opus 4.8 and GPT 5.6 Sol. Agent Arena scores models on real long-horizon agentic work: web search, filesystem, and terminal tools running full workflows like writing code, building apps, and analyzing documents, measured across thousands of live sessions.

A few things stood out in the breakdown. Kimi K3 ranks first on confirmed task success, the explicit "yes that worked" signal from users, and posts a strong result on praise versus complaint. It still trails the field on steerability and on recovering from CLI errors, so it is not the pick for every job yet, but on actually getting long tasks finished it sits right at the front.

The part we care about here: Kimi K3 is live on Atlas Cloud today, through the same OpenAI-compatible API as the rest of the frontier. Much of that Agent Arena top list, Opus 4.8, GPT 5.6 Sol, Sonnet 5, GLM 5.2, Grok 4.5, and Kimi K3, sits behind one endpoint and one key, so you can route each task to whichever model wins it instead of locking to a single provider. Kimi K3's open weights are expected around July 27, and if they land on schedule it becomes the top open-weight model on the board.

Full leaderboard and methodology are in Arena.ai's post. You can try Kimi K3 on Atlas Cloud here: https://www.atlascloud.ai/models/moonshotai/kimi-k3

u/atlas-cloud — 1 month ago
▲ 15 r/MiniMaxH3AI+2 crossposts

A new Seedance 2.5 collaboration with Michael Owen brings his career-defining moments back to life, a preview of what the model is built to do

A new collaboration built on Seedance 2.5 is working with Michael Owen to revisit some of the defining moments of his career, the signature moves, the goals, the instincts that made him one of the most recognizable strikers of his generation. From those moments on the pitch to what else becomes possible once a model can work with a real career instead of a single clip, it's built around unlocking what used to be impossible.

It's also a good preview of where Seedance 2.5 is headed more broadly, and the direction is a real jump forward. Original generations are set to run up to 30 seconds, the longest single generation the line has offered so far, enough room to hold an actual narrative arc instead of a quick clip. It's also shaping up to be the first video model in the lineup released at 720p resolution, a real step up in fidelity for a line that keeps pushing further with each release.

Character consistency is getting a real push too, holding the same face and identity recognizable across an entire sequence of different moments instead of drifting scene to scene, which is exactly what a piece built around one real person's career actually demands. On top of that, the generations are shaping up to be more editable after the fact, adjusting a shot without having to regenerate the whole sequence from scratch.

Classic moments, now within reach, and just an early look at what the model can do once it's treating an entire career as material to work with.

u/atlas-cloud — 7 days ago
▲ 33 r/AtlasCloudAI+1 crossposts

Magic Pen: a marker pen touches the real street and things snap into flat-2D anime, one continuous POV shot on Seedance 2.0

Fun one. A first-person street-magic vlog where a marker pen touches real things and they snap into flat-2D anime, one continuous handheld shot, no cuts. Made on Seedance 2.0 (Mini works too). Full prompt below.

STYLE: live-action plus flat 2D anime-sticker composite, first-person POV street-magic vlog. Photoreal detail with the texture of real phone rear-camera footage, strong contrast between the photoreal city and the flat cartoon characters. One continuous handheld phone shot, no cuts, no scene transitions.

CAMERA: raw unstabilized handheld the whole way, walking bounce, slight arm sway, occasional autofocus hunting, real exposure shifts between bright sky and building shade, natural phone HDR color. Between targets the camera moves in quick whip pans that follow the pen.

LIGHT: late-afternoon sun from one consistent direction, every real object and every animated character drops a soft contact shadow matching that sun.

SCENE: one continuous city block walked end to end, an elevated track overhead at the start, a tree-lined sidewalk with pigeons, a bus lane, and a bus-stop bench further down. Lived-in everyday street, background passersby with natural motion.

THE PEN (magic law): the vlogger's real hand holds a black marker pen, always in frame, the visual guide connecting every beat. Every transformation follows the same sequence: the pen points, the voice says "Biu!", blue hand-drawn sketch lines wrap the target, ink spreads, the target becomes a flat-2D cel anime character with bold cartoon outlines, keeping the original's exact size, position, speed, direction and perspective, perfectly anchored to the real street. Characters keep flat sticker shading, never re-lit by real-world light, but everything they touch reacts with real physics.

BEATS:

00:00-00:03 walking POV under the elevated track, the hand raises the pen.

00:03-00:06 pen swings to a pigeon on the pavement, "Biu!", it becomes a cute hand-drawn cartoon bird, hops twice, flaps and flies off leaving sketch feathers that dissolve into ink particles.

00:06-00:10 the pen follows the particles to the bus lane where a bus passes, "Biu!", a fast sketch outline covers the bus and it becomes a giant flat-2D orange tabby cat keeping the bus's exact size, speed and perspective, padding down the lane, ears twitching, soft contact shadows on the asphalt.

00:10-00:12 the pen swings to a woman on the bus-stop bench, "Biu!", sketch lines trace her and she is redrawn as a vibrant anime cel illustration, same pose, she smiles warmly and waves once.

00:12-00:15 the vlogger flings the marker spinning into the sky with a final "Biu!", the camera tilts up as the pen draws a blue ink spiral, the ink blooms until the whole real sky becomes a hand-drawn anime sky, cel-shaded clouds, a hand-drawn sun, while the photoreal skyline stays real along the bottom. A small handwritten "THE END!" doodle pops in among the clouds and the frame freezes.

Made on Seedance 2.0: https://www.atlascloud.ai/models/bytedance/seedance-2.0/text-to-video (Seedance 2.0 Mini works too, about 10 seconds).

u/atlas-cloud — 1 month ago

Seedream 5.0 Pro vs GPT Image 2: cost, reference images, and editing control

Seedream 5.0 Pro vs GPT Image 2 is not winner-takes-all, so here is the honest split after running both. GPT Image 2 is attractive if you want a general-purpose image model inside the OpenAI ecosystem, and for cheap output-only images at low quality it is hard to beat.

Seedream 5.0 Pro gets more interesting the moment the workflow uses reference images and local edits. Where the two split:

  • output-only, simple, low-quality one-offs: GPT Image 2 Low is cheaper
  • reference-heavy product images, portraits, ad variants, and editing workflows: Seedream's reference handling and editing control pull ahead

The part that surprised me was the billing model, not the output. Seedream 5.0 Pro is a flat rate per image, the same whether you run text-to-image or image editing. GPT Image 2 is token-metered, and it always processes reference images at high fidelity, so edit-heavy requests run roughly 2 to 3 times the baseline. For the reference-driven editing work Seedream is built for, that gap compounds fast.

Seedream 5.0 Pro GPT Image 2
Billing model flat rate per image token-metered (image input + output tokens)
Per-image rate $0.054 flat, edit or text-to-image ~$0.006 low / ~$0.053 medium / ~$0.211 high (calculator estimates, not list)
Reference images up to 10 blended in one pass processed at high fidelity, adds tokens to every edit
Edit-heavy cost behavior stays flat runs ~2 to 3x baseline
Unit-cost predictability fixed per image varies with size, quality, and retries

So the useful benchmark is not cost per image, it is cost per usable edit. For a throwaway low-quality render, GPT Image 2 Low wins on raw cost. For high-quality, reference-heavy, iterative editing, Seedream 5.0 Pro comes in well under GPT Image 2's high tier and the flat rate keeps unit economics predictable.

We run Seedream 5.0 Pro on Atlas at that flat rate for both editing and text-to-image: https://www.atlascloud.ai/models/seedream-5.0-pro

reddit.com
u/atlas-cloud — 1 month ago

Two weeks with Seedream 5.0 Pro vs Nano Banana Pro. NBP still wins realism, but the editing gap surprised me

I kept seeing "Seedream 5.0 is a downgrade" after the Lite launch, so I went in expecting nothing and ran the Pro tier against Nano Banana Pro for two weeks of real work. Two things up front. Pro is a different model from the Lite build people were dunking on. And no, it doesn't dethrone NBP. But the comparison isn't the one everyone keeps having.

Where NBP still wins: pure photoreal realism, especially skin and faces, camera-effect believability, and single-shot hero images. If the deliverable is one gorgeous realistic frame, NBP is still my pick, and on faces it's not close.

Where Seedream 5.0 Pro pulled ahead, and where I didn't expect it to:

Dimension Seedream 5.0 Pro Nano Banana Pro
Photoreal realism (skin/faces) good, not class-leading best
Point-and-edit (change one region, rest locked) box / arrow / coordinates, multi-region in one pass prompt-level, less surgical
Layer separation (PSD-style) background + transparent element layers no
Reference images blended up to 10 fewer
In-image / small text 15 languages, small text usable good
Per-image price from $0.054 higher
Permissiveness more permissive stricter

The thing that actually changed my workflow was editing, not generation. On NBP I re-prompt the whole image and hope the part I liked survives the reroll. On Seedream 5.0 Pro I mark the one region I want changed and the rest stays put. For iterative client revisions that is the whole game, and the layer separation, where it splits a finished image into movable transparent layers, folds a Photoshop step into the generation itself.

Honest verdict: NBP for the hero realistic shot, Seedream 5.0 Pro for anything edit-heavy, iterative, multilingual, or cost-sensitive. If you got burned by 5.0 Lite, the Pro is worth a second look specifically for the editing, not to win a realism benchmark.

I run both off one key so switching between them is a model_id swap, nothing else changes. Seedream 5.0 Pro is here.

u/atlas-cloud — 1 month ago

Seedream 5.0 Pro is live: image generation you actually edit instead of reroll The new

Seedream 5.0 Pro is live on Atlas. It's ByteDance's flagship image model, and the reason it's worth a post isn't that the outputs look nicer, it's that the whole interaction changes. Image models have been slot machines: write a prompt, roll four, reroll if you don't like it. 5.0 Pro turns that into actual editing, where you point at what you want changed and the rest stays put.

Point-and-edit. Mark a region with a box, an arrow, even coordinates, and only that region changes. You can stack several edits in one pass, one instruction per marked area, and they don't bleed into each other.

Anchor positioning for grids. On a chessboard or a product shelf you can say "the piece bottom-left, one column in" and it edits exactly that cell. Positional editing like this used to be a disaster zone for image models.

Layer separation. It splits a finished image into a background layer plus N element layers, each a transparent PNG you can move, scale, recompose, or reuse in another scene. For ecommerce detail pages or marketing assets this folds the "cut the layers apart in Photoshop" step into generation. (This one rolls out within a week of launch.)

Material and color swap. Give it a hex code or a material, wood grain, leather, glass, satin, and it applies that to the target region with the structure preserved. Product colorways without fighting adjectives.

Multi-image blend. Up to 10 reference images, merging their objects, style and material into one target per your instruction.

Two more worth flagging. Dense text and infographics took a real jump, small-text rendering is much better now (not zero-typo, but usable for posters and ecommerce pages), and it does native prompts and in-image text across 15 languages including Arabic, Korean and Thai, the ones that used to come out as garbled glyphs.

And a useful one if you also make video: Seedream 5.0 output is trusted input into the Seedance family (2.5, 2.0, Fast, Mini). The image carries an invisible same-ecosystem marker, so Seedance skips the input-side real-person check instead of blocking it. Text-to-image output is auto-trusted, image-to-image after your account clears KYC. The usual flow is to generate a character or first frame in Seedream and feed it straight into Seedance for video, one clean pipeline. One boundary stays hard: this covers virtual, AI-generated subjects only, real human faces as input are still prohibited, and the review on the final video output runs as normal.

On tiers: Pro is the 2K control-and-quality tier, up to 2048x2048 at 1:1 and around 2.7K on the long edge at 16:9, from $0.054 an image, up to 10 reference images. If you need native 4K that's Seedream 5.0 Lite, which goes to 2K/3K/4K and takes up to 14 references.

It's on Atlas now on the same key as the rest of the image, video and LLM lineup, so adding it is a model_id swap, not a new integration. Model page with params, pricing and an in-browser try: https://www.atlascloud.ai/models/seedream-5.0-pro

u/atlas-cloud — 1 month ago