Skip to content
AETRONConstructor
My neuronetsRegistrySpacesDocsFAQ
Sign in
AETRON

A constructor for AI services on AETRON’s decentralised network of GPU miners.

ProductCreate a neuronetModel catalogueDatasetsSpacesMy payouts
ResourcesDocumentationAPITimings & statusesFAQ
Networkaetron.ioPrivacy & verification
© 2026 AETRONTestnet
Models/qwen-image
Q

Qwen-Image

Alibaba Qwen · Qwen-Image
Create a neuronet with this model →
DiffusionApache-2.0publicTrusted20B params
text-to-imagediffusionmmdittypographymultilingualdatacenter

20B MMDiT text-to-image. Best-in-class text rendering, English and Chinese.

Overview

Qwen-Image is a 20B Multimodal Diffusion Transformer (MMDiT) from the Qwen team, paired with a Qwen2.5-VL-7B text encoder. Its standout capability is text rendering inside the image — English and especially Chinese typography, where it leads the open models by a wide margin. Apache-2.0, fully commercial.

Deployment notes

  • 80GB VRAM class only (A100 80G / H100). bf16 weights alone are ~55GB: 20B transformer + 7.6B text encoder + VAE, and the runner loads the whole pipeline onto one device without offload.
  • ~58GB of weights to download before the first job.
  • Native resolution is 1328×1328; this network caps the request at 1024 — the window where diffusion verification is calibrated.
  • CFG knob: this pipeline uses true_cfg_scale, not guidance_scale. Supply a negative prompt and it runs at the pipeline default (4.0); leave it empty and guidance is off.
Create a neuronet with this model →
Overview
Task
Diffusion · DiffusionV1
Family
Qwen-Image
Tier
Trusted
Status
active
Runtime spec
dtype
Bf16
attention
Sdpa
compile
Eager
seed
0
gpu_arch
Sm90
Runs on
Datacenter
Hardware
Min VRAM
80 GB
Recommended
80 GB
Tensor parallel
≥1
TEE archs
Sm90
Verification defaults
Tier
B
Trust
Level1
Security
Standard
Privacy
P0
Min pulse score
95
License
spdx
Apache-2.0 ↗
Commercial use
✓
Redistribution
✓
Source & provenance
Repo
Qwen/Qwen-Image ↗
Revision
75e0b4be04
Added
2026-07-27
Verified
—