{"slug": "qwen-3-0-image-pro", "title": "Qwen 3.0 Image Pro", "summary": "Alibaba Cloud's Qwen team released Qwen-Image-3.0-Pro, an image generation model supporting up to 4.5k tokens of input, text rendering as small as 10px, and native support for 12 languages and 20+ fonts. The model is priced at $0.003 per 1K or 2K image input and $0.04 per 1K image output or $0.075 per 2K image output, with a rate limit of 1 request per minute. It features prefix completion, function calling, context cache, and structured outputs, positioning it as a productivity tool for generating complex layouts like newspapers and menus.", "body_md": "### Qwen-Image-3.0-Pro\n\n[Try AI](https://www.qwencloud.com/try-ai/chat?models=qwen-image-3.0-pro)\n\n[Add to Compare](/compare?models=qwen-image-3.0-pro)\n\n## Overview\n\nRich content: Supports input of up to 4.5k tokens and dense information layout with images-within-images, enabling complex layouts like newspapers, storyboards, menus, and exam papers to be generated in a single pass. Authentic detail: Supports precise rendering of text as small as 10px, and vividly reproduces fine details such as micro-expressions, pores, and individual strands of hair—approaching the quality of real photography. Deep knowledge: Supports native rendering of 12 languages and 20+ fonts, realistic simulation of mainstream interfaces such as web pages, games, and live streams, fully incorporating external knowledge. Qwen-Image-3.0-Pro isn't just pursuing \"good looks\"—it's pursuing \"usefulness\", making image generation a truly deployable productivity tool.\n\n#### Input\n\n#### Output\n\n## Features\n\n#### Prefix Completion\n\nEnable Partial Mode when calling the Qwen API to make the model continue strictly from your provided prefix text.[View docs](https://docs.qwencloud.com/developer-guides/text-generation/partial-mode)\n\n#### Function Calling\n\nUse function calling to connect large language models with external tools and systems.[View docs](https://docs.qwencloud.com/developer-guides/text-generation/function-calling)\n\n#### Cache\n\nContext Cache stores shared prefixes for long-context requests to reduce repeated computation, improve latency, and lower cost.[View docs](https://docs.qwencloud.com/developer-guides/text-generation/context-cache#implicit-cache)\n\n#### Structured Outputs\n\nStructured Outputs help ensure the model returns a JSON string in the expected format.[View docs](https://docs.qwencloud.com/developer-guides/text-generation/structured-output)\n\n## Pricing\n\n- 1K Image Input$0.003Per image\n- 2K Image Input$0.003Per image\n- 1K Image Output$0.04Per image\n- 2K Image Output$0.075Per image\n\n## Rate Limits\n\n- RPMRequests Per Minute1\n\n## API Reference\n\n[Call API](https://home.qwencloud.com/api-keys)\n\n```\ncurl --location 'https://dashscope-intl.aliyuncs.com/api/v1/services/aigc/multimodal-generation/generation' \\\n--header 'Content-Type: application/json' \\\n--header \"Authorization: Bearer $DASHSCOPE_API_KEY\" \\\n--data '{\n    \"model\": \"qwen-image-3.0-pro\",\n    \"input\": {\n        \"messages\": [\n            {\n                \"role\": \"user\",\n                \"content\": [\n                    {\n                        \"text\": \"A vertical outdoor portrait photograph with a warm, film-like afternoon street atmosphere, featuring a beautiful young adult woman looking back over her shoulder at the camera with a joyful toothy smile, her long thick wavy black hair catching the golden rim light, her fair skin, delicate eyebrows, bright eyes, and soft coral-red lips creating a radiant expression. She wears a simple black backless dress with thin spaghetti straps, showcasing her back, and cradles a large, lush bouquet of orange, apricot, pink, and pale peach roses in her arms, creating a sharp contrast against her dress. The top-left of the frame is covered with dark green vines and small orange flowers draping naturally, partially obscuring a matte dark blue signboard with the white Gothic text \\\"Il Messaggero\\\". Below the sign is a blurred glass newsstand window with black metal frames showing hints of newspapers and magazines. The right background features strong golden hour backlighting streaming down a warm-toned, sun-drenched city street, with buildings blurred into soft beige-gray shapes, creating a beautiful bokeh effect and a blurry red traffic sign in the far distance. The entire image has a cinematic, romantic, and bright urban stroll atmosphere, characterized by soft contrast, fine film grain, a shallow depth of field, and stunning backlit highlights.\"\n                    }\n                ]\n            }\n        ]\n    },\n    \"parameters\": {\n        \"prompt_extend\": true\n    }\n}'\n```\n\n", "url": "https://wpnews.pro/news/qwen-3-0-image-pro", "canonical_source": "https://www.qwencloud.com/models/qwen-image-3.0-pro", "published_at": "2026-08-05 14:53:29+00:00", "updated_at": "2026-08-05 15:51:58.641410+00:00", "lang": "en", "topics": ["generative-ai", "ai-products", "artificial-intelligence"], "entities": ["Alibaba Cloud", "Qwen", "Qwen-Image-3.0-Pro"], "alternates": {"html": "https://wpnews.pro/news/qwen-3-0-image-pro", "markdown": "https://wpnews.pro/news/qwen-3-0-image-pro.md", "text": "https://wpnews.pro/news/qwen-3-0-image-pro.txt", "jsonld": "https://wpnews.pro/news/qwen-3-0-image-pro.jsonld"}}