Model
Qwen-Image-3.0 Released: Supports 4.5k Token Input and 10px Small Text Rendering
Alibaba Tongyi Qianwen launches the third-generation foundational image generation model Qwen-Image-3.0, with core capabilities focused on "realism." It supports instruction input of up to 4.5k tokens and can generate a 3×3 grid layout containing 9 complex infographics in one go. The model can accurately render text as small as 10px and natively supports 12 languages. It can simulate mainstream interfaces such as web pages, games, and live streams, maintaining readability in dense layout scenarios like academic papers and newspapers.
Read the original (opens in a new tab)
News stream data aggregated by AI HOT