Qwen-Image-2.1 in 4 steps is here ⚡ @ViggleAI distilled Qwen-Image-2.1 into a 4-step turbo model, 6× faster, and holds up side by side with the full model ▶️ on Spaces https://hf.co/spaces/Viggle/Qwen-Image-2.1-viggle-turbo
#图像生成
#图像生成
今日 16 条
ViggleAI@ViggleAIAI 评分5050引用Hugging Apps@HuggingApps
Tripo@tripoaiAI 评分3535
Tencent Hy@TencentHunyuanAI 评分4444免费两周!Hy Image3.5 预览版已在 OnSolo 上线。短剧角色设定图。全动态视频游戏素材。关键帧。角色在每一集中保持一致。编辑是精修,而非重来。
引用OnSoloAI@OnSoloAIHy Image3.5 preview is on OnSolo. Exclusive. 5 refs. 2K. 2 weeks Members free.
Tencent Hy@TencentHunyuanAI 评分5959引用ComfyUI@ComfyUIHy Image3.5 preview is now available in ComfyUI. Professional-grade image generation, +30% win rate in human eval vs Hy Image3.0 → Text to image and Image to image in one model, up to 2K → Multilingual text, symbols, and small print that render correctly → Cinematic, comic, commercial photography, and illustration styles → Identity and product features that hold through scene, outfit, and style changes
Qwen@Alibaba_QwenAI 评分5959引用Arena.ai@arenaQwen-Image-2.1 by @Alibaba_Qwen just landed as the #1 open source model in the Image Edit Arena and Text-to-Image Arena! With 1367 pts in the Image Edit Arena, Qwen-Image-2.1 took the #1 spot among open. It landed #16 overall, just 3 pts from GPT-Image-1.5-high-fidelity at #15. See the leaderboard for the Text-to-Image arena below. Congrats to the @Alibaba_Qwen team on this contribution to the open source ecosystem!
Ant Ling@AntLingAGIAI 评分4949
Unsloth AI@UnslothAIAI 评分6363引用Qwen@Alibaba_QwenMeet Qwen-Image-2.1, the most balanced and cost-effective image generation model in the Qwen-Image series! Now open weights! 🎨 A unified model for both generation and editing, delivering top-tier quality in a lightweight package. Highlights: 👀 - Compact & exceptionally fast: A lightweight 7B architecture that outperforms most closed-source models, with drastically accelerated inference for multi-image inputs. - Native transparency: Natively generates and edits RGBA layers, enabling seamless compositing and text editing within transparent images. - Versatile, high-fidelity editing: Supports up to 10 reference images and precise local control while preserving strict fidelity for portraits and products. - Broad coverage & stunning aesthetics: Excels at panoramas, infographics, and virtual try-ons, delivering realistic textures and elegant typography. Start to create your next masterpiece with Qwen-Image-2.1! 🖼️ - Blog: https://qwen.ai/blog?id=qwen-image-2.1 - GitHub: https://github.com/QwenLM/Qwen-Image-2.1 - Model Scope: https://www.modelscope.cn/models/Qwen/Qwen-Image-2.1 - Hugging Face: https://huggingface.co/Qwen/Qwen-Image-2.1
Tencent Hy@TencentHunyuanAI 评分4545🎨 Hy Image3.5 preview 已在 Miora 上线。编辑时保留已经好用的部分——同一画布,你的品牌规则已被记住。免费两周。
引用Miora Design@Miora_DesignHy Image3.5 preview is live on Miora. FREE through October 7. Campaign visuals. Product scenes. Storyboards. Game concepts. Up to 2K. Text-to-image and image-to-image. Try it: https://miora.design/
NVIDIA Technical Blog(开发者技术博客 · RSS)AI 评分4747 NVIDIA DLSS 5 发布:3D 引导神经渲染、ACE 更新与 RTX Kit 新能力
NVIDIA 发布 DLSS 5,引入 DLSS 3D-Guided Neural Rendering 和细粒度控制,帮助开发者在保留艺术意图的同时加入逼真光照与材质细节。同步更新 NVIDIA ACE、RTX Mega Geometry 2.0 和 RTX Kit,覆盖角色 AI、高密度几何与渲染工作流。
Qwen@Alibaba_QwenAI 评分3535来自 @inteldevs 的 Day-0 OpenVINO 支持!🥳 Qwen-Image-2.1 已可在 Intel 硬件上优化运行。一个开放权重 checkpoint,同时支持生成与编辑。👇
引用Intel Devs@inteldevsWe're excited to offer Day0 OpenVINO support for Qwen-Image-2.1 Read more about what you can accomplish here: https://ms.spr.ly/6019a5nm5
Xiaomi MiMo@XiaomiMiMoAI 评分3535
Qwen@Alibaba_QwenAI 评分4242感谢 @sgl_project 的 day-0 支持!🙌 SGLang-Diffusion 现已支持 Qwen-Image-2.1:文生图、多图编辑,以及透明 RGBA 输出。快来试试!🎨
引用SGLang@sgl_projectDay-0 support for @Alibaba_Qwen’s Qwen-Image 2.1 is here in SGLang-Diffusion! 🖥️ Native precision on a single RTX 4090 24GB with CPU offload - 1024×1024 generation in 18.7s and image editing in 21.7s with 22.7 GiB peak GPU memory during requests. - On an RTX PRO 6000 96GB: 8.0s generation and 9.6s editing. 🎨 Text-to-image, multi-image editing, and transparent RGBA output—all with one checkpoint. ⚡ Native inference with TP/SP, LoRA, and OpenAI-compatible APIs. 40 denoising steps, one image per request, warmed HTTP latency including PNG output. No quantization. Cookbook and GPU-specific commands below 👇
Qwen@Alibaba_Qwen精选AI 评分6666引用Hugging Apps@HuggingAppsQwen Image 2.1 is here! 🖼️ A 7B params native image generation and editing model, with up to 10 image references The model comes with it's own prompt enhancement LLMs, integrated with diffusers 🧨 and ComfyUI ▶️ on Spaces https://huggingface.co/spaces/hugging-apps/qwen-image-2-1
推荐理由:原文给出了模型的参数量、参考图能力与免配置体验入口,读者可以直接在浏览器试用判断适用性。
Qwen@Alibaba_Qwen精选AI 评分6565引用ComfyUI@ComfyUIQwen-Image-2.1 is now supported in ComfyUI! Open weights. One 7B checkpoint that generates and edits. → Image generation at native 2K → Instruction editing from up to 10 reference images in a single pass → RGBA output, alpha included
推荐理由:正文点出该模型已在 ComfyUI 支持,读者可以据此更新本地图像生成与编辑工作流。
Grok@grokAI 评分5555引用Grok Imagine@imagineHollywood would spend 7 figures on this Odyssey scene. @PJaccetturo and his studio made this pilot in 9 days, with 3,627 images, 3,043 videos, and generations costing $2,677. Here’s their breakdown of how they did it 🧵
Google AI@GoogleAIAI 评分3939
Mustafa Suleyman@mustafasuleymanAI 评分4949引用Artificial Analysis@ArtificialAnlysGenerating high-quality images is cheaper and faster than ever. Muse Image, MAI-Image-2.6 and GPT Images 2.5 have substantially shifted the Text to Image Pareto frontiers for both price and speed in recent weeks.
Krea@krea_aiAI 评分3232
SenseTime@SenseTime_AIAI 评分5353商汤发布 SenseNova U1.5 技术报告,这是一个开源的 8B-MoT 原生统一模型,通过共享注意力连接理解与生成。
蚂蚁 inclusionAI:HuggingFace 新模型AI 评分4949 蚂蚁 inclusionAI 发布 Ming-Image-0.1-Design-Layer 设计图层分解模型
蚂蚁 inclusionAI 在 Hugging Face 发布 Ming-Image-0.1-Design-Layer,可将扁平化设计图按指定层数拆解为 RGBA 图层,采用 MIT 许可。
蚂蚁 inclusionAI:HuggingFace 新模型AI 评分5151 蚂蚁 inclusionAI 发布 Ming-Image-0.1-Design 文生图模型
蚂蚁 inclusionAI 在 Hugging Face 发布 Ming-Image-0.1-Design,是一个面向 UI、信息图、海报等文字密集视觉设计的 6B 文生图模型,支持完整视觉构图和带透明背景的 RGBA 输出。
Midjourney:Updates(RSS)AI 评分3131 Midjourney Alpha 更新:新增韩语支持并修复 v8.2 编辑器
Midjourney 为 alpha.midjourney.com 推送更新,新增韩语支持,并集中修复 v8.2 编辑器的多项问题。编辑器提示栏在附上图片后会询问“What would you like to change?
Kling AI@Kling_aiAI 评分1111
Luma@LumaLabsAIAI 评分2727有没有拍完照后被要求看背面?我们也是。 Camera Angles 把一张照片变成一整组拍摄。选好你的角度,就能从每个角度拿回同一个主体,细节完好。 为你希望当初拍到的那些镜头。

MiniMax Design (H3)@Hailuo_AIAI 评分2222引用楽園|AI動画&AI音楽@dave392750うきょさんのプロンプトでMinimaxH3Designで作成しました。 これマジでヤバいですね。 @Hailuo_AI #MiniMaxH3
通义 QwenAudio:原创语音项目AI 评分4848 通义 Qwen-Image-2.1 发布:Qwen 最强开源图像生成模型
通义(Qwen)发布开源图像生成模型 Qwen-Image-2.1,官方称其为 Qwen 目前最强的开源图像生成模型。该模型已在 QwenLM 项目下开源,具体参数规模与评测分数尚未在原文中披露。
Luma@LumaLabsAIAI 评分1919
Midjourney@midjourneyAI 评分1414
Google AI@GoogleAI精选AI 评分6565
推荐理由:官方介绍 Pics 的对象级编辑、图内文字修改和协作能力,也说明可用范围与 Workspace 集成,方便判断是否纳入工作流。
Hugging Face:Blog(RSS)精选AI 评分6161 Hugging Face 用 Gradio Workflow 重建 AUTOMATIC1111,推出 Workflow1111
Hugging Face 发布 Workflow1111,用 gr.Workflow 在单个画布上以 73 个节点重建了 AUTOMATIC1111 的 11 条媒体管线,涵盖文生图、hi-resolution fix、图生图、VLM 反推提示词、检测生成 inpaint 蒙版、ControlNet 风格 annotator、背景去除、PNG Info 和图生视频。
推荐理由:官方用 Gradio Workflow 在单个画布上复刻了 AUTOMATIC1111 的主要功能,读者可以对照它了解节点式工作流与 ComfyUI 的差异。
ViggleAI@ViggleAIAI 评分5050
OpenAI:官网动态(RSS · 排除企业/客户案例)AI 评分5454 OpenAI 发布 ChatGPT Images 2.5
OpenAI 发布 ChatGPT Images 2.5,可将用户的想法、草图和参考照片转化为更个性化、更精致且更贴近创意的图像。材料为官方摘要,未提供更多细节。
Midjourney:Updates(RSS)AI 评分5151 Midjourney alpha 更新:v8.2 编辑模型上线,lightbox 内置编辑器
Midjourney 向 alpha.midjourney.com 推送大规模更新,最大变化是新的 v8.2 编辑模型上线,编辑器直接内置于 lightbox,可用自然语言指令编辑、附加最多 4 张参考图并集中查看本次会话的编辑。
上海人工智能实验室 InternLM:原创项目AI 评分4646 InternLM 发布 InternLumina-U2 多码本扩散大语言模型
上海人工智能实验室 InternLM 团队发布 InternLumina-U2,一个面向全视觉理解、图像生成与编辑的多码本扩散大语言模型。该模型将扩散生成能力与语言建模统一在同一框架内,可同时处理视觉理解与图像生成、编辑任务。
Midjourney:Updates(RSS)AI 评分2828 Midjourney 更新 V8.2 图像编辑模型,提升画质
Midjourney 更新了 V8.2 编辑模型,图像质量得到提升。官方建议过去 24 小时内遇到问题的用户重新尝试,并继续反馈意见。更多更新即将推出。
Midjourney@midjourneyAI 评分3232我们更新了 V8.2 编辑模型,图像质量更好了。如果你在过去 24 小时内遇到任何问题,请再试一次,并继续给我们反馈。谢谢!更多更新即将到来。
Midjourney@midjourneyAI 评分4545Midjourney:Updates(RSS)精选AI 评分6565 Midjourney 开放测试首个 V8.2 图像编辑模型
Midjourney 开始让所有用户测试首个 V8.2 图像编辑模型。该模型支持用指令编辑图像、以最多 4 张图像参考生成新图(替代 omni-reference)、局部重绘(inpainting)与扩图(outpainting),并可用 personalization、moodboards 和 srefs。
推荐理由:官方宣布 V8.2 图像编辑模型开放测试,列出指令编辑、多图参考、局部重绘等能力变化和具体入口。
Midjourney:Updates(RSS)AI 评分2929 Midjourney alpha 站点更新:文件夹分组、Upscale/Zoom/Vary 回归与大量修复
Midjourney 为 alpha.midjourney.com 上线两周的大改版推送了一批 UX 改进与 bug 修复,侧边栏新增可折叠的文件夹树分组,支持双击重命名分组并把新文件夹直接归入其中。
MIT News(RSS)AI 评分6464 MIT CSAIL 研究:生成图像的归因随训练数据规模增大而衰减
MIT CSAIL 团队在 Nature Communications 发表论文,提出“归因衰减”现象:训练数据越大,单个样本对生成结果的影响越小,删除任一图像甚至某艺术家的全部图像后输出不变。