image generation
18 articles tagged with image generation
SenseNova Releases U1.5-8B-MoT, an Open-Weight Unified Model for Image Generation and Editing
SenseNova has released SenseNova-U1.5-8B-MoT, an open-weight native multimodal model built on its NEO-unify architecture for image generation, editing, and native 4K output. The model is available on Hugging Face under an Apache 2.0 license, with no inference pricing yet since it must be self-hosted.
OpenAI Adds Transparent Background Generation to GPT-Image-2 API
OpenAI is previewing a transparent background feature for GPT-Image-2 through its API, letting developers generate PNGs with no background baked in at generation time. The company claims this produces cleaner results than traditional background removal, particularly on difficult edges like glass or thin fibers.
xAI's Imagine Image 2.0 Scores Second Place Behind GPT-Image-2 in Arena Benchmarks
xAI's new Imagine Image 2.0 lands in second place globally on both the Image Edit and Text-to-Image Arena leaderboards, trailing OpenAI's GPT-Image-2. The model ships with editing tools including Magic Wand, Multi-Ref Editing, and Smart Resize.
Black Forest Labs Unveils FLUX.2 [klein]: A Distilled Model for Interactive Image Generation
Black Forest Labs has released FLUX.2 [klein], a lightweight variant of its FLUX.2 image generation model family designed for faster, more interactive use. The company frames the release as a step toward 'interactive visual intelligence,' though detailed benchmarks and pricing have not yet been disclosed.
Microsoft Releases Mage-Flow, a 4B Open-Weight Model That Matches 20B+ Rivals on Image Generation and Editing
Microsoft has released Mage-Flow, a 4B-parameter open-weight foundation model for text-to-image generation and instruction-based editing. The company claims it matches or beats much larger open systems like Qwen-Image (20B) and FLUX.2 (32B) while running faster and using less memory.
Alibaba Releases Qwen-Image-3.0, an Image Generator That Renders 10-Pixel Text and 3x3 Infographic Grids in One Pass
Alibaba's Qwen team has released Qwen-Image-3.0, an image generator that accepts prompts up to 4,500 tokens and can render legible text as small as ten pixels, complex LaTeX formulas, and twelve languages in a single pass. The model is currently invite-only via API, and unlike its predecessor, it likely won't ship with open weights.
Envato Generates 51 Million Images Using FLUX Models, Accounting for 25% of Platform Volume
Envato has generated over 51 million images using Black Forest Labs' FLUX models since establishing a direct partnership in 2023. FLUX now accounts for approximately 25% of total image generation volume on the creative platform, with FLUX.2 download rates running 16% above platform averages.
Meta launches Muse Image, a free AI image generator integrated across Instagram, WhatsApp, and Facebook Marketplace
Meta has launched Muse Image, a new AI image generator from its Meta Superintelligence Labs division. The model is available free for Instagram Stories, WhatsApp, and the Meta AI app, with integration into Facebook Marketplace for visualizing used furniture in home settings.
Meta launches Muse Image model with Instagram account prompts and QR code generation
Meta has launched Muse Image, the first AI image generation model from Meta Superintelligence Labs, now available in the US through Meta AI app, Instagram, and WhatsApp. The model accepts Instagram accounts as prompts to incorporate users' likenesses and claims to generate functional QR codes with legible styled text.
Meta releases Muse Image generator, powers Instagram Stories and WhatsApp chat features
Meta has released Muse Image, its first image generation model from Meta Superintelligence Labs. The model is now available in the Meta AI app and powers new creative features in Instagram Stories and WhatsApp direct chats, with over 30 AI-powered effects for Instagram Stories launching in limited countries.
Google releases Nano Banana 2 Lite: 4-second image generation at $0.034 per 1,000 images
Google released Nano Banana 2 Lite, an AI image generator that produces images in 4 seconds and costs $0.034 per 1,000 images. The model is optimized for high-volume workflows and replaces the original Nano Banana as Google's entry-level image generation offering.
Google releases Nano Banana Pro image generation model with 2K/4K output and five-subject identity preservation
Google has released Nano Banana Pro, an advanced image generation and editing model built on Gemini 3 Pro. The model supports 2K/4K output resolution, preserves identity across up to five subjects, and includes real-time Search grounding for context-rich visual synthesis.
Ideogram AI releases FP8-quantized image generation model on Hugging Face alongside Google's Gemma-4-12B text models
Three new models appeared on Hugging Face: Ideogram AI's FP8-quantized version of its Ideogram-4 image generation model and Google's Gemma-4-12B text models in both base and instruction-tuned variants. The releases mark continued expansion of model availability through Hugging Face's platform.
OpenAI releases ChatGPT Images 2.0 with accurate text rendering and brand-style matching
OpenAI launched ChatGPT Images 2.0, upgrading from decorative images to full-page graphics with detailed text rendering. The update is available to all ChatGPT tiers, with advanced features requiring paid subscriptions that access the Thinking model. Hands-on testing shows significant improvements in text accuracy and brand-style replication, though factual errors still occur.
OpenAI Releases GPT-5.4 Image 2 with 272K Context Window and Image Generation
OpenAI has released GPT-5.4 Image 2, combining the GPT-5.4 reasoning model with image generation capabilities. The multimodal model features a 272K token context window and is priced at $8 per million input tokens and $15 per million output tokens.
OpenAI releases ChatGPT Images 2.0 with integrated reasoning and text-image composition
OpenAI has released ChatGPT Images 2.0, which integrates reasoning capabilities to generate complex visual compositions combining text and images. The model supports aspect ratios from 3:1 to 1:3 and outputs up to 2K resolution, with advanced features available to Plus, Pro, Business, and Enterprise users.
Google connects Gemini chatbot to personal Google Photos for AI-generated images
Google announced Thursday that users can connect their Google Photos library to the Gemini chatbot for personalized image generation through its Nano Banana feature. Users must opt in to Personal Intelligence, and the feature will roll out to paid subscribers in the coming days.
Canva launches agentic AI assistant that automatically calls design tools from text prompts
Canva has released Canva AI 2.0, an agentic assistant that automatically calls design tools based on text prompts and creates editable layered designs. The update includes integrations with Slack, Gmail, Google Drive, Calendar, and Zoom for context building, plus web research and task scheduling capabilities.