Skip to contentSkip to stories

Updated

#Image generation

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri20 items
  1. MagnificAI score22

    Magnific shares MOON fashion brand image prompts for campaign shots

    AIMagnific (@magnific) posts three example image prompts for a fictional fashion brand called MOON, covering an editorial cover, a candid cafe photo and a product breakfast scene. The prompts specify a shared ivory, powder blue and burgundy palette, preserved reference garments and jewelry, and fully clothed subjects. The accompanying post says the workflow uses Brand Kit and Magnific One to build a consistent campaign from first image to video.

  2. IdeogramAI score45

    Ideogram 4.5 keeps edited images intact across 30 consecutive edits

    AIArtificial Analysis ran 30 consecutive real estate staging edits through four image editing models, and Ideogram 4.5 kept most of each image unchanged while others drifted. Ideogram 4.5 and FLUX 3 left 95% or more of the image untouched on small edits, while GPT Image 2.5 Sunburst re-rendered most of the image and left only about a fifth unchanged. Nano Banana 2.1 kept its edits local but gradually darkened the rest of the image.

  3. Artificial AnalysisAI score50

    Ideogram 4.5 keeps edited photos intact over 30 consecutive edits

    AIArtificial Analysis ran four image editing models through 30 consecutive edits of the same real estate photo, with each model editing its previous output. Ideogram 4.5 and FLUX 3 left 95% or more of the image essentially untouched on small edits, while GPT Image 2.5 Sunburst re-rendered most of the image each time, leaving only about a fifth unchanged. Nano Banana 2.1 edited locally but shifted and gradually darkened the rest of the image.

    Video from @ArtificialAnlys's post
  4. Qwen · new models on Hugging FaceAI score49

    Qwen releases Qwen-Image-2.1-Turbo, an 8-step accelerated image generation checkpoint

    AIQwen has published Qwen-Image-2.1-Turbo on Hugging Face, an accelerated checkpoint of Qwen-Image-2.1 for text-to-image generation and image editing with 8 denoising steps. The checkpoint uses the same 7B visual generation architecture, loads directly with QwenImage21Pipeline in Diffusers, and includes its recommended sampling schedule. It defaults to CFG=1 and uses prefix KV caching to reuse text and reference-image context across steps.

  5. ModelScopeAI score60

    Qwen-Image-2.1-Turbo cuts image generation and editing to 8 denoising steps

    AIModelScope announces Qwen-Image-2.1-Turbo, an accelerated checkpoint that keeps the 7B visual architecture and runs image generation and editing in 8 denoising steps. The source says it uses CFG=1 and prefix KV caching to reuse text and reference-image context across steps, supports 2048 resolution with square, portrait, landscape, and widescreen presets, and loads through QwenImage21Pipeline in Diffusers. It is released under the Qwen Research License Agreement.

    Why it matters: The source names a concrete speedup path, 8 sampling steps and CFG=1 with prefix KV caching, which matters to anyone weighing image generation latency.

    Image from @ModelScope2022's post
  6. LeiphoneAI score42

    Doubao Work adds Canvas feature and Doubao 2.1 Lite model

    AIDoubao Work has added a Canvas feature for complex creative tasks, placing materials, design plans and outputs on one infinite canvas where users can keep editing text, colors and layout after images are generated. The update also integrates the lightweight Doubao 2.1 Lite model, aimed at everyday Q&A, document writing, spreadsheets and PPT creation, with optimized response speed and usage consumption.

Oct 8

Oct 8Thu
  1. Anil Chandra Naidu MatchaAI score29

    Open-Voyager releases open-source creative-work agent harness

    AIOpen-Voyager is announced as a free, open-source harness for creative work, described as a Codex or Claude Code equivalent. The post says it integrates with 600+ models and links to a GitHub repository for self-hosting. The background post says the original Voyager is built for video, graphics, and games, and can drive tools such as Blender, Resolve, and After Effects.

    Video from @matchaman11's post
  2. Artificial AnalysisAI score38

    Grok Imagine Video 1.5 Lite leads in architecture, consumer, and knowledge-work use cases

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier in Architecture & Real Estate, Consumer, and Productivity & Knowledge Work use cases. It sits furthest from the frontier in Live-Action Film and Frontier use cases. Against Grok Imagine Video 1.5, Lite matches it in Social Media & Creator Content and trails it on the other nine use cases.

    Image from @ArtificialAnlys's post
  3. Artificial AnalysisAI score31

    Grok Imagine Video 1.5 Lite leads on quality and speed benchmark

    AIAmong 12 models on AA-Video-T2V-Silent v2.0, Grok Imagine Video 1.5 Lite is the only one that is both fastest and highest quality, with no model beating it on both measures. It generates a 10-second 1080p clip in a median of 60.5 seconds. Kling 3.0 1080p (Pro) scores slightly higher but takes 94 seconds for a 5-second clip, while Vidu Q3 Turbo is 9 seconds faster on a 5-second 720p clip yet scores well below it.

    Image from @ArtificialAnlys's post
  4. Midjourney UpdatesAI score46

    Midjourney Tests Thinking Mode for Image Generation on Alpha Site

    AIMidjourney is testing a "Thinking Mode" on its Alpha website, where users can click "Rerun (Thinking)" in a job's lightbox to regenerate an image. The company says early tests show gains in prompt accuracy, typography, and coherence, and it is asking users to share feedback in its #ideas-and-features channel. It may later offer the mode broadly or as an option to add more thinking after a job.

  5. Midjourney UpdatesAI score25

    Midjourney Adds Shared Folders and Thinking Mode in Alpha Update

    AIMidjourney's alpha site now lets users share folders with others as Collaborators or Viewers, with sharing by link also available. A new thinking mode lets users rerun jobs made with 8.2 standard and edit models to fix missed prompt details such as objects, layout, anatomy, and text. Sharing does not change image privacy, so non-stealth images can still appear on Explore and profiles.

  6. elvisAI score42

    Voyager: an open harness for creative AI work across video and games

    AIElvis Saravia argues that creative work needs domain-specific agent harnesses rather than coding-oriented ones, and he highlights Voyager as an open harness for video, graphics, and games. According to the quoted post, Voyager lets agents work with local files and drive apps such as Blender, DaVinci Resolve, and Unity, and it is designed to work with models like Opus, Astra, and DeepSeek.

    Video from @omarsar0's post
  7. MagnificAI score4

    Magnific shares a surreal giant-egg image prompt with a hand emerging

    AIMagnific posts a detailed text prompt for a surreal landscape image featuring a 3-meter ivory-white egg cracking open in a shallow pink-water pool, with a single human hand reaching out between two limestone walls. The prompt specifies frontal symmetrical composition, dramatic natural sunlight, and a medium-format analog film look with organic 35mm grain.

    Image from @magnific's post
  8. MagnificAI score5

    Magnific shares a frontal theatrical image prompt with a giant black prop hand

    AIMagnific (@magnific) posts a detailed image-generation prompt for a strictly frontal, cinematic scene of 7–9 women in white dresses dancing atop a pale rocky stage. The prompt specifies a single spotlight on one central dancer, a muted blue sunset sky, and a giant matte-black practical-prop hand rising from behind the rock.

    Image from @magnific's post