Skip to contentSkip to stories

Updated

#Model release

Showing low-relevance items too. Hide low-relevance items

Sep 23

Sep 23Wed
  1. QwenOfficialAI score62

    Qwen-Audio-3.1 upgrades ASR, TTS and Realtime and adds two new models

    AIAlibaba's Qwen team released Qwen-Audio-3.1, upgrading its ASR, TTS and Realtime models and adding TTS-Next and ASR-Next. The post says TTS prices fell about 70%, Realtime about 85%, and ASR up to 95%. More APIs are coming soon.

    Why it matters: The post lists the new audio models and price changes together, which helps developers compare the upgraded lineup with their current speech workflows.

    Image from @Alibaba_Qwen's post
  2. ModelScopeOfficialAI score62

    Xiaomi MiMo-V2.6 open-sourced as a multimodal agent model family under MIT License

    AIXiaomi has released MiMo-V2.6 as an open model family under the MIT License, designed for large-scale reinforcement learning. MiMo-V2.6-Pro scores 46 on the Artificial Analysis Intelligence Index, with 71.9 on DeepSWE v1.1, 89.9 on Terminal-Bench 2.1, and 82.0 on OSWorld-Verified. The 1.02T-parameter MoE activates 42B parameters and supports text, image, video, and audio input with a 1M-token context.

    Why it matters: The post links benchmark results, parameter scale, and a multi-agent RL training run, giving readers concrete figures to compare against other open models.

    Image from @ModelScope2022's post
  3. KrASIA · Big TechNewsAI score46

    Tencent Hy Image 3.5 preview refined through its consumer and business products

    AITencent has released a preview of its Hy Image 3.5 image generation model, which product teams across Yuanbao, WorkRally, Ima, and other services are helping refine through co-design. Tencent Cloud prices the model at USD 0.024 per 2K output image, and it supports text-to-image and image-to-image generation with up to five reference images. Tencent said an internal blind evaluation found it on par with ByteDance's Seedream 5.0 Pro and slightly better than Nano-Banana Pro and Qwen-Image-3.0 Pro.

  4. swyxXAI score29

    Opus 5.5 becomes new default for AINews, writing more concise

    AIswyx reports that Claude Opus 5.5 is now the default model for Latent Space's AINews, after side-by-side testing against Sol showed more concise and tasteful reporting with less "slopese" than Opus 5. The post links to the AINews issue titled "Claude Opus 5.5: The New Default."

    Image from @swyx's post
  5. ModelScopeOfficialAI score62

    Shanghai AI Lab and SJTU release open-weight 8.9B NCP-ArchPreview model under Apache 2.0

    AIShanghai AI Lab and SJTU's LUMIA Lab released NCP-ArchPreview, an 8.9B open-weight language model under Apache 2.0. The model reportedly reaches OLMo-3-7B's final Stage 1 loss using 51.3% of the tokens from the 5.73T Dolma 3 corpus, a 1.95× convergence gain. Its concept module jointly predicts tokens and concepts, and domain adaptation updates only its 17M parameters while the token backbone stays frozen.

    Why it matters: The post pairs an Apache 2.0 open-weight release with training-efficiency figures, showing how the concept module adapts to new domains with few trainable parameters.

    Image from @ModelScope2022's post

Sep 22

Sep 22Tue
  1. ModelScopeOfficialAI score62

    inclusionAI open-sources Ming-Image-0.1-Design models for visual design

    AIinclusionAI open-sources the Ming-Image-0.1-Design family, two complementary 6B models for visual-design workflows, under an MIT License. Design generates complete UIs, dashboards, infographics, and posters up to 2048×2048 with native transparent RGBA output, and Layer decomposes flattened graphics into independently editable RGBA layers.

    Why it matters: The post separates a design-generation model from a layer-decomposition model, letting users compare two distinct visual-design workflows under one MIT license.

    Image from @ModelScope2022's post
  2. Fireworks AI BlogOfficialAI score65

    Fireworks releases Ember-1, a Kimi K3 variant that cuts reasoning tokens by about 40%

    AIFireworks Research released Ember-1, a specialized model built on Kimi K3 that it says delivers the same quality with 40% fewer tokens. Across five industry benchmarks, Ember-1 matched K3 max quality at a fraction of the cost, and in two customer A/B tests it used about 35% fewer tokens per task. It is available as a Research Preview on Serverless, and Fireworks is also launching training support for customized models.

    Why it matters: The source gives benchmark and A/B results for cutting reasoning tokens while holding quality, which bears on cost planning for coding and agent workloads.

  3. Tibor BlahoXAI score88

    OpenAI launches GPT-6 Sol and Luna while Anthropic releases Claude Opus 5.5

    AIOpenAI released GPT-6 Sol and Luna, with API prices cut in half, while Anthropic released Claude Opus 5.5 at roughly Fable 5.1 level for 40% less than Opus 5. GPT-6 Sol and Luna cost $2/$10 and $0.10/$0.50 per million tokens, versus GPT-5.6 promotional pricing, and Opus 5.5 costs $4/$20 per million tokens. Sonnet 5.5 and Haiku 5.5 are announced for the coming weeks.

    Why it matters: The post links OpenAI's GPT-6 Sol and Luna pricing with Anthropic's Claude Opus 5.5 launch, which helps readers compare the two vendors' current frontier offerings.

    Image from @btibor91's post
  4. Tri DaoXAI score44

    Rigel: 2.3B hybrid Mamba-2 MoE nears Llama-3.2-3B with <1% FLOPs

    AIMayank's Rigel, a 2.3B-parameter MoE (360M active) hybrid Mamba-2 model, was pretrained across H100, A100, V100 GPUs and TPU v5p/v6e on one codebase. The model lands within a few points of Llama-3.2-3B while using under 1% of its pretraining FLOPs. Tri Dao praised the work's engineering effort and the model's strength for its small size.

  5. Simon WillisonXAI score60

    OpenAI launches GPT-6 Sol and Luna with 50% lower API prices

    AIOpenAI introduced GPT-6 Sol and GPT-6 Luna, which build on advances behind GPT-6 Astra. The company says it cut API prices 50% for Sol and Luna compared with GPT-5.6 promotional pricing, passing on caching and inference efficiency gains. Simon Willison notes GPT-6 Luna costs half of GPT-5.6 Luna and calls Luna his favorite model for building product features because of its cost and speed.

  6. Kilo (acq. by Anaconda)OfficialAI score17

    Kilo compares Grok 4.7 and Opus 5.5 on a grass-touching simulator

    AIKilo Code tested Grok 4.7 against Opus 5.5 on a touching-grass simulator, with Grok's one-shot result costing 47% less. Per the quoted post, Grok 4.7 cost $3.52 versus $7.35 for Opus 5.5, and both models are available in Kilo now.

  7. Greg BrockmanXAI score81

    OpenAI launches GPT-6 Sol and Luna with 50% lower API prices than GPT-5.6

    AIOpenAI introduced GPT-6 Sol and GPT-6 Luna, which it says bring much of the strength of GPT-6 Astra into faster and more affordable models. The company also reports more efficient caching and inference, with API prices 50% lower than GPT-5.6 promotional pricing.

    Why it matters: The quoted announcement names specific pricing and access changes for Sol and Luna, which matter for teams weighing cost against the Astra tier.

  8. Alex AlbertXAI score18

    Opus 5.5 builds a 1906 San Francisco street scene in Blender

    AIAlex Albert says Opus 5.5 has improved 3D modeling and vision for Blender work, letting users build an entire world from a single prompt. He shares a historically accurate render of San Francisco's Market Street in 1906, before the earthquake.

    Video from @alexalbert__'s post
  9. Noam BrownXAI score78

    OpenAI releases GPT-6 Sol and Luna at 50% lower API prices

    AIOpenAI has released GPT-6 Sol and GPT-6 Luna, which it says build on GPT-6 Astra and offer faster, more affordable performance. API prices for Sol and Luna are 50% lower than GPT-5.6 promotional pricing, and Luna now costs $0.10 input and $0.50 output per 1M tokens. The author also notes an earlier 80% Luna price cut at the end of July, with output dropping from $6 to $0.50 within two months.

    Why it matters: The source gives concrete API price cuts across two model tiers, making the cost trend across recent releases easy to track for developers.

  10. Sherwin WuXAI score46

    GPT-6 Luna launches at $0.10 and $0.50 per million tokens

    AIOpenAI's GPT-6 Luna is priced at $0.10 per 1M input tokens and $0.50 per 1M output tokens, according to Sherwin Wu. Per OpenAI Devs, Luna and GPT-6 Sol launch today with API prices 50% lower than GPT-5.6. Wu jokes that per-billion-token pricing may soon be needed.

  11. ChatGPTOfficialAI score62

    OpenAI rolls out GPT-6 Sol and GPT-6 Luna in ChatGPT Work and Codex

    AIOpenAI announced GPT-6 Sol and GPT-6 Luna, rolling out today in ChatGPT Work and Codex. The rollout covers Plus, Pro, Business, Enterprise, and Edu users.

    Why it matters: The post names two new GPT-6 variants and their rollout to specific ChatGPT and Codex plan tiers, which shows how access is being staged.

    Video from @ChatGPT's post
  12. ChatGPTOfficialAI score72

    GPT-6 Luna Rolls Out to Free and Go Users in the ChatGPT Desktop App

    AIOpenAI's official ChatGPT account says Free and Go users can try GPT-6 Luna in the desktop app, with rollout starting today. The post links to OpenAI's announcement introducing GPT-6 Sol and Luna.

    Why it matters: The post names the access tier and platform for the new model, which is the detail readers need to judge whether it applies to them.

  13. Felix RiesebergXAI score47

    Anthropic's Opus 5.5 praised for natural writing and computer art

    AIAnthropic's Felix Rieseberg says the new Opus 5.5 model writes more naturally than earlier models. He also highlights its strong generative computer art and drawing ability, noting it is not an image model yet produces attractive visuals.

  14. Mike KriegerXAI score67

    Anthropic launches Claude Opus 5.5, leading in coding and knowledge work

    AIAnthropic has launched Claude Opus 5.5, the first model in its new Claude 5.5 family. According to the quoted launch post, it performs at the level of Claude Fable 5.1 for most tasks and costs 40% less to run than Opus 5. The author says it leads in coding and knowledge work and praises its writing quality.

    Why it matters: The quoted launch post gives a concrete cost comparison, useful for weighing Opus 5.5 against earlier Opus and Fable 5.1 models for routine work.

  15. Lydia Hallie ✨XAI score23

    Claude paid plans get a usage limit reset until Oct 22

    AIAnthropic's Pro, Max, and Team subscribers can claim a usage limit reset in Settings → Usage, available until Oct 22. Opus 5.5 is now the default for paid plans and is priced lower than Opus 5, so 5-hour and weekly limits go 25% further.

  16. v0OfficialAI score52

    Claude Opus 5.5 is now available in v0

    AIv0 announced that users can now use Claude Opus 5.5 in v0, with a link to try it. The quoted Anthropic post says Opus 5.5 is the first model in its new Claude 5.5 family, performs at the level of Claude Fable 5.1 on most tasks, and costs 40% less to run than Opus 5.

  17. Boris ChernyXAI score62

    Claude Opus 5.5 ports HAProxy to Rust faster and cheaper than Fable 5.1

    AIAnthropic introduced Claude Opus 5.5 as the first model in its Claude 5.5 family, saying it performs at the level of Claude Fable 5.1 for most tasks at 40% lower run cost than Opus 5. Boris Cherny reports that Opus 5.5 and Fable 5.1 each ported HAProxy from C to Rust and both passed nearly all of its tests, with Opus 5.5 finishing in 9.5 hours versus 12 hours and at 51% less cost.

    Why it matters: The author reports a same-task comparison in which Opus 5.5 finished a HAProxy C-to-Rust port faster and cheaper than Fable 5.1, offering a concrete cost and time benchmark.

  18. Lydia Hallie ✨XAI score62

    Opus 5.5 released, claimed about 30% faster and 40% cheaper per task than Opus 5

    AIOpus 5.5 is now available, and the author says it is about 30% faster and about 40% cheaper per task than Opus 5. The author describes it as feeling like Fable and invites readers to try it.

    Why it matters: The post gives concrete speed and cost comparisons against Opus 5, the most useful part for judging whether the model fits a workload.

    Video from @lydiahallie's post
  19. Alex AlbertXAI score62

    Anthropic introduces Claude Opus 5.5 as first model in Claude 5.5 family

    AIAnthropic has introduced Claude Opus 5.5, the first model in its new Claude 5.5 family. The quoted announcement says it performs at the level of Claude Fable 5.1 for most tasks and costs 40% less to run than Opus 5. Alex Albert's post praises the model as smart, clear, fast, and cheaper, but offers no independent test results.

    Why it matters: The quoted announcement gives a concrete cost comparison against Opus 5, which helps readers weigh the model's value beyond the author's praise.

  20. catXAI score62

    Claude Opus 5.5 becomes the default model in Claude Code and Claude app

    AIClaude Opus 5.5 is now the default model in Claude Code and the Claude app, including Cowork, for Pro, Max, and Team plans. Anthropic is defaulting to effort medium across products, which it says is comparable to Fable 5.1 on intelligence but faster. Rate limits will go 25% further on Opus 5.5 compared to Opus 5.

    Why it matters: The post names concrete default changes across Claude Code and the Claude app, plus a specific effort setting and rate-limit difference, useful for judging day-to-day cost and speed.

  21. Sam BowmanXAI score75

    Anthropic's Sam Bowman says Claude Opus 5.5 is safer, reducing misalignment risk

    AISam Bowman says Claude Opus 5.5 is sufficiently safer than its predecessors that releasing it more likely than not reduces misalignment risks. The quoted @claudeai post introduces Claude Opus 5.5 as the first model in the Claude 5.5 family, performing at the level of Claude Fable 5.1 on most tasks at 40% lower run cost than Opus 5.

    Why it matters: The post links a safety judgment to a model release, which is useful for readers weighing how Anthropic frames release decisions against misalignment risk.

  22. AnthropicOfficialAI score71

    Anthropic releases Claude Opus 5.5, the first model in its Claude 5.5 family

    AIAnthropic has made Claude Opus 5.5 available today, introducing it as the first model in its new Claude 5.5 family. According to the quoted @claudeai post, it performs at the level of Claude Fable 5.1 on most tasks and costs 40% less to run than Opus 5.

    Why it matters: The post gives a concrete cost comparison against Opus 5 and names the model family, helping readers gauge the trade-off between price and performance.

  23. Lovable BlogOfficialAI score38

    Lovable Adds Opus 5.5, Cutting Build Steps by a Third to Half at Same Quality

    AILovable now offers Opus 5.5, which it says matches Opus 5's results while finishing builds in a third to half fewer steps. Internal benchmarks showed Opus 5.5 scoring 4 to 6% ahead of Opus 5 on verification discipline, with step reductions of 26% to 57% and input token reductions of 21% to 59% across tasks.

  24. Black Forest Labs · new models on Hugging FaceOfficialAI score62

    Black Forest Labs releases FLUX 3 Action, a 7B open-weights robot world action model

    AIBlack Forest Labs released FLUX 3 Action, an open-weights 7B world action model that outputs robot joint commands from camera frames, robot state, and a text instruction. On the RoboLab-120 benchmark it reports 42.92% task success, ahead of Cosmos3-Nano-Policy at 36.8% and π0.5 at 28.0%. The model is fine-tuned on DROID, is distributed under the FLUX Kommunity License v.1.0, and runs in about 32 GB of GPU memory in bfloat16.

    Why it matters: The model card gives a benchmark comparison, parameter counts, and an action contract, so readers can judge how it compares with existing robot policies.

  25. Black Forest Labs · new models on Hugging FaceOfficialAI score58

    Black Forest Labs releases open-weights FLUX 3 Action SO-101 robot policy

    AIBlack Forest Labs has published FLUX 3 Action SO-101 on Hugging Face as an open-weights 7B world action model. It takes two camera frames, the robot state, and a text instruction, then returns the next 42 actions with predicted video frames, with 32 executed at 30 Hz before replanning. The card also provides a rank-32 LoRA fine-tuning recipe for user datasets and states that the application must enforce joint velocity, force, and workspace limits.

  26. Black Forest Labs · new models on Hugging FaceOfficialAI score60

    Black Forest Labs releases FLUX 3 Action base weights for robot adaptation

    AIBlack Forest Labs has released flux-3-action-base, an open-weights 7B world action model that takes camera frames, robot state, and a text instruction to output the next action chunk. The release is an adaptation component rather than a complete robot policy, and new embodiments require their own action heads. The source says the weights are paired with shared video VAE and Qwen3-VL-4B-Instruct text encoders and is governed by the FLUX Kommunity License v.1.0.

    Why it matters: The source separates the adaptation base from full robot policies and states the shared encoders and new-embodiment requirements, which clarifies what developers must still build for their robots.

  27. AI SupremacyBlogAI score45

    TypeSafe AI's Jev Is a Non-LLM Probabilistic Classifier for Fast Software Decisions

    AITypeSafe AI released Jev, a transformer-based System-1 model that outputs calibrated probabilistic decisions instead of generating tokens, returning answers in 70–500 ms at $0.042 per million input tokens. The model is built for typed Choice, Score, and yes/no questions inside software pipelines, and it is available to everyone without a waitlist, with $5 in starting credits. Vercel, Cloudflare, LangChain, and Langfuse have added Jev to their platforms.