Skip to contentSkip to stories

Updated

#Model release

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 30

Sep 30Wed
  1. Ant LingOfficialAI score38

    Ling-3.1-flash ports C image library to Rust with 8.015× speedup

    AIAnt Ling reports that its Ling-3.1-flash model completed a roughly 20-hour Rust port of a C image library. After a performance regression caused by busy-waiting workers and a parallelism adjustment, the model recovered and reached an 8.015× speedup. All 30 correctness checks passed.

    Image from @AntLingAGI's post
  2. Ant LingOfficialAI score46

    Ant Ling releases Ling-3.1-flash with 1M-token context, plans open-source

    AIAnt Ling introduced Ling-3.1-flash, a model with about 560B total parameters, about 25B active per token, and up to a 1M-token context window. The company plans to open-source the model soon. It reports 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional across work, coding, and healthcare tasks.

    Image from @AntLingAGI's post
  3. IdeogramOfficialAI score46

    Ideogram 4.5 launches with four quality modes at native 2K resolution

    AIIdeogram 4.5 comes in four quality modes, ranging from 0.8¢ to 22¢ per image, all at native 2K resolution. Available now on our launch partners: @superscale_ai @Picsart @cfabricacom @luminal_ai @arena @runware @florafaunaai @krea_ai @trymoda @LeonardoAi @runwayml @pika_labs @fal @ComfyUI @magnific @GammaApp @LumaLabsAI @DesignArena

    Image from @ideogram_ai's post
  4. IdeogramOfficialAI score38

    Ideogram 4.5 launches as a precise image edit model

    AIIdeogram released Ideogram 4.5, which it calls the most precise edit model, claiming it avoids the artifacts, pixel shifts, and color changes that leading models add with each edit. The company says this eliminates artifact buildup and makes multi-turn editing possible. It is live in Ideogram, via the API, and with launch partners, with open weights promised soon.

    Video from @ideogram_ai's post
  5. Aidan GomezXAI score38

    Cohere launches Embed 5 Pro and Embed 5 Fast embedding models

    AICohere has introduced Embed 5, a new family of state-of-the-art embedding models, with Embed 5 Pro for frontier capabilities and Embed 5 Fast for low-latency performance. The post says the models are extremely scalable, offer SOTA accuracy, and can be deployed privately and accessed through Model Vault.

  6. ModelScopeOfficialAI score46

    IndexTeam releases Index-Translate multilingual translation model family

    AIIndexTeam has released Index-Translate, a multilingual family covering text, speech, dubbing, and long-document translation across 150 languages. Its 9B model scores 0.8789 on FLORES, 0.8209 on instTrans, and 0.7387 on MEME, and the 2B and 9B models are released under Apache 2.0.

    Video from @ModelScope2022's post
  7. SenseTimeOfficialAI score23

    SenseTime previews Dynamic Design, animating static images with SenseNova 6.8 Flash

    AISenseTime previewed Dynamic Design, powered by SenseNova 6.8 Flash, which turns static images into animated visuals. The system decides which elements stay static and which to animate, chooses HTML/CSS, SVG, transparent images, or video for each element, and choreographs text reveals and subject motion. SenseNova 6.8 Flash is coming soon, and SenseTime is offering a limited beta.

    Video from @SenseTime_AI's post
  8. ModelScopeOfficialAI score62

    InSpatio-World 1.5 turns images and videos into real-time explorable 4D worlds

    AIInSpatio-World 1.5 from InSpatio_AI turns a single image, four images, a panorama, or a video into a navigable scene with wide viewpoint changes. The 1.3B model scores 68.72 on WorldScore-Dynamic, ranking first among evaluated real-time and interactive methods, with speeds up to 24 FPS. The post says the code is released under Apache 2.0 and that dependencies keep their own licenses.

    Why it matters: The post gives specific benchmark, speed, and input details, so readers can judge how the model handles real-time scene exploration from images or video.

    Video from @ModelScope2022's post
  9. ollamaOfficialAI score23

    Ollama adds Nimble, Tev1 4B, and Tev1 0.8B models

    AIOllama now lists three models, Nimble, Tev1 4b, and Tev1 0.8b, for use with /v1/systemone. The post provides a library link for each model, with the Tev1 variants listed at 4b and 0.8b.

  10. ollamaOfficialAI score30

    Ollama adds local decision models like Nimble via new API

    AIOllama now supports decision models such as Nimble locally, usable for tasks like ticket triaging, model routing, and content moderation. The model can be installed with ollama pull nimble and accessed through the new local /v1/systemone API, as shown in a real-time Ollama racer demo.

    Video from @ollama's post
  11. Kling AI BlogOfficialAI score49

    Kling 4.0 Extends Native Video to 30 Seconds With Up to 10 Keyframes

    AIKling 4.0 extends native single-pass video generation from 15 to 30 seconds and adds Multiple Keyframes supporting up to 10 keyframe images, versus Start & End Frames in Kling 3.0. It also expands reference inputs to up to 15 combined assets, including up to 5 videos totaling 30 seconds, and adds 10-bit HDR at 1080p and 4K. The all-new Kling 4.0 will officially launch in October, and Kling 4.0 Flash became available to a limited group of early-access users on September 28.

  12. Artificial Analysis ArticlesOfficialAI score39

    Upstage Releases Solar Mini 4 Reasoning Model, Scoring 24 on Intelligence Index

    AIKorean AI lab Upstage has released Solar Mini 4, a proprietary reasoning model that scores 24 on the Artificial Analysis Intelligence Index with 35B total and 3B active parameters. It is priced at $0.10/$0.40 per 1M input/output tokens and has a 1M-token context window, but averages 7.1 minutes per task due to heavy output token use. Its weights are not released, and its size cannot be independently verified.

  13. Artificial Analysis ArticlesOfficialAI score75

    Gemini 4 Argon matches GPT-6 Astra on intelligence index at lower cost

    AIArtificial Analysis reports that Google's Gemini 4 Argon scores 53 on its Intelligence Index with high reasoning, matching GPT-6 Astra (max) and one point ahead of GPT-6.1 Sol (max). At the current 50% launch discount, its cost per task is $1.99, about 60% of GPT-6 Astra's $3.26, but the discount's end date is unconfirmed and standard pricing would raise it to $3.98. The model is being rolled out to selected users and is not publicly available.

    Why it matters: The benchmark compares Gemini 4 Argon's cost per task and hallucination rate with GPT-6 Astra, showing where its value depends on a temporary 50% discount.

  14. Kling AI BlogOfficialAI score62

    Kling 4.0 enters early access with 30-second native video generation

    AIKling 4.0 is entering early access, with a wider rollout planned for October, and Kling 4.0 Flash opens to Ultra Yearly subscribers on September 28. The update generates videos up to 30 seconds in a single pass, accepts up to 15 reference assets, and supports up to 10 keyframe images. Upcoming features include 10-bit HDR output at 4K and 1080p and video extension up to 2 minutes.

    Why it matters: The post specifies concrete capability limits such as 30-second native generation, up to 15 references, and 10 keyframes, which help users judge fit for production workflows.

Sep 29

Sep 29Tue
  1. v0OfficialAI score42

    GPT-6.1 Sol now available in v0

    AIGPT-6.1 Sol is now live in v0, with access via the v0 app link provided in the post. The quoted Vercel post says it is also on AI Gateway and improves on GPT-6 Sol for coding, computer use, multi-step workflows, and complex document analysis.

  2. Tibor BlahoXAI score78

    OpenAI's DevDay 2026 brings dots agents, GPT-6.1 Sol, and Ultrafast speed tier

    AIOpenAI announced more than 20 updates at DevDay 2026, including dots always-on agents, GPT-6.1 Sol, Ultrafast token generation, ChatGPT Space, and a $500/month Pro 500 plan. GPT-6.1 Sol is priced at $2 input and $10 output per 1M tokens and is available in the API as gpt-6.1-sol. Ultrafast generates tokens up to 8x faster in Codex and up to 6x faster in the API.

    Why it matters: The post lists dozens of OpenAI DevDay 2026 changes across models, agents, plans, and APIs, useful for scanning what shipped and who gets access.

    Image from @btibor91's post
  3. DatabricksOfficialAI score34

    Databricks adds GPT-6.1 Sol and Grok 4.7 on Unity Gateway

    AIDatabricks has made OpenAI's GPT-6.1 Sol and xAI's Grok 4.7 available on Unity Gateway the day they launched. The post says GPT-6.1 Sol leads the cost-quality Pareto frontier on OfficeQA Pro v2, while Grok 4.7 reaches the frontier on enterprise document parsing. Unity Gateway also offers access to 60+ other frontier and open models on Databricks.

    Video from @databricks's post
  4. Liquid AIOfficialAI score32

    Liquid AI launches d1, first decision model, beating Jev on HF index

    AILiquid AI announced d1, its first decision model, which it says is the first to outperform Jev on Hugging Face's Decision Index. The company claims d1 wins on multilingual evals, resists prompt injection better, handles longer inputs more effectively, and is built for fast, structured decision-making in software environments. It is available via the Liquid API at console.liquid.ai, with OpenRouter availability coming soon.

    Image from @liquidai's post
  5. Junyang LinXAI score22

    Junyang Lin hopes a model will surpass Opus 5.5

    AIJunyang Lin said he hopes a model will be smarter than Claude Opus 5.5. The post gives no benchmark, price, or release details, and it is a hope rather than a claim of achievement.

  6. Sam AltmanXAI score60

    OpenAI introduces 6.1 Sol near Astra intelligence at one fifth the price

    AIOpenAI's Sam Altman introduces 6.1 Sol, a model with near-Astra intelligence priced at one fifth of Astra and a 95% cache read discount. The quoted post says ultrafast mode gives 8X speed for Astra today and will come to 6.1 Sol soon.

    Why it matters: The source gives concrete pricing and cache discount figures for a near-Astra model, which helps readers compare its cost against Astra for their own workloads.

  7. OpenAIOfficialAI score42

    OpenAI reopens Pro 200 subscriptions with GPT-6.1 Sol and Astra access

    AIOpenAI is reopening Pro 200 subscriptions, providing continued access to frontier models such as Astra. The company also introduced GPT-6.1 Sol, which it says brings near-Astra capabilities to a model usable every day. OpenAI further committed not to reintroduce the 5-hour usage limit, so subscribers can use their full weekly allowance when they want.

  8. BAAI · new models on Hugging FaceOfficialAI score62

    BAAI releases AREX-2, a 27B agent model for self-improving long-horizon tasks

    AIBAAI released AREX-2, a 27B-parameter long-horizon agent model that improves solutions over multiple test-time rounds by proposing, measuring, reflecting, and revising. It was trained on machine-learning and algorithmic-programming tasks with verifiable feedback, and the source reports that this self-improvement transfers to deep research. The model is Apache License 2.0 licensed and has a 262,144-token context length.

    Why it matters: The source compares AREX-2 against closed and open models on coding and deep-research benchmarks, showing how test-time self-improvement is measured across task types.

  9. CognitionOfficialAI score40

    GPT-6.1 Sol Now Available in Devin at Lower Cost

    AIGPT-6.1 Sol is now available in Devin, scoring 60.4% on FrontierCode 1.1, close to GPT-6 Sol's 60.7%. At medium reasoning effort it costs $0.31 per task, 81% less than GPT-6 Sol at max effort.

    Image from @cognition's post
  10. OpenAIOfficialAI score62

    OpenAI makes GPT-6.1 Sol available to Plus, Pro, Business, Enterprise, and Edu users

    AIOpenAI says GPT-6.1 Sol is available starting today to all Plus, Pro, Business, Enterprise, and Edu users. The model is offered in ChatGPT Work and Codex, and the post links to OpenAI's introduction page.

    Why it matters: The post specifies which plan tiers gain GPT-6.1 Sol and in which products, showing where the new model reaches users directly.

  11. OpenAIOfficialAI score37

    GPT-6.1 Sol improves alignment and transparency over GPT-6 Sol

    AIOpenAI reports that GPT-6.1 Sol shows major alignment improvements over GPT-6 Sol in its evaluations, moving closer to GPT-6 Astra. The model is more transparent about its limitations and more reliable at respecting user intent and safety constraints.

    Image from @OpenAI's post
  12. ModelScopeOfficialAI score54

    IQuest-Q1 released as 320B MoE model for long-horizon coding agents

    AIModelScope announced IQuest-Q1, a 320B MoE model with 15B active parameters and a 512K context window for agentic coding. The post reports scores of 84.5 on CyberGym, 83.2 on Terminal-Bench 2.1, 64.6 on DeepSWE v1.1, and 63.0 on NL2Repo, and says weights are released under the IQuest-Q1 License.

    Image from @ModelScope2022's post
  13. ModelScopeOfficialAI score44

    Intern-Decision multimodal models scale structured decisions at 0.8B–4B

    AIShanghai AI Laboratory's Intern-Decision family of 0.8B, 2B, and 4B multimodal models averages 79.38, 84.68, and 90.02 across seven decision benchmarks. Intern-Decision-4B scores 88.74, surpassing Jev while achieving better probability calibration. Reported mean latency is 33.98, 33.28, and 44.16 ms, versus 109.70 ms for Jev in the same local HF setup.

    Image from @ModelScope2022's post
  14. InternLM (Shanghai AI Lab) · new models on Hugging FaceOfficialAI score40

    InternLM releases AdvancedMathBench-AutoVerifier to grade natural-language math proofs

    AIInternLM's AutoVerifier, built on Qwen3_5MoeForConditionalGeneration with about 68 GiB of weights across 40 safetensors shards, evaluates natural-language mathematical proofs, explains errors, and identifies the earliest incorrect step. It serves as the automatic grader for AdvancedMathBench's ProverBench, which accepts a proof only when all eight judgments report -1. The model is a learned grader rather than a formal proof checker and can make errors.

  15. vLLMOfficialAI score58

    IQuest-Q1 320B MoE coding model gets day-0 support in vLLM

    AIvLLM announced day-0 support for IQuest-Q1, a 320B-parameter MoE model with 15B active per token, 256 experts with 8 active, and a 524,288-token context. The post credits existing vLLM features such as the hybrid KV cache coordinator, sinks attention path, and EAGLE speculative decoding with probabilistic draft sampling. The linked material includes a Docker image and vllm serve commands, with and without recursive MTP.

    Image from @vllm_project's post
  16. Together AIOfficialAI score22

    Qwen3.8-Flash gets 40% off through month's end on Together AI

    AITogether AI is offering 40% off Qwen3.8-Flash through the rest of the month, a window it suggests for running evaluations. Alibaba's Qwen3.8-Flash is designed for high-volume applications such as coding and coworking assistants, with an emphasis on quality at low cost.

    Image from @togethercompute's post
  17. Artificial Analysis ArticlesOfficialAI score78

    GPT-6.1 Sol replaces GPT-6 Sol with near-Astra intelligence at lower cost

    AIArtificial Analysis reports that GPT-6.1 Sol replaces GPT-6 Sol after seven days and scores 1 point below GPT-6 Astra on the Intelligence Index. At max effort it costs $0.72 per Intelligence Index task, compared with $3.26 for GPT-6 Astra and $1.05 for GPT-6 Sol. Its pricing matches GPT-6 Sol at $2/$10 per million input/output tokens, but it uses about 10-30% more output tokens.

    Why it matters: The source compares GPT-6.1 Sol against GPT-6 Sol, GPT-5.6 Sol, and GPT-6 Astra on cost per task and token use, helping readers weigh performance against price.

Sep 28

Sep 28Mon
  1. ModelScopeOfficialAI score44

    Audio8 ASR Infinite enables unlimited-length streaming speech transcription with bounded memory

    AIAudio8 ASR Infinite transcribes Chinese and English audio of unlimited length using a rolling KV Cache that avoids accumulated drift. At a 480 ms delay, it reports 1.75 CER on AISHELL-1, 2.89 on AISHELL-4, and 3.04/6.81 WER on LibriSpeech test-clean/test-other. The preview release is under Apache 2.0, with deployment through an adapted vLLM stack.

    Video from @ModelScope2022's post
  2. SpaceXAIOfficialAI score46

    Grok 4.7 is now available on Amazon Bedrock

    AIaccording to the SpaceXAI account, which is operated by xAI and Grok. The post gives no further details on pricing, context length, or benchmarks.

    Video from @SpaceXAI's post
  3. Simon WillisonXAI score55

    Sonnet 5.5 becomes the free-tier model on claude.ai

    AISimon Willison says Claude Sonnet 5.5 now powers the free tier on claude.ai, so free users can run the kinds of experiments he describes. He contrasts this with ChatGPT's free tier, which he says still runs the less capable GPT-5.6 Luna.