Skip to contentSkip to stories

Updated

#Model release

Showing low-relevance items too. Hide low-relevance items

Oct 2

Oct 2Fri
  1. Ant LingOfficialAI score27

    Ling-3.1-flash now available free on OpenRouter

    AIAnt Ling has made Ling-3.1-flash available on OpenRouter at no cost, inviting users to try it and share feedback. The post provides no details on model size, benchmarks, context length, or pricing beyond the free access.

  2. Kilo (acq. by Anaconda)OfficialAI score36

    Ling 3.1 Flash is free in Kilo Code until October 13

    AIKilo Code is offering Ling 3.1 Flash for free until October 13, with the model served by Novita Labs. Ant Ling's background post describes the model as roughly 560B total parameters with about 25B active per token and a context window of up to 1M tokens. Ant Ling says it scores 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional, and plans to open-source it soon.

  3. Hugging Face BlogOfficialAI score70

    Ai2 open-sources AstaBrief 8B, a fast model for generating cited research reports

    AIAi2 released AstaBrief 8B, an open-weights model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. The model runs as Fast mode in Asta, averaging 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The post also describes filtering synthetic training data by citation density and building DPO pairs judged by two models that agreed.

    Why it matters: The post explains how supervised fine-tuning, preference data, and citation-density filtering were used to build a cited-report model, which is useful for teams training their own models.

  4. Ai2 (Allen Institute for AI)OfficialAI score67

    Ai2 open-sources AstaBrief 8B, a fast open-weights scientific report model

    AIAi2 released AstaBrief 8B, a model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. In Asta's Generate a report feature, Fast mode averages 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The model is built on Qwen3-8B with supervised fine-tuning and DPO, and institutions can run its open weights on their own infrastructure.

    Why it matters: The post explains the data filtering and one-pass generation choices behind a fast open-weights report model, showing what worked and what did not.

  5. TinkerOfficialAI score33

    Tinker praises Fulcrum's cheap, effective style-customization training approach

    AITinker says Fulcrum trains its Echo writing model by building on a base model that already writes well, tailoring both SFT and RL to separate the default LLM voice from authors' voices. The post calls this customization approach both cheap and effective. Fulcrum says Echo beats frontier models at writing tasks such as fiction and technical explanations, at a training cost under $5K.

Oct 1

Oct 1Thu
  1. Peter Steinberger 🦞XAI score42

    Cloudflare releases Clef and Clef-flash decision models

    AICloudflare says it is releasing two decision models it trained, Clef and Clef-flash. The main post links to a blog post with details, but the text provided gives no further specifications, benchmarks, or pricing.

  2. NVIDIA · new models on Hugging FaceOfficialAI score44

    NVIDIA releases PixelUMM, an encoder-free model for pixel-space image and video tasks

    AINVIDIA has released PixelUMM, an encoder-free unified multimodal model with 15,199,672,064 parameters that handles text, image, and video understanding and generation directly in pixel space. It represents images as 16-by-16 RGB pixel patches on a Qwen3-8B language backbone, with iterative denoising for generation. The checkpoint is licensed for non-commercial research or evaluation only, while the source code is under Apache License 2.0.

  3. KreaOfficialAI score12

    Krea adds BFL image model link to its platform

    AIKrea (@krea_ai) shared a link to a page for an image model it refers to as "bfl-3-image." The post gives no further details on capabilities, pricing, or release terms.

  4. Vercel DevelopersOfficialAI score24

    Laya decision model free on Vercel AI Gateway through October 31

    AIVercel says the Laya decision model is available free on AI Gateway through October 31, in partnership with Boundless. The post suggests using it to route agent work, triage support requests, and check guardrails.

  5. Higgsfield AI 🧩OfficialAI score28

    Higgsfield launches FLUX 3 Image model on its platform

    AIHiggsfield AI has made FLUX 3 Image available to try on its platform, linking directly to its image generation page with the flux-3-image model selected. The post provides no details on capabilities, pricing, or benchmarks.

  6. Higgsfield AI 🧩OfficialAI score28

    Higgsfield adds Black Forest Labs' FLUX 3 Image model for editing and 4K generation

    AIHiggsfield has launched FLUX 3 Image, a new Black Forest Labs image model, now available on its platform. The model lets users edit a single detail without altering the rest of the image, place objects using bounding boxes, combine up to 10 reference images, and generate output at up to 4K resolution.

    Video from @higgsfield's post
  7. Mike KnoopXAI score44

    Qwen3.8-27B verified on ARC-AGI, nearly fitting Kaggle runtime limits

    AIMike Knoop notes that a verification of the roughly 27B-parameter model is notable because it is about as large as fits within the official ARC Prize Kaggle competition runtime. ARC Prize reports Qwen3.8-27B from Alibaba's Qwen team scored 42.4% on ARC-AGI v2 at $0.45 per task and 87.5% on v1 at $0.22 per task.

    Image from @mikeknoop's post
  8. SpaceXAIOfficialAI score46

    Grok 4.7 now available on Gemini Enterprise Agent Platform

    AIGrok 4.7 is now available on the Gemini Enterprise Agent Platform, according to the post from SpaceXAI, the account owned by xAI and Grok. The post gives no further details on pricing, context length, or capabilities.

    Video from @SpaceXAI's post
  9. ReplicateOfficialAI score47

    FLUX 3 Image launches with native 4K generation and bounding-box layout control

    AIBlack Forest Labs has released FLUX 3 Image, which generates native 4K images and supports hyper-specific layouts using bounding boxes. It can make multiple targeted edits at once while staying consistent across them, and up to 10 references can be used to compose an image. The first week is 50% off, and an open-weights version is coming in the coming weeks.

    Image from @replicate's post
  10. Black Forest LabsOfficialAI score54

    Black Forest Labs introduces FLUX 3 Image with precise editing controls

    AIBlack Forest Labs announces FLUX 3 Image, which supports multi-turn edits that leave other pixels unchanged, layout control via bounding boxes, generation up to 4K, and up to 10 reference images. Commercial weights are available for companies running image generation at scale, and an open weights version is launching in the coming weeks.

    Video from @bfl_ai's post
  11. ReplicateOfficialAI score46

    Replicate powers Tavus's Griffin, a video Turing test-passing model

    AIReplicate says it is powering Griffin from Tavus on its platform. Tavus describes Griffin as the first model to pass the video Turing test, with 48% of live conversation participants believing it was a real human. The model ranks first on NVIDIA's full-duplex AI video benchmark.

  12. Microsoft CopilotOfficialAI score34

    Microsoft Copilot adds GPT-6.1 Sol and Claude Sonnet 5.5 models

    AIMicrosoft Copilot begins rolling out OpenAI's GPT-6.1 Sol and Anthropic's Claude Sonnet 5.5 today, joining Claude Opus 5.5 and GPT-6 Sol added earlier this month. Users can pick the model suited to each task, with Work IQ grounding responses in their files, meetings, and chats within existing permissions. The rollout starts today in Copilot Cowork and Copilot Studio, with Word, Excel, PowerPoint, and Chat following in phases over the coming week.

  13. OpenRouter · New modelsBlogAI score36

    Apodex 1.1 Mini Released as Free Reasoning Model for Research Tasks

    AIApodex has released Apodex 1.1 Mini, a free reasoning-first model designed for complex, long-horizon research and forecasting tasks. According to the source, it works directly with files, data, code, and tools to produce verifiable results.

  14. Microsoft AIOfficialAI score36

    Microsoft's MAI models now available through Vercel AI Gateway

    AIMicrosoft AI's MAI models are now accessible to developers via Vercel, including the newest releases MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash. The partnership brings these models into Vercel's AI Gateway as another route for building Microsoft AI models into applications.

  15. Vercel DevelopersOfficialAI score32

    Vercel adds Microsoft AI speech and transcription models to AI Gateway

    AIVercel says it partnered with Microsoft AI to make MAI-Voice-2.1 and MAI-Transcribe-2-Streaming available through AI Gateway today. MAI-Voice-2.1 handles long-form and fast-reply speech, while MAI-Transcribe-2-Streaming provides transcription.

  16. Mustafa SuleymanXAI score40

    Microsoft AI launches MAI-Transcribe-2-Streaming, claiming top real-time transcription accuracy

    AIMicrosoft AI launched MAI-Transcribe-2-Streaming, which Artificial Analysis ranks #1 of 38 models for final transcript accuracy at 2.5% WER, returned 0.13s after end of speech. Artificial Analysis lists its streaming price at $0.54 per hour of audio, at the higher end among leading streaming models. Microsoft's post claims the model is 55% faster and 60% cheaper than ElevenLabs and invites developers to build agents on its platform.

  17. Microsoft AIOfficialAI score15

    Microsoft AI announces three new models now available today

    AIMicrosoft AI says three models are available today, with more details provided in a linked post. The post does not name the models or give specifications, benchmarks, or pricing.

  18. Cloudflare Blog · AIOfficialAI score58

    Cloudflare releases open-source Clef decision models and an RL fine-tuning service

    AICloudflare released Clef and Clef-flash, two decision models hosted on Workers AI and open-sourced on Hugging Face under Apache 2.0, and launched a reinforcement learning fine-tuning service. In Cloudflare's tests, Clef classified a domain in 2.2s versus 4.7s for gpt-oss-120b, and the models are Jev-API compatible. The company is offering fine-tuning first through a forward-deployed engineering team, with a self-serve platform planned later.

  19. Don't Worry About the Vase (Zvi Mowshowitz)BlogAI score62

    AI #188: Gemini 4 Argon, GPT-6.1 Sol, and Anthropic's IPO Filing

    AIGoogle says Gemini 4 Argon is rolling out at $2/$10 per million tokens, though the author has not yet been able to access the model to test it. OpenAI pulled GPT-6.1 Astra over alignment failures and released GPT-6.1 Sol, which it prices at the same $2/$10 and says shows substantial alignment improvements over GPT-6 Sol. The post also covers Anthropic's leaked IPO prospectus, which reportedly lists roughly $518 billion in compute commitments, and a court ruling upholding the Department of War's supply chain risk designation of Anthropic.

  20. OpenRouter · New modelsBlogAI score36

    Pareto 26.10 Preview: A Multimodal Model for Research, Coding and Agents

    AIPareto 26.10 Preview is a multimodal composite model built for research, coding, and agentic workflows. It is described as delivering frontier-level performance across a broad range of general-purpose tasks, though the source excerpt is a preview and provides no benchmark scores, parameter counts, pricing, or availability details.