Ling-3.1-flash now free on OpenCode with 262K context
AIOpenCode is offering inclusionAI's latest model, Ling-3.1-flash, for free. The model has 560B total parameters, 25B active parameters, and a 262K context window.
Updated
Updated
Showing low-relevance items too. Hide low-relevance items
AIOpenCode is offering inclusionAI's latest model, Ling-3.1-flash, for free. The model has 560B total parameters, 25B active parameters, and a 262K context window.
AIAnt Ling has made Ling-3.1-flash available on OpenRouter at no cost, inviting users to try it and share feedback. The post provides no details on model size, benchmarks, context length, or pricing beyond the free access.
AIKilo Code is offering Ling 3.1 Flash for free until October 13, with the model served by Novita Labs. Ant Ling's background post describes the model as roughly 560B total parameters with about 25B active per token and a context window of up to 1M tokens. Ant Ling says it scores 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional, and plans to open-source it soon.
AIAi2 released AstaBrief 8B, an open-weights model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. The model runs as Fast mode in Asta, averaging 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The post also describes filtering synthetic training data by citation density and building DPO pairs judged by two models that agreed.
Why it matters: The post explains how supervised fine-tuning, preference data, and citation-density filtering were used to build a cited-report model, which is useful for teams training their own models.
AIinclusionAI has released Ling 3.1 Flash, a hybrid reasoning mixture-of-experts model with 560B total parameters, of which 25B are active. The source does not provide benchmark scores, pricing, context length, or availability details.
AIKarl's AI Watts (@aiwarts) says Opus 5.5 is very strong and that goodcase.ai delivers over 20% improvement across all network conditions and multiple platforms. The post gives no benchmark data or methodology to support the 20% figure.

AIAi2 released AstaBrief 8B, a model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. In Asta's Generate a report feature, Fast mode averages 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The model is built on Qwen3-8B with supervised fine-tuning and DPO, and institutions can run its open weights on their own infrastructure.
Why it matters: The post explains the data filtering and one-pass generation choices behind a fast open-weights report model, showing what worked and what did not.
AITinker says Fulcrum trains its Echo writing model by building on a base model that already writes well, tailoring both SFT and RL to separate the default LLM voice from authors' voices. The post calls this customization approach both cheap and effective. Fulcrum says Echo beats frontier models at writing tasks such as fiction and technical explanations, at a training cost under $5K.
AIOpenAI says GPT-6.1 Sol is now good, cheap, and fast, after slow speeds caused by a load spike in its first two days. A global reset for all paid ChatGPT accounts is scheduled for Friday, October 2, at 10AM PT.
AIcheap typed answers for the small calls inside a harness (routing, approvals, judging), big model for the rest
AIGLM 5.3 and GLM 5.3 Flash are now available in Cursor. GLM 5.3 Max is the best-scoring open-weight model on CursorBench 4.0.

AICloudflare says it is releasing two decision models it trained, Clef and Clef-flash. The main post links to a blog post with details, but the text provided gives no further specifications, benchmarks, or pricing.
AINVIDIA has released PixelUMM, an encoder-free unified multimodal model with 15,199,672,064 parameters that handles text, image, and video understanding and generation directly in pixel space. It represents images as 16-by-16 RGB pixel patches on a Qwen3-8B language backbone, with iterative denoising for generation. The checkpoint is licensed for non-commercial research or evaluation only, while the source code is under Apache License 2.0.
AIPerplexity open-sourced pplx-decider-27b, a state-of-the-art multimodal Decision Model, and offers it through a new Decisions API. Input tokens cost $0.04 per million and output tokens are free, with further price cuts promised in the coming days.
AIKrea (@krea_ai) shared a link to a page for an image model it refers to as "bfl-3-image." The post gives no further details on capabilities, pricing, or release terms.
AIVercel says the Laya decision model is available free on AI Gateway through October 31, in partnership with Boundless. The post suggests using it to route agent work, triage support requests, and check guardrails.
AIHiggsfield AI has made FLUX 3 Image available to try on its platform, linking directly to its image generation page with the flux-3-image model selected. The post provides no details on capabilities, pricing, or benchmarks.
AIHiggsfield has launched FLUX 3 Image, a new Black Forest Labs image model, now available on its platform. The model lets users edit a single detail without altering the rest of the image, place objects using bounding boxes, combine up to 10 reference images, and generate output at up to 4K resolution.
AIMike Knoop notes that a verification of the roughly 27B-parameter model is notable because it is about as large as fits within the official ARC Prize Kaggle competition runtime. ARC Prize reports Qwen3.8-27B from Alibaba's Qwen team scored 42.4% on ARC-AGI v2 at $0.45 per task and 87.5% on v1 at $0.22 per task.

AIARC Prize has published full ARC-AGI results for Alibaba's Qwen3.8-27B, with the testing run hosted by Baseten. The post links the public leaderboard, the open-source benchmarking repository for reproducing the results, and the testing policy.
AIGrok 4.7 is now available on the Gemini Enterprise Agent Platform, according to the post from SpaceXAI, the account owned by xAI and Grok. The post gives no further details on pricing, context length, or capabilities.
AIBlack Forest Labs has released FLUX 3 Image, which generates native 4K images and supports hyper-specific layouts using bounding boxes. It can make multiple targeted edits at once while staying consistent across them, and up to 10 references can be used to compose an image. The first week is 50% off, and an open-weights version is coming in the coming weeks.

AIBlack Forest Labs says customers can contact the company for commercial weights for its FLUX 3 image model. The post also points users to a playground to try the model and to a product page with more information.
AIBlack Forest Labs is discounting its FLUX 3 Image model by 50% through its API until October 8th. The company is also making FLUX 3 Image available under a commercial weights license, letting companies fine-tune it and deploy it on their own infrastructure.
AIBlack Forest Labs announces FLUX 3 Image, which supports multi-turn edits that leave other pixels unchanged, layout control via bounding boxes, generation up to 4K, and up to 10 reference images. Commercial weights are available for companies running image generation at scale, and an open weights version is launching in the coming weeks.
AIReplicate says it is powering Griffin from Tavus on its platform. Tavus describes Griffin as the first model to pass the video Turing test, with 48% of live conversation participants believing it was a real human. The model ranks first on NVIDIA's full-duplex AI video benchmark.
AILiquid AI's d1 model is now available through OpenRouter, as the post says demand for it has been very high. The post notes that users can now access it easily via the OpenRouter platform.
AIPerplexity has published the weights for pplx-decider-v1-27b, a model fine-tuned from Qwen3.8-27B with a 250k-token context window. The weights are available on Hugging Face, and the post points developers to the Perplexity Decisions API quickstart for getting started.
AIMicrosoft Copilot begins rolling out OpenAI's GPT-6.1 Sol and Anthropic's Claude Sonnet 5.5 today, joining Claude Opus 5.5 and GPT-6 Sol added earlier this month. Users can pick the model suited to each task, with Work IQ grounding responses in their files, meetings, and chats within existing permissions. The rollout starts today in Copilot Cowork and Copilot Studio, with Word, Excel, PowerPoint, and Chat following in phases over the coming week.
AICloudflare has open-sourced Clef and Clef-Flash, its first homegrown decision models, now available on Workers AI. Harrison Chase praised the release as an open-weights entrant and argued that agent harnesses should be able to swap their decision model as easily as their main model.
AIApodex has released Apodex 1.1 Mini, a free reasoning-first model designed for complex, long-horizon research and forecasting tasks. According to the source, it works directly with files, data, code, and tools to produce verifiable results.
AIMicrosoft AI's MAI models are now accessible to developers via Vercel, including the newest releases MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash. The partnership brings these models into Vercel's AI Gateway as another route for building Microsoft AI models into applications.
AICloudflare's Workers AI team has released Clef, its first models, two fast decision models that it says top benchmarks for quality and latency. They are available hosted on Workers AI or as open weights on Hugging Face.
AIVercel says it partnered with Microsoft AI to make MAI-Voice-2.1 and MAI-Transcribe-2-Streaming available through AI Gateway today. MAI-Voice-2.1 handles long-form and fast-reply speech, while MAI-Transcribe-2-Streaming provides transcription.
AIMicrosoft AI launched MAI-Transcribe-2-Streaming, which Artificial Analysis ranks #1 of 38 models for final transcript accuracy at 2.5% WER, returned 0.13s after end of speech. Artificial Analysis lists its streaming price at $0.54 per hour of audio, at the higher end among leading streaming models. Microsoft's post claims the model is 55% faster and 60% cheaper than ElevenLabs and invites developers to build agents on its platform.
AIMicrosoft AI says three models are available today, with more details provided in a linked post. The post does not name the models or give specifications, benchmarks, or pricing.
AIMicrosoft AI announced three new models: MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash. The source says the streaming transcription model is accurate and aims for natural speech with less waiting between conversational turns for building voice agents.

AICloudflare released Clef and Clef-flash, two decision models hosted on Workers AI and open-sourced on Hugging Face under Apache 2.0, and launched a reinforcement learning fine-tuning service. In Cloudflare's tests, Clef classified a domain in 2.2s versus 4.7s for gpt-oss-120b, and the models are Jev-API compatible. The company is offering fine-tuning first through a forward-deployed engineering team, with a self-serve platform planned later.
AIGoogle says Gemini 4 Argon is rolling out at $2/$10 per million tokens, though the author has not yet been able to access the model to test it. OpenAI pulled GPT-6.1 Astra over alignment failures and released GPT-6.1 Sol, which it prices at the same $2/$10 and says shows substantial alignment improvements over GPT-6 Sol. The post also covers Anthropic's leaked IPO prospectus, which reportedly lists roughly $518 billion in compute commitments, and a court ruling upholding the Department of War's supply chain risk designation of Anthropic.
AIPareto 26.10 Preview is a multimodal composite model built for research, coding, and agentic workflows. It is described as delivering frontier-level performance across a broad range of general-purpose tasks, though the source excerpt is a preview and provides no benchmark scores, parameter counts, pricing, or availability details.