Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Sep 29

Sep 29Tue
  1. Azure BlogOfficialAI score40

    SQL Server on Azure Local Becomes Generally Available for Connected and Disconnected Use

    AIMicrosoft has made SQL Server on Azure Local generally available for connected and disconnected deployments, letting organizations run SQL Server in their own datacenters and edge locations. Disconnected operations continue locally where external connectivity is restricted or unavailable. Eligible existing SQL Server licenses can be used, and Foundry Local on Azure Local, currently in preview, brings AI inference alongside SQL Server data.

  2. Meta NewsroomOfficialAI score34

    Meta Launches Forum, a Standalone App for Browsing Facebook Groups

    AIMeta is testing Forum, a standalone iOS and Android app in the US that syncs with users' Facebook Groups to consolidate their conversations in one place. The update adds a new top-contributor role replacing previous badges, an AI-powered Ask feature that surfaces group posts and comments, and topic labels for exploring interests.

  3. PerplexityOfficialAI score60

    Perplexity open-sources Bumblebee to scan developer machines for risky packages

    AIPerplexity has open-sourced Bumblebee, a read-only scanner for macOS and Linux that checks developer machines for risky packages, extensions, and AI tool configurations. When connected to Computer, it can trigger deeper scans whenever a new supply-chain risk emerges. The post says Computer reviews findings from Bumblebee and Numbat to propose better detection rules, and humans approve every change before it ships.

  4. Microsoft ResearchOfficialAI score75

    Microsoft Research introduces Quine, a multimodal biology world model and research harness

    AIMicrosoft Research introduced Quine, an experimental research system combining a multimodal world model of biology with an interactive harness that connects models, scientific tools, literature, and researchers. In a pancreatic cancer study with the Broad Institute, Quine prioritized compounds that shifted tumor cell states, and several top-ranked candidates were validated in wet-lab assays. Access is initially limited to the Quine Fellows program and select collaborations, and the system is intended for research use only, not clinical use.

    Why it matters: The post shows how a multimodal biology world model is wired into a harness, grounded in one wet-lab cancer example and a limited fellows-program access path.

  5. ModelScopeOfficialAI score12

    Qwen-Image-2.1 users can submit issues via official feedback form

    AIModelScope invites users experiencing problems with Qwen-Image-2.1 to submit feedback through an official form that connects them directly to the Qwen Image research team. Users are asked to share prompts, images, or workflows to help the team investigate and improve the model. Submissions in English or Chinese are welcome.

    Image from @ModelScope2022's post
  6. IEEE Spectrum · AINewsAI score14

    IC-STAR Brings Full-Flow Autonomous AI to Digital and Analog Chip Design

    AIThe webinar presents IC-STAR, an autonomous AI approach that shifts silicon engineers from manually managing tools and handoffs to defining objectives and supervising AI-driven execution across the chip development lifecycle. It covers four enabling technologies and includes a look at Ambiq's production deployment of autonomous AI. The source provides no performance figures or availability details.

  7. ModelScopeOfficialAI score44

    Intern-Decision multimodal models scale structured decisions at 0.8B–4B

    AIShanghai AI Laboratory's Intern-Decision family of 0.8B, 2B, and 4B multimodal models averages 79.38, 84.68, and 90.02 across seven decision benchmarks. Intern-Decision-4B scores 88.74, surpassing Jev while achieving better probability calibration. Reported mean latency is 33.98, 33.28, and 44.16 ms, versus 109.70 ms for Jev in the same local HF setup.

    Image from @ModelScope2022's post
  8. InternLM (Shanghai AI Lab) · new models on Hugging FaceOfficialAI score40

    InternLM releases AdvancedMathBench-AutoVerifier to grade natural-language math proofs

    AIInternLM's AutoVerifier, built on Qwen3_5MoeForConditionalGeneration with about 68 GiB of weights across 40 safetensors shards, evaluates natural-language mathematical proofs, explains errors, and identifies the earliest incorrect step. It serves as the automatic grader for AdvancedMathBench's ProverBench, which accepts a proof only when all eight judgments report -1. The model is a learned grader rather than a formal proof checker and can make errors.

  9. OpenBMBOfficialAI score34

    MiniCPM-o 4.5 now runs in SGLang Omni v0.1.7 for developers

    AIOpenBMB announced that MiniCPM-o 4.5 is now supported in SGLang Omni v0.1.7, giving developers more flexibility to run and build with the model. The background release notes add that MiniCPM-o 4.5 brings multimodal input and speech output to the runtime. MiniCPM-o and MiniMax-Music3 also gained Intel XPU support in the same release.

  10. vLLMOfficialAI score58

    IQuest-Q1 320B MoE coding model gets day-0 support in vLLM

    AIvLLM announced day-0 support for IQuest-Q1, a 320B-parameter MoE model with 15B active per token, 256 experts with 8 active, and a 524,288-token context. The post credits existing vLLM features such as the hybrid KV cache coordinator, sinks attention path, and EAGLE speculative decoding with probabilistic draft sampling. The linked material includes a Docker image and vllm serve commands, with and without recursive MTP.

    Image from @vllm_project's post
  11. Azure BlogOfficialAI score46

    Microsoft Fabric and Copilot Integration: New Data Foundation Features for Agents

    AIMicrosoft is bringing business context from Fabric IQ into Microsoft Copilot, with Fabric IQ in Copilot Chat and Cowork generally available and integration into the new Code experience coming soon through the Frontier program. Power BI is also gaining agentic app creation in Power BI Desktop, letting users generate applications from trusted semantic models and publish them to Microsoft Fabric.

  12. SGLangOfficialAI score53

    SGLang adds Day-0 support for IQuest-Q1 with a single-node serve command

    AISGLang says it has Day-0 support for IQuest-Q1, an open-source sparse MoE model with 320B total and 15B active parameters for coding and agentic tasks. The post includes a single-node serving command for H200 GPUs in BF16, using tensor parallelism of 8, EAGLE speculative decoding, and the iquest_q1 reasoning and tool-call parsers. The image marks the command as not verified.

    Image from @sgl_project's post
  13. Together AIOfficialAI score22

    Qwen3.8-Flash gets 40% off through month's end on Together AI

    AITogether AI is offering 40% off Qwen3.8-Flash through the rest of the month, a window it suggests for running evaluations. Alibaba's Qwen3.8-Flash is designed for high-volume applications such as coding and coworking assistants, with an emphasis on quality at low cost.

    Image from @togethercompute's post
  14. howie.seriousXAI score13

    Claude Code account bans may stem from poor IP quality

    AIThe post says Claude Code account bans often result from low-quality IP addresses, and recommends the IPCheck.ing tool to check IP quality. It adds that many promoted home broadband and VPS services should be verified with such a tool before trusting them. Background from @jason5ng32 says IPCheck.ing updated its IP quality scoring to v3 and that one VPS provider widely recommended on X scored 54 in a sampled IP check.

  15. Mastra BlogOfficialAI score42

    Mastra Adds Memory Hooks to Observe and Modify Agent Memory Cycles

    AIMastra has added memory hooks that let developers monitor or alter an agent's observational memory cycles. Lifecycle hooks such as onObservationStart and onReflectionEnd report on each cycle, including token usage for spotting cost spikes, while transform hooks like beforeObservation and afterReflection can prune, remove, or redact memory data.

  16. Manus BlogOfficialAI score50

    Manus Flex lets users connect their own API keys to the Manus workspace

    AIManus is launching Manus Flex, a module that lets users power Manus agents with their own API key from a supported inference provider. Model inference is billed directly by that provider, while other services used in Manus tasks still consume Manus credits. OpenRouter, Fireworks, and Modal are announced as initial inference partners for the Flex Inference Partner Program.

  17. Artificial Analysis ArticlesOfficialAI score78

    GPT-6.1 Sol replaces GPT-6 Sol with near-Astra intelligence at lower cost

    AIArtificial Analysis reports that GPT-6.1 Sol replaces GPT-6 Sol after seven days and scores 1 point below GPT-6 Astra on the Intelligence Index. At max effort it costs $0.72 per Intelligence Index task, compared with $3.26 for GPT-6 Astra and $1.05 for GPT-6 Sol. Its pricing matches GPT-6 Sol at $2/$10 per million input/output tokens, but it uses about 10-30% more output tokens.

    Why it matters: The source compares GPT-6.1 Sol against GPT-6 Sol, GPT-5.6 Sol, and GPT-6 Astra on cost per task and token use, helping readers weigh performance against price.

Sep 28

Sep 28Mon
  1. ModelScopeOfficialAI score44

    Audio8 ASR Infinite enables unlimited-length streaming speech transcription with bounded memory

    AIAudio8 ASR Infinite transcribes Chinese and English audio of unlimited length using a rolling KV Cache that avoids accumulated drift. At a 480 ms delay, it reports 1.75 CER on AISHELL-1, 2.89 on AISHELL-4, and 3.04/6.81 WER on LibriSpeech test-clean/test-other. The preview release is under Apache 2.0, with deployment through an adapted vLLM stack.

    Video from @ModelScope2022's post
  2. KreaOfficialAI score12

    Krea launches a video generation tool at

    AIKrea, an AI creative platform, has published a dedicated video page at The post provides only this link, so no features, models, pricing, or capabilities are confirmed.

  3. KreaOfficialAI score22

    Seedance 2.5 Draft Mode now available on Krea

    AIKrea has launched Draft Mode for Seedance 2.5, letting users experiment with 480p generations before switching to 1080p once a scene is right. The post directs readers to try the feature on Krea's platform.

    Video from @krea_ai's post
  4. Amp NewsOfficialAI score34

    Amp Adds Plaid Speed for GPT-6 Astra Modes at 6x Speed and Cost

    AIAmp now supports Plaid speed for modes that use GPT-6 Astra, using OpenAI's ultrafast tier to run inference up to 6× faster at 6× cost per token. Plaid works only with Amp-provided inference, not linked ChatGPT subscriptions, and subagents and non-Plaid inference fall back to fast or standard speed.

  5. DatabricksOfficialAI score38

    Claude Sonnet 5.5 now available on Databricks across AWS, Azure, GCP

    AIDatabricks now offers Anthropic's Claude Sonnet 5.5 on AWS, Azure, and GCP, governed through Unity Gateway. The post says Sonnet 5.5 is more efficient than Sonnet 5 for coding and agentic use and reaches Opus 5-level accuracy on document understanding, parsing, and search. It joins Claude Opus 5.5, Claude Fable 5.1, and 60+ other open-source and frontier models on the platform.

    Video from @databricks's post
  6. François CholletXAI score32

    K3-Node: a Keras 3 GNN library running on JAX, PyTorch, and TF

    AIK3-Node is a graph neural network library built natively on Keras 3, with models that run on JAX, PyTorch, and TensorFlow with hardware acceleration including Apple Silicon and TPU. According to the post, it achieves 100% public API parity with PyG and incorporates foundation models and architectures from Spektral and StellarGraph.

  7. Hugging FaceOfficialAI score14

    Hugging Face asks developers for Gradio feature requests

    AIHugging Face is inviting developers to suggest features for Gradio, its tool for building AI app frontends. The post follows a quoted message from Abid Lab noting that Gradio's original purpose has shifted and that the team is redesigning it from scratch around developers' current challenges with AI apps.

  8. Perplexity DevelopersOfficialAI score44

    Perplexity adds reusable custom agents to its Agent API

    AIPerplexity says developers can now build custom reusable agents in its Agent API using Profiles, Skills, and managed connectors. Agents are configured once in the API Portal and can then be reused across applications and workflows.

    Video from @perplexitydevs's post
  9. Google AIOfficialAI score44

    Google Labs expands experimental CC agent into a family group assistant

    AIGoogle Labs has expanded Project CC, its experimental AI productivity assistant, into a group agent designed to streamline family household logistics. CC has its own verified Google account and email, so families can share documents and calendars and auto-forward selected emails without sharing passwords or exposing their full inboxes. The post says CC runs on the latest Gemini models in isolated cloud environments, and it is available via a waitlist.

  10. Mike KriegerXAI score67

    Anthropic releases Claude Sonnet 5.5, 30% faster and up to 30% cheaper than Sonnet 5

    AIAnthropic has released Claude Sonnet 5.5, the second model in the Claude 5.5 family. The company says it is more than 30% faster than Sonnet 5 and costs up to 30% less for most work.

    Why it matters: The post gives concrete speed and price changes against Sonnet 5, which helps readers judge whether the upgrade fits their workloads and budgets.