Skip to contentSkip to stories

Updated

#Product update

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 22

Sep 22Tue
  1. LlamaIndex 🦙OfficialAI score22

    LiteParse v2.14.6 parses text PDFs about 25% faster locally

    AILlamaIndex released LiteParse v2.14.6, an open-source PDF-to-Markdown parser that processes text-based PDFs about 25% faster. On realistic documents it handled pages at 2.8ms per page, 1.5 times faster than the next-fastest local parser. It runs locally in Python, Node.js, Rust, or directly in the browser.

    Image from @llama_index's post
  2. Mike KriegerXAI score62

    Anthropic cuts API prices to $4 and $20 per million tokens

    AIThe company cut API pricing to $4 per million input tokens and $20 per million output tokens, which it says is 20% less than Opus 5. Cache reads are also 60% cheaper, and subscribers get higher five-hour rate limits plus a banked rate limit reset.

  3. Mike KriegerXAI score67

    Anthropic launches Claude Opus 5.5, leading in coding and knowledge work

    AIAnthropic has launched Claude Opus 5.5, the first model in its new Claude 5.5 family. According to the quoted launch post, it performs at the level of Claude Fable 5.1 for most tasks and costs 40% less to run than Opus 5. The author says it leads in coding and knowledge work and praises its writing quality.

    Why it matters: The quoted launch post gives a concrete cost comparison, useful for weighing Opus 5.5 against earlier Opus and Fable 5.1 models for routine work.

  4. Lydia Hallie ✨XAI score23

    Claude paid plans get a usage limit reset until Oct 22

    AIAnthropic's Pro, Max, and Team subscribers can claim a usage limit reset in Settings → Usage, available until Oct 22. Opus 5.5 is now the default for paid plans and is priced lower than Opus 5, so 5-hour and weekly limits go 25% further.

  5. v0OfficialAI score52

    Claude Opus 5.5 is now available in v0

    AIv0 announced that users can now use Claude Opus 5.5 in v0, with a link to try it. The quoted Anthropic post says Opus 5.5 is the first model in its new Claude 5.5 family, performs at the level of Claude Fable 5.1 on most tasks, and costs 40% less to run than Opus 5.

  6. AnthropicOfficialAI score71

    Anthropic releases Claude Opus 5.5, the first model in its Claude 5.5 family

    AIAnthropic has made Claude Opus 5.5 available today, introducing it as the first model in its new Claude 5.5 family. According to the quoted @claudeai post, it performs at the level of Claude Fable 5.1 on most tasks and costs 40% less to run than Opus 5.

    Why it matters: The post gives a concrete cost comparison against Opus 5 and names the model family, helping readers gauge the trade-off between price and performance.

  7. StepFunOfficialAI score52

    StepFun releases Step Code v0.1.0 as an open-source coding CLI

    AIStepFun has released Step Code v0.1.0, an open-source command-line tool under the MIT License that covers reading and editing code, running tests, and shipping from one CLI. The post reports 80.9% on Terminal-Bench 2.1 and 73.3% on Multi-Frame, a 150-task long-horizon benchmark from StepFun. It also includes one-command static site publishing with StepPage and links the GitHub repository.

    Image from @StepFun_ai's post
  8. Daniel HanXAI score42

    Qwen-Image-2.1 runs locally in Unsloth Desktop via INT8, FP8, GGUF

    AIDaniel Han says Qwen-Image-2.1 works in Unsloth Desktop through INT8, FP8, and GGUF builds, with Unsloth also releasing dynamic GGUFs for it. Pinned RAM offloading lets INT8 and FP8 fit under 6–8GB of VRAM while remaining relatively fast. The linked Unsloth post says the 7B model runs on 12GB VRAM and performs on par with Nano Banana 2.0.

  9. Tencent HyOfficialAI score34

    Tencent Hunyuan's Hy Image3.5 preview launches free on Miora for two weeks

    AITencent Hunyuan has released a preview of Hy Image3.5 on Miora, a design platform, and is offering it free for two weeks. The model keeps existing canvas workflows and remembers users' brand rules while making edits. Miora's background post says the free period runs through October 7 and supports up to 2K output for text-to-image and image-to-image.

  10. OpenBMBOfficialAI score59

    VoxWeft runs real-time interpretation locally on Apple Silicon using VoxCPM2

    AIOpenBMB highlights VoxWeft, an open-source simultaneous interpretation system for Apple Silicon built by developer @HenryZ30734018 on an MLX implementation of VoxCPM2. The system turns live speech into translated speech on-device, with first audio streaming in about 170 ms on an M5 MacBook. VoxCPM2 generates speech in 30 languages, supports direct language-pair interpretation, and clones a target voice from about 5 seconds of reference audio.

    Video from @OpenBMB's post
  11. TechNode · AINewsAI score60

    Alibaba's T-Head unveils Zhenwu V900 AI chip with full-stack system design

    AIT-Head, Alibaba's chip subsidiary, unveiled the Zhenwu V900 AI chip for training and inference at the 2026 Apsara Conference in Hangzhou. The company claims three times the performance of its predecessor, the Zhenwu M890, with 216GB of memory, 1,200GB/s inter-chip bandwidth, and mass production expected in the first quarter of 2027.

  12. Kimi.aiOfficialAI score46

    Kimi launches browser extension for chatting, automating web tasks

    AIKimi has released its Kimi Browser Extension, formerly Kimi WebBridge, which runs in the browser sidebar to navigate websites and fill out forms. Users can record repetitive steps once and save them as a skill for Kimi to reuse later. The extension is available now on the Chrome Web Store.

    Video from @Kimi_Moonshot's post
  13. AI SupremacyBlogAI score45

    TypeSafe AI's Jev Is a Non-LLM Probabilistic Classifier for Fast Software Decisions

    AITypeSafe AI released Jev, a transformer-based System-1 model that outputs calibrated probabilistic decisions instead of generating tokens, returning answers in 70–500 ms at $0.042 per million input tokens. The model is built for typed Choice, Score, and yes/no questions inside software pipelines, and it is available to everyone without a waitlist, with $5 in starting credits. Vercel, Cloudflare, LangChain, and Langfuse have added Jev to their platforms.

  14. KrASIA · Big TechNewsAI score60

    Alibaba unveils Zhenwu V900 chip and targets over 20 GW data center capacity by 2032

    AIAlibaba is preparing its Zhenwu V900 proprietary accelerator for commercial release in the first quarter of 2027, claiming three times the performance of the M890. The company also targets more than 20 gigawatts of global data center capacity by 2032. Its capital expenditures rose 75% year-on-year in the June quarter, while AI cloud revenue grew 44.9% and cloud adjusted EBITA margin reached about 12%.

  15. X.PINXAI score42

    Alibaba targets 20GW cloud capacity by 2032, unveils Zhenwu V900 chip

    AIAlibaba CEO Eddie Wu said on September 22 that strong AI demand is driving infrastructure investment, targeting over 20GW of global cloud data-center capacity by 2032. T-Head unveiled the Zhenwu V900, claiming 3× the compute performance of the M890 and support for clusters of up to 500,000 accelerator cards. Servers using V900 chips are scheduled to launch in Q1 2027.

    Image from @thexpin's post
  16. Gemini API ChangelogOfficialAI score62

    Gemini 3.8 Flash TTS and Flash-Lite TTS become generally available with a new Voices endpoint

    AIGoogle made the Gemini 3.8 Flash TTS and Gemini 3.8 Flash-Lite TTS models generally available, along with the Gemini API Voices endpoint. Flash TTS is positioned for studio-grade voice fidelity and long-form multi-turn stability, while Flash-Lite TTS targets high-throughput, real-time voice agents and replaces gemini-3.1-flash-tts-preview. The update adds voice design, voice replication with consent verification, and access to 150+ prebuilt and custom voices.

Sep 21

Sep 21Mon
  1. Kimi.aiOfficialAI score34

    Kimi K3 now available on Amazon Bedrock

    AIMoonshot AI's Kimi K3 is now available on Amazon Bedrock for coding, document analysis, and extended agent workflows. Bedrock provides access, encryption, and auditing controls, and explicit prompt caching is supported.

    Image from @Kimi_Moonshot's post
  2. Tencent HyOfficialAI score38

    Tencent Hunyuan releases Hy Image3.5 preview for image generation

    AITencent Hunyuan has launched a preview of Hy Image3.5, which it says wins 30% more often than Hy Image3.0 in human evaluation. The model supports text-to-image and image-to-image generation at up to 2K resolution with improved consistency, and is priced at $0.024 per image on the Tencent Cloud API, with reference images free. Two weeks of free access is offered through OnSolo and Miora.

    Video from @TencentHunyuan's post
  3. xAI News (Grok)OfficialAI score46

    How SpaceXAI uses Grok Bot to scale customer support without new hires

    AISpaceXAI says its combined support team handled a 175% rise in tickets without hiring, crediting Grok Bot, which it says would otherwise have required about 200 additional staff. The company reports resolving tickets for $0.20 to $0.30 each, versus the $1 to $4 per resolution it attributes to traditional AI support tools. Grok Bot is also reported to resolve 99% of refund requests without human intervention.

  4. vLLM BlogOfficialAI score60

    vllm-metal brings concurrent vLLM serving to Apple Silicon Macs

    AIvllm-metal ports vLLM's scheduler, paged KV cache, and OpenAI-compatible server to Apple Silicon, with MLX and Metal handling execution. The v0.28.0 release added batched MTP, GGUF and hybrid-model support, and faster prefill on M5, and v0.29.0 is installable through Homebrew.

    Why it matters: The post explains how vllm-metal packs requests and pages KV cache on Apple Silicon, with benchmarks showing where concurrent serving gains and tradeoffs appear.

  5. Amp NewsOfficialAI score36

    Amp Runners Add Git Worktree Creation and Secret Injection

    AIAmp runners can now create Git worktrees from the directory picker, with each new worktree placed as a sibling folder on a new branch off the current HEAD while uncommitted changes stay in the original checkout. Runners can also opt in with --amp-env to inject Secrets & Env Vars configured on ampcode.com into thread shell commands, MCP servers, and plugins, with changes applied to the next thread without a restart.

  6. Google Developers BlogOfficialAI score38

    Google Colab premium benefits now included in Google AI plans

    AIGoogle AI subscribers now get premium Colab benefits, including priority access to faster accelerators and more powerful machines. Google AI Ultra subscribers also get uninterrupted background execution and Premium GPU access for long training runs. The benefits roll out over the next few weeks in Colab-supported countries, and existing Colab subscriptions are unchanged.

  7. Together AI BlogOfficialAI score36

    Together AI's canary rollouts upgrade production models without downtime

    AITogether AI's canary rollouts shift production traffic between two model deployments on the same endpoint in staged percentages, with optional metric gates between steps. Operators can choose canary, blue-green, or rolling strategies, and a rollout starts only when explicitly launched; it can be paused, canceled, or reversed. The platform scales the target before moving traffic and waits for routing to converge before draining the source.

  8. Waymo BlogOfficialAI score28

    Waymo launches transit rewards in San Francisco Bay Area, paying riders Waymo Cash for transit trips

    AIWaymo is launching a transit rewards program in the San Francisco Bay Area that issues $2.85 in Waymo Cash when riders take a Waymo trip and public transit within 2 hours, using a linked Visa card. The program covers all 27 Bay Area transit agencies that accept contactless Visa tap-to-pay, and employees get access first before a public rollout in the coming weeks. Waymo will also lease 40 dedicated Caltrain station parking spaces for staged vehicles.

  9. Xiaomi MiMoOfficialAI score42

    Xiaomi launches MiMo-V2.6 Pro UltraSpeed with up to 20× faster generation

    AIXiaomi's MiMo-V2.6-Pro UltraSpeed is now available in MiMo Desktop and through the API, delivering up to 20× faster generation. MiMo Desktop and membership plans launch alongside Pro and Flash, and users can run both models via subscription or their own API key. API pricing remains unchanged from V2.5.

  10. Xiaomi MiMoOfficialAI score36

    Xiaomi MiMo-V2.6 unifies code, design, and tool use across creative outputs

    AIXiaomi's MiMo-V2.6 combines code, design, and tool use to build frontend interfaces, presentations, Figma-linked visual assets, and videos. The post says MiMo-V2.5-TTS supports narration in video production, and that the model can compose music, including an orchestral piece for around ten instruments that can be converted to MIDI. On Design Arena, the Pro version reportedly performs comparably to Claude Opus 5 and GPT-5.6 Sol.

    Image from @XiaomiMiMo's post