Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 7

Oct 7Wed
  1. ThariqOfficialAI score67

    Claude Haiku 5.5 returns as a cheaper, faster small model

    AIAnthropic has released Claude Haiku 5.5, which it describes as the cheapest, fastest, and most capable small model it has released. On average it costs around 75% less to run than Claude Haiku 4.5, and the author says it is 10x cheaper than Haiku 4.5 under 100k tokens. It can be tried with computer use, workflows, and the API.

    This story has a top pick“Anthropic releases Claude Haiku 5.5 as its cheapest, fastest small model”

  2. Satya NadellaOfficialAI score72

    Windows adds on-device agents, local coding models, and Hybrid Intelligence

    AIMicrosoft says Windows will bring unmetered intelligence to PCs, letting agents work securely on-device. The post lists MAI-Code-1.1 Flash, a 137B parameter coding model with a 256K context window optimized to run on PCs, and GitHub Copilot handoffs to local models. It also describes Hybrid Intelligence, which lets Copilot act on the PC and keep sensitive work local, and Code in Copilot for building software without cloud token spend, on devices such as Surface Laptop Ultra powered by NVIDIA RTX Spark.

    Why it matters: The post lists concrete Windows agent, local model, and device changes, showing how coding work may shift from cloud tokens toward on-device execution.

    Image from @satyanadella's post
  3. Satya NadellaXAI score20

    Microsoft outlines Windows direction for hybrid intelligence

    AISatya Nadella linked to a Windows Experience blog post describing what Microsoft announced about building Windows for hybrid intelligence. The post itself gives no further details, so the specific features and scope cannot be confirmed from this source.

  4. Vercel DevelopersOfficialAI score29

    Claude Haiku 5.5 is now live on Vercel AI Gateway

    AIVercel says Anthropic's Claude Haiku 5.5, model ID anthropic/claude-haiku-5.5, is now available on AI Gateway. The company describes it as the fastest Claude model at standard speed, built for subagents and summarization, and the first Haiku to offer effort levels, with ZDR supported.

  5. OpenRouterOfficialAI score38

    Cloudflare's Clef decision models now available on OpenRouter

    AICloudflare's open-source Clef (27B) and Clef Flash (9B) decision models are available on OpenRouter. They accept text, JSON, or images and return typed answers with probabilities rather than generated text. Pricing is $0.24 per M input tokens for Clef and $0.09 per M for Clef Flash, with output free.

  6. CursorOfficialAI score38

    Claude Haiku 5.5 priced at $0.10/M input tokens, Sonnet cache cut

    AIAnthropic's Claude Haiku 5.5 is priced at $0.10 per million input tokens and $0.50 per million output tokens, rising to $0.50 and $2.50 above 100k input tokens. Claude Sonnet 5.5 cache reads have also dropped from $0.20 to $0.10 per million tokens. Cursor points readers to its CursorBench evaluations to compare Haiku 5.5.

  7. CursorOfficialAI score32

    Claude Haiku 5.5 is now available in Cursor

    AICursor now offers Claude Haiku 5.5, which costs 10x less than Claude Haiku 4.5 on shorter requests. Users can enable it under Cursor Settings > Models.

    Image from @cursor_ai's post
  8. Replit ⠕OfficialAI score34

    Replit previews desktop app with Microsoft for local Windows builds

    AIReplit announced a preview of its desktop app, built with Microsoft, that builds and runs apps locally on Windows. Each build runs in its own sandbox powered by Microsoft Execution Containers and NVIDIA OpenShell. Early access is available through a waitlist at replit.com.

    Video from @Replit's post
  9. LangChainOfficialAI score42

    LangChain releases Managed Deep Agents v0.9 with schedules and per-run configuration

    AILangChain says Managed Deep Agents v0.9 lets agents create their own reminders, follow-ups, and recurring tasks mid-conversation through a Schedules SDK. Per-run configuration lets users choose the model, skills, MCP servers, and sandbox for each run, so one deployment can serve multiple teams or repos. The update also adds Slack Reactions, where agents react to messages as soon as they start a run.

  10. 🚨 AI News | TestingCatalogXAI score62

    Anthropic releases Claude Haiku 5.5, its fastest and cheapest model

    AIAnthropic has released Claude Haiku 5.5, which the author describes as its fastest and cheapest model to date. The source says it costs about 75% less to run than Claude Haiku 4.5 and is the first Haiku model with an adjustable effort setting. The attached benchmark table reports Haiku 5.5 scores on tasks including computer use (OSWorld 2.1 offline subset, 72.4%) and Terminal-Bench 4.0 (39.2%), compared with Haiku 4.5 and other models.

    Image from @testingcatalog's post
  11. Claude Code · GitHub ReleasesOfficialAI score36

    Claude Code v2.1.293 adds Claude Haiku 5.5 and fixes dozens of bugs

    AIClaude Code v2.1.293 adds Claude Haiku 5.5 (claude-haiku-5-5), now the default Haiku model on the Anthropic API, with 1M context and pricing of $0.10/$0.50 per Mtok ($0.50/$2.50 for prompts over 100K). The release also adds agentType to the subagentStatusLine payload and isDeferred to $.tool.register, and fixes numerous issues including a memory leak in HTTP MCP connections.

  12. ClaudeDevsOfficialAI score29

    Anthropic Announces Tiered Token Pricing Above and Below 100K Tokens

    AIAnthropic's ClaudeDevs account states pricing for prompts under 100K tokens at $0.10 input and $0.50 output per million tokens, with cache reads at $0.01. For prompts above 100K tokens, the rates rise to $0.50 input and $2.50 output per million tokens, with cache reads at $0.05.

  13. ClaudeDevsOfficialAI score62

    Claude Platform rolls out monthly API credits for Max and Team plans

    AIAnthropic's ClaudeDevs account announced monthly Claude Platform API credits for Max and Team plans. Max 5x receives $100, Max 20x receives $200, and Team receives up to $500 in pooled credits. The credits work on any model, including Haiku 5.5, in users' own code or third-party harnesses.

  14. ClaudeDevsOfficialAI score42

    Claude Haiku 5.5 released, costing about 75% less than Haiku 4.5

    AIAnthropic has made Claude Haiku 5.5 available on the Claude Platform and in Claude Code, costing around 75% less to run than Haiku 4.5. The post recommends pairing it with Opus 5.5 or Sonnet 5.5 as a subagent for high-volume, cost-sensitive tasks such as summaries, compactions, or database queries.

    Video from @ClaudeDevs's post
  15. OpenAIOfficialAI score82

    GPT-6 with Intelligent UI rolls out to ChatGPT Chat tab across tiers

    AIOpenAI is rolling out GPT-6 with Intelligent UI globally to Plus, Pro, Business, and Enterprise users today, with Free and Go users following starting tomorrow. Plus, Pro, Business, and Enterprise tiers are powered by GPT-6 Sol, while Free and Go tiers use GPT-6 Luna, and both are tuned for everyday conversation. The update applies only to the Chat tab, and the models powering Work and Codex are not changing.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  16. OpenAIOfficialAI score62

    OpenAI's GPT-6 Intelligent UI composes responses with visuals and interactive elements

    AIOpenAI says GPT-6 can now compose responses using text, visuals, and interactive elements, choosing how they fit together based on the user's question. Responses can include graphics and charts that help explain an idea, along with tappable buttons, forms, and interactive experiences usable directly in the conversation.

    Video from @OpenAI's post
  17. Mike KriegerOfficialAI score44

    Anthropic launches Claude Haiku 5.5, a faster, cheaper small model

    AIAnthropic introduced Claude Haiku 5.5, which it calls the cheapest, fastest, and most capable small model it has released. On average, it costs about 75% less to run than Claude Haiku 4.5. The post positions Haiku 5.5 for high-volume work alongside Opus handling heavier reasoning tasks.

  18. ClaudeOfficialAI score40

    Claude Haiku 5.5 now available on all major cloud platforms

    AIAnthropic has released Claude Haiku 5.5, available now across all platforms including Amazon Web Services, Google Cloud, and Microsoft Azure. The post links to Anthropic's announcement page for further details.

  19. ClaudeOfficialAI score26

    Claude Haiku 5.5 offers strong value on tasks under 100,000 tokens

    AIAnthropic says Claude Haiku 5.5 is especially good value for tasks under 100,000 tokens, which made up around 90% of requests to its previous Haiku model. The post points to cost-effectiveness for the typical workload rather than providing pricing or benchmark figures.

    Image from @claudeai's post
  20. ClaudeOfficialAI score38

    Anthropic's Haiku 5.5 targets high-volume, cost-sensitive tasks

    AIAnthropic's Haiku 5.5 is built for high-volume, cost-sensitive work such as summaries and classification. It can serve as a subagent alongside Claude Opus 5.5 and Sonnet 5.5 on coding tasks. It is also fast enough for live customer support and browser use.

  21. NVIDIA AI DeveloperOfficialAI score22

    NVIDIA announces what's new in CUDA 13.4

    AINVIDIA's developer account promoted a broadcast covering what's new in CUDA 13.4. The post provides only the release name and a link, so no specific features, benchmarks, or details are stated.

  22. Wired · AINewsAI score62

    OpenAI's ChatGPT Intelligent UI generates interactive visuals for answers

    AIOpenAI unveiled an Intelligent UI update for ChatGPT that generates custom visual elements when they help answer a question. The update is powered by GPT-6, rolling out to paid users immediately and to free users the next day. A reviewer's test produced an annotated slug diagram, an apartment affordability calculator with sliders, and a clickable airplane seat explorer, and OpenAI says users can ask for fewer visual outputs.

  23. IThome · AINewsAI score42

    Nvidia unveils DGX Station for Windows, a desktop AI supercomputer for running trillion-parameter models

    AINvidia announced DGX Station for Windows, a desktop AI supercomputer built on the NVIDIA GB300 Grace Blackwell Ultra Desktop Superchip with up to 748GB of unified memory, able to run models of up to about one trillion parameters locally. The machine offers up to 20 PFLOPS of AI compute and combines 252GB of HBM3e GPU memory with 496GB of LPDDR5X CPU memory. It is scheduled to go on sale in the fourth quarter of 2026.

  24. NVIDIAOfficialAI score38

    Jaguar Type 01 launches powered by NVIDIA Hyperion and Halos systems

    AIThe new Jaguar Type 01 is powered by NVIDIA Hyperion, a computer and sensor platform that processes what the car sees and senses to support real-time decisions. It is paired with NVIDIA Halos, a safety system covering chips through software, and the software passed 150,000 tests over tens of thousands of hours before reaching the road. Over-the-air updates will continue improving the vehicle after it leaves the showroom.

    Video from @nvidia's post
  25. Georgi GerganovXAI score44

    llama.cpp adds ggml RPC for distributing inference across heterogeneous devices

    AIllama.cpp can distribute inference across heterogeneous devices through the ggml RPC backend, according to Georgi Gerganov. He says it is currently an advanced setting, but he expects it to become more accessible to regular users over time. A related post reports MiMo 2.6 Flash running across an RTX 6000 GPU and an M5 laptop over 10 GbE at about 40 tokens/sec.

  26. GitHub Blog · AI & MLOfficialAI score57

    GitHub argues secret protection must scale with AI-driven code growth

    AIGitHub reports that one in three pull requests now involves an AI agent, and that public secret exposures rise with the volume of pushes rather than from declining developer care. It introduces a ModernBERT-based classifier with Microsoft Applied Sciences that evaluates candidate secrets in under two milliseconds and could more than double the secrets prevented at push time. The feature is in private preview, with availability for GitHub Secret Protection customers later this month.