Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 7

Oct 7Wed
  1. RunwayOfficialAI score36

    Runway launches an ability to work directly inside ChatGPT

    AIRunway says users can brief its tool, let it work, and give notes from the same chat window with Runway directly inside ChatGPT Astra. The post invites readers to get started through a linked page.

    Video from @runwayml's post
  2. Harrison ChaseXAI score27

    deepagents now dynamically loads tools when skills are loaded

    AILangChain's deepagents now supports dynamically loading tools when a skill is loaded, so skills requiring specific tools no longer need those tools always available. With OpenAI and Anthropic models, this can be done without breaking the prompt cache.

  3. AWS Machine Learning BlogOfficialAI score56

    Claude Haiku 5.5 becomes available on Amazon Bedrock and Claude Platform on AWS

    AIAnthropic's Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic, it is the fastest and most efficient model in the Claude 5.5 family and costs around 75 percent less than Claude Haiku 4.5 for most tasks. The post also covers pairing it with Claude Opus 5.5 as a subagent layer and provides Boto3, Converse, and Anthropic SDK examples for calling the model.

  4. Hacker News · AI (150+ points)BlogAI score52

    Meta and Microsoft cut employee use of Anthropic's Claude as they shift to in-house coding tools

    AIAccording to The Information, Meta and Microsoft are reducing employee use of Anthropic's Claude while moving toward their own coding tools. Microsoft's expected internal Anthropic spending above $1 billion a year has fallen by more than a third, and its monthly AI spending limits per employee were reportedly cut from $100,000 to about $10,000 in most cases. Meta's Claude Code users reportedly fell from about 60,000 to 30,000, though it still reportedly spent over $105 million on Claude Code over 28 days.

  5. OpenRouterOfficialAI score47

    Anthropic's Claude Haiku 5.5 launches on OpenRouter with lower prices

    AIAnthropic's Claude Haiku 5.5 is now available on OpenRouter as the first Haiku model supporting reasoning efforts. OpenRouter says it is about 75% cheaper than its predecessor, runs at over 100 tokens per second, and shows a clear benchmark capability gain.

    Image from @OpenRouter's post
  6. NVIDIA BlogOfficialAI score67

    NVIDIA and Microsoft Launch RTX Spark Laptops and DGX Station for Windows AI Agents

    AINVIDIA and Microsoft announced RTX Spark laptops and compact desktops that run the full NVIDIA AI stack locally, with laptop preorders open today and sales from October 16. Microsoft also announced general availability of Microsoft Execution Containers (MXC), an OS-level infrastructure for agents to run securely in the background, while NVIDIA previewed DGX Station for Windows with 748GB of coherent memory and up to 20 petaFLOPS of FP4 compute.

    Why it matters: The announcement pairs Windows agent infrastructure with local hardware, showing how agents may move onto personal computers and enterprise desktops rather than only cloud services.

  7. DatabricksOfficialAI score22

    Databricks' Genie Ontology infers business context without a fully built ontology

    AIDatabricks' Genie Ontology infers relevant business context across data and assets to answer open-ended questions without waiting for a fully built ontology. Advancing Analytics' Simon Whiteley examines how OntoRank decides what to trust and why certified assets still matter for data governance.

    Video from @databricks's post
  8. Tibor BlahoXAI score62

    ChatGPT adds Intelligent UI powered by GPT-6 Instant

    AIChatGPT now includes Intelligent UI, which lets GPT-6 Instant generate interface elements inside chat. The quoted OpenAI engineer says the team aimed to keep HTML's power while making the interface feel fast and native, and that post-training GPT-6 to judge when an interface helps remains an open challenge.

  9. 👩‍💻 Paige BaileyOfficialAI score36

    Google's EmbeddingGemma 2 model released on Hugging Face for science

    AIGoogle released EmbeddingGemma 2 on Hugging Face, and Paige Bailey called it a step toward open models for open science. The post links the model and cites earlier EmbeddingGemma-based projects, including medical, geographic, oncology, and PubMed embedding models.

  10. AWS Machine Learning BlogOfficialAI score38

    AWS Adds Real-Time Access Checks to RAG in Amazon Quick and Bedrock Knowledge Bases

    AIAWS has added real-time access control list checks to Amazon Quick and Amazon Bedrock Knowledge Bases, verifying user permissions directly with sources like Google Drive at query time. The two-stage design first runs semantic search with cached ACLs, then confirms each candidate document against the authoritative source before passing passages to the LLM. This closes gaps where permissions changed between periodic syncs.

  11. Design ArenaOfficialAI score44

    Claude Haiku 5.5 is now available on Design Arena

    AIDesign Arena has added Anthropic's Claude Haiku 5.5, which the post describes as the company's fastest and most capable small model yet. It is aimed at high-volume, cost-sensitive work such as coding, classification, summarization, database queries, and speed-sensitive workflows like customer support and browser use. Per the background post from Claude, it costs around 75% less to run than Claude Haiku 4.5.

    Image from @DesignArena's post
  12. Gizmodo · AINewsAI score46

    Google Launches SynthID.com to Check Images and Videos for AI Watermarks

    AIGoogle launched SynthID.com, letting users upload an image or video to check whether it was made with AI. The tool detects only content created with tools from Google, OpenAI, Nvidia, and Kakao, and it requires signing in with a Google, Apple, or ChatGPT account. Gizmodo's tests found Gemini and Grok gave inaccurate or unsupported answers about AI-generated images, so the results should not be treated as definitive.

  13. Vaibhav (VB) SrivastavOfficialAI score72

    ChatGPT rolls out GPT-6 with interactive in-conversation tools

    AIOpenAI is rolling out GPT-6 and Intelligent UI in ChatGPT, adding interactive diagrams, calculators, and tools directly inside conversations. GPT-6 can also begin answering while still thinking, and access starts today for Plus, Pro, Business, and Enterprise users, with Free and Go users following tomorrow.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  14. MarkTechPostNewsAI score60

    Liquid AI releases open-weight d1-3B and d1-omni-600M decision models

    AILiquid AI released two open-weight multimodal decision models, d1-3B and d1-omni-600M, which return probability answers in one forward pass with zero output tokens. d1-3B scores 48.57 on Decision Index v0.2.1 and answers one question in 8 ms on an RTX 4090, while the models are licensed free for commercial use below $10 million in annual revenue.

  15. Greg BrockmanOfficialAI score62

    OpenAI introduces Intelligent UI for interactive answers in ChatGPT

    AIOpenAI is rolling out Intelligent UI in ChatGPT, a capability for answering quickly with fully interactive user interfaces. According to the quoted post, it makes everyday questions more visual, helps explain complex topics, and provides interactive tools for tasks on the spot. The quoted post also says GPT-6 is rolling out to everyone alongside the feature.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  16. NVIDIA Technical BlogOfficialAI score34

    NVIDIA Teaches Robots to Assemble GB300 Tester Trays

    AINVIDIA's technical blog describes how its team taught robots to assemble GB300 tester trays, a task that currently requires skilled manual labor in factories. The post also discusses lessons about robot learning, mechanical intelligence, and engineering. The source excerpt provides limited detail beyond this.

  17. Lydia Hallie ✨XAI score34

    Set Haiku 5.5 autocompact to 100K to stay in cheaper tier

    AIAnthropic's Lydia Hallie says API-billed users can set Haiku 5.5's autocompact window to 100K to remain in the cheaper token pricing tier. The setting is saved per model, so it applies only to Haiku, including subagents, and is configured with /model haiku followed by /autocompact 100k. Per the background post, prompts under 100K tokens cost $0.10/$0.50 per million tokens with $0.01 cache reads, versus $0.50/$2.50 with $0.05 cache reads above 100K.

  18. ThariqOfficialAI score67

    Claude Haiku 5.5 returns as a cheaper, faster small model

    AIAnthropic has released Claude Haiku 5.5, which it describes as the cheapest, fastest, and most capable small model it has released. On average it costs around 75% less to run than Claude Haiku 4.5, and the author says it is 10x cheaper than Haiku 4.5 under 100k tokens. It can be tried with computer use, workflows, and the API.

    This story has a top pick“Anthropic releases Claude Haiku 5.5 as its cheapest, fastest small model”

  19. Satya NadellaOfficialAI score72

    Windows adds on-device agents, local coding models, and Hybrid Intelligence

    AIMicrosoft says Windows will bring unmetered intelligence to PCs, letting agents work securely on-device. The post lists MAI-Code-1.1 Flash, a 137B parameter coding model with a 256K context window optimized to run on PCs, and GitHub Copilot handoffs to local models. It also describes Hybrid Intelligence, which lets Copilot act on the PC and keep sensitive work local, and Code in Copilot for building software without cloud token spend, on devices such as Surface Laptop Ultra powered by NVIDIA RTX Spark.

    Why it matters: The post lists concrete Windows agent, local model, and device changes, showing how coding work may shift from cloud tokens toward on-device execution.

    Image from @satyanadella's post
  20. Satya NadellaXAI score20

    Microsoft outlines Windows direction for hybrid intelligence

    AISatya Nadella linked to a Windows Experience blog post describing what Microsoft announced about building Windows for hybrid intelligence. The post itself gives no further details, so the specific features and scope cannot be confirmed from this source.

  21. Vercel DevelopersOfficialAI score29

    Claude Haiku 5.5 is now live on Vercel AI Gateway

    AIVercel says Anthropic's Claude Haiku 5.5, model ID anthropic/claude-haiku-5.5, is now available on AI Gateway. The company describes it as the fastest Claude model at standard speed, built for subagents and summarization, and the first Haiku to offer effort levels, with ZDR supported.

  22. OpenRouterOfficialAI score38

    Cloudflare's Clef decision models now available on OpenRouter

    AICloudflare's open-source Clef (27B) and Clef Flash (9B) decision models are available on OpenRouter. They accept text, JSON, or images and return typed answers with probabilities rather than generated text. Pricing is $0.24 per M input tokens for Clef and $0.09 per M for Clef Flash, with output free.

  23. Google Cloud TechOfficialAI score16

    Google Cloud Run hackathon on Product Hunt launches October 14

    AIGoogle Cloud Tech is hosting a Cloud Run hackathon on Product Hunt, asking developers to deploy side projects to production with Cloud Run. Participants launch their projects on Product Hunt on October 14 to take part. More details are available at the linked page.

    Video from @GoogleCloudTech's post
  24. CursorOfficialAI score38

    Claude Haiku 5.5 priced at $0.10/M input tokens, Sonnet cache cut

    AIAnthropic's Claude Haiku 5.5 is priced at $0.10 per million input tokens and $0.50 per million output tokens, rising to $0.50 and $2.50 above 100k input tokens. Claude Sonnet 5.5 cache reads have also dropped from $0.20 to $0.10 per million tokens. Cursor points readers to its CursorBench evaluations to compare Haiku 5.5.

  25. CursorOfficialAI score32

    Claude Haiku 5.5 is now available in Cursor

    AICursor now offers Claude Haiku 5.5, which costs 10x less than Claude Haiku 4.5 on shorter requests. Users can enable it under Cursor Settings > Models.

    Image from @cursor_ai's post
  26. Replit ⠕OfficialAI score34

    Replit previews desktop app with Microsoft for local Windows builds

    AIReplit announced a preview of its desktop app, built with Microsoft, that builds and runs apps locally on Windows. Each build runs in its own sandbox powered by Microsoft Execution Containers and NVIDIA OpenShell. Early access is available through a waitlist at replit.com.

    Video from @Replit's post
  27. LangChainOfficialAI score42

    LangChain releases Managed Deep Agents v0.9 with schedules and per-run configuration

    AILangChain says Managed Deep Agents v0.9 lets agents create their own reminders, follow-ups, and recurring tasks mid-conversation through a Schedules SDK. Per-run configuration lets users choose the model, skills, MCP servers, and sandbox for each run, so one deployment can serve multiple teams or repos. The update also adds Slack Reactions, where agents react to messages as soon as they start a run.

  28. 🚨 AI News | TestingCatalogXAI score62

    Anthropic releases Claude Haiku 5.5, its fastest and cheapest model

    AIAnthropic has released Claude Haiku 5.5, which the author describes as its fastest and cheapest model to date. The source says it costs about 75% less to run than Claude Haiku 4.5 and is the first Haiku model with an adjustable effort setting. The attached benchmark table reports Haiku 5.5 scores on tasks including computer use (OSWorld 2.1 offline subset, 72.4%) and Terminal-Bench 4.0 (39.2%), compared with Haiku 4.5 and other models.

    Image from @testingcatalog's post
  29. Claude Code · GitHub ReleasesOfficialAI score36

    Claude Code v2.1.293 adds Claude Haiku 5.5 and fixes dozens of bugs

    AIClaude Code v2.1.293 adds Claude Haiku 5.5 (claude-haiku-5-5), now the default Haiku model on the Anthropic API, with 1M context and pricing of $0.10/$0.50 per Mtok ($0.50/$2.50 for prompts over 100K). The release also adds agentType to the subagentStatusLine payload and isDeferred to $.tool.register, and fixes numerous issues including a memory leak in HTTP MCP connections.

  30. ClaudeDevsOfficialAI score29

    Anthropic Announces Tiered Token Pricing Above and Below 100K Tokens

    AIAnthropic's ClaudeDevs account states pricing for prompts under 100K tokens at $0.10 input and $0.50 output per million tokens, with cache reads at $0.01. For prompts above 100K tokens, the rates rise to $0.50 input and $2.50 output per million tokens, with cache reads at $0.05.

  31. ClaudeDevsOfficialAI score62

    Claude Platform rolls out monthly API credits for Max and Team plans

    AIAnthropic's ClaudeDevs account announced monthly Claude Platform API credits for Max and Team plans. Max 5x receives $100, Max 20x receives $200, and Team receives up to $500 in pooled credits. The credits work on any model, including Haiku 5.5, in users' own code or third-party harnesses.