Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. meng shaoXAI score49

    LangChain adds three Deep Agents Skills upgrades: tool binding, pinning, reloading

    AILangChain has added three engineering upgrades to Skills in its Deep Agents framework: tool-binding Skills, pinned Skills, and mid-thread reloading. Tool-binding lets a SKILL.md declare tools via metadata.include_tools, so tools are injected only when the Skill is read, and pinned Skills inject full instructions before the next model call, skipping a round trip. Setting skills_metadata to None rescans the Skills library mid-thread without restarting, at the cost of invalidating the cache.

    Image from @shao__meng's post
  2. The DecoderNewsAI score72

    Claude Haiku 5.5 cuts prices but uses more tokens than GPT-6 Luna

    AIAnthropic released Claude Haiku 5.5, its fastest and most affordable small model, at prices up to 90 percent lower for most prompts under 100,000 tokens. Artificial Analysis ranks it first among small-class models on its Intelligence Index with a score of 43, but it consumes about three times the output tokens per task that GPT-6 Luna needs at maximum effort.

    This story has a top pick“Anthropic releases Claude Haiku 5.5 as its cheapest, fastest small model”

  3. TechRadar · AINewsAI score25

    HP survey finds three in five UK business leaders say AI has created new roles

    AIA HP survey of nearly 20,000 desk-based workers globally found three in five UK business leaders say AI adoption has created new roles or teams in their organizations. Only 10% said AI is primarily replacing or reducing roles, while 52% of UK workers use employer-provided AI tools daily or weekly, up from 38% last year. Still, 38% of UK knowledge workers worry AI could replace their roles.

  4. QbitAINewsAI score49

    Manus Returns to Beijing, Hiring 17 Roles After Raising Over $500M

    AIManus parent company Butterfly Effect has completed a new financing round of over $500 million, led by Boyu Capital and IDG Capital, and is rebuilding a team in Beijing to develop AI Agent products for the Chinese market. Its recruitment page lists 17 open positions, up from 11 before the National Day holiday, including AI Agent product manager, Agent Harness engineer, Agent evaluation engineer, and LLM algorithm engineer roles.

  5. GuizangXAI score22

    Anthropic's cheaper Haiku 5.5 draws backlash over China rivals

    AIGuizang (@op7418) says he posted news of Anthropic's price cut for Haiku 5.5 and was attacked by commenters who view the pricing as aimed at Chinese models. He compares Haiku 5.5 with DeepSeek-V4.1 flash and Zhipu's GLM 5.3 flash, finding it cheaper than both and scoring one point higher than GLM 5.3 flash on Terminal Bench 4.0 in AA's test.

  6. The DecoderNewsAI score72

    AI hacking tools let a likely single attacker breach multiple South Korean banks

    AIA suspected Chinese-speaking attacker breached several South Korean financial institutions between late September and early October 2026, reportedly stealing over 25,000 records from Shinhan Bank alone. The attacker used ARTEX, a Chinese open-source tool that uses AI language models to automate finding security flaws, and models named in the report include DeepSeek v4.1-flash, GLM-5.3, and Grok 4.6.

    Why it matters: The case shows how AI-driven penetration tools let one attacker breach several banks in a short window, a risk experts had warned about.

  7. Sakana AIOfficialAI score3

    Sakana AI launches "Sakana AI Insider" newsletter with releases and events

    AISakana AI is accepting sign-ups for its newsletter "Sakana AI Insider," which will share product release information, R&D updates, and priority notices for Swag giveaways and events. Subscribers also get early access to event registration, following a recent event that drew many applicants.

    Image from @SakanaAILabs's post
  8. QbitAINewsAI score32

    Geely unveils AI-powered Geely Smart Charging with 2250 kW peak charging power

    AIGeely Automobile Group launched its Geely Smart Charging technology on September 23, 2026, reaching a 2250 kW peak single-gun charging power and keeping maximum temperature at or below 65°C. The system, co-developed with StepFun and built on its PowerMind energy model, reportedly raises battery cycle life by more than 20% and targets county-level coverage by the end of 2027.

  9. The Guardian · AINewsAI score42

    Co-op places legal services staff under AI monitoring of customer phone calls

    AIThe Co-op's legal services arm is using AI models to record and score staff phone calls with customers seeking probate and will advice, with a model from OpenAI assessing more than 50 aspects of each call. Managers use pass and fail scores to analyse employee performance, and the Co-op says the system is a support tool, not a decision-maker. Trade unions and a whistleblower have criticised the monitoring as oppressive.

  10. Gizmodo · AINewsAI score45

    OpenAI Reportedly Regains Nearly Half of AI Compute Market Share From Anthropic in 2026

    AIAccording to a Wall Street Journal report citing OpenRouter data, OpenAI's share of AI compute routed through the platform rose from under 25% at the start of 2026 to nearly 50% last month. The data comes largely from AI-native startups, with some legacy tech companies also included. The report comes as OpenAI reportedly shifted focus from Sora and an erotica generator toward productivity and business customers.

  11. MarkTechPostNewsAI score55

    Architect launches Liquid Inference, an auction-based router for LLM inference

    AIArchitect Financial Technologies launched Liquid Inference, an LLM router where providers bid to serve each request and the buyer pays the lowest offer meeting its rules. Developers can switch by changing the base URL, and the first 500 users get $20 of free inference. The source notes that fees, the provider list, and latency data are not yet public.

  12. MarkTechPostNewsAI score45

    NVIDIA's PivotOPD Trains Multi-Turn AI Agents to Recover From Pivotal Mistakes

    AINVIDIA, Princeton University, and the University of Maryland introduced PivotOPD, an on-policy distillation method that teaches multi-turn LLM agents to recover from their most damaging early mistake. Tested on Qwen3-1.7B and Qwen3-8B students, it posts the best average against 13 baselines on ALFWorld, WebShop, and Search-based QA. It recovers from 72.7% of replayed pivotal mistakes, versus 20.3% for standard OPD, with no added inference cost.

  13. Teknium 🪽XAI score22

    Hermes Agent coming to ASUS RTX Spark PCs soon

    AITeknium announced that Hermes Agent will be available on the ASUS RTX Spark PC when it releases soon. The post links the upcoming support to ASUS's ProArt P16 and P14 creator laptops, which ASUS says are powered by NVIDIA and Hermes AI.

  14. MIT Technology Review · AINewsAI score26

    AVEVA's Arti Garg outlines a safer path to autonomous industrial AI

    AIAVEVA chief technologist Arti Garg argues industrial AI should augment rather than replace human supervisors in critical decisions, with guardrails defining where automated systems can act. She says organizations must rethink business processes and safeguards as foundation models, physical AI, and agentic AI enable more complex automation.

  15. Hacker News · AI (150+ points)BlogAI score40

    OpenAI withdraws three math papers over a sign error in a proof

    AIOpenAI has withdrawn three math manuscripts, including "Algebraicity of Weil classes on split abelian eightfolds," after a sign error invalidated a stabilization-trace cancellation argument. The withdrawals affect two dependent papers, and the withdrawn papers now carry notices linking to archived manuscripts. OpenAI also revised 14 other manuscripts with proof repairs and corrections, and added six formalizations.

  16. CNBC · TechnologyNewsAI score46

    Huawei Mate 90 phones use in-house LogicFolding chip as EV sales slow

    AIHuawei released its first smartphone with its own LogicFolding chip, the Mate 90 series, on October 1, as China's smartphone and car markets slow. Huawei only sells several million smartphones outside China each year, and Huawei-powered vehicle deliveries fell 29% year-on-year in September.

  17. Ant LingOfficialAI score22

    Ant Ling's Ling-3.1-flash now live on AI/ML API

    AIAnt Ling announced a day-zero collaboration with AI/ML API, making Ling-3.1-flash available there for agentic and cowork scenarios. AI/ML API describes it as a 560B-parameter MoE model with about 25B active per token and up to 1M context, built for agents, coding, and long documents. The model is free to try on AI/ML API until October 13.

  18. Ant LingOfficialAI score34

    Ling-3.1-flash is free on OpenCode for a limited time

    AIAnt Ling's Ling-3.1-flash is now free on the OpenCode coding harness for a limited week-long period. The model has 560B total parameters, 25B active parameters, and a 262K context window, according to the OpenCode background post. Ant Ling says it shows strong coding performance and thanks OpenCode for day-zero support.

  19. PandailyNewsAI score45

    KargoBot Launches Mixed Autonomous Freight Network in Ordos With Cabless Robots

    AIKargoBot has started a scaled AI freight network in Qipanjing, Ordos, combining human-driven trucks, autonomous trucks with cabs, and cabless transport robots on one system. The company says cabless robots could raise economic gain per vehicle from 20 percent to more than 30 percent, a target it has not audited. Platooning reportedly improves gross margin by about 10 to 18 percent versus manned haulage, with one lead driver able to head up to five follower trucks.

  20. PandailyNewsAI score37

    Tencent WorkBuddy Builds a WeChat Mini-Program From a Prompt to Preview

    AITencent's WorkBuddy agent can take a plain-language request through to a WeChat mini-program preview and a publish request, according to a hands-on product test reported on October 8. In the test, the agent built a voice notebook with cloud login, storage and files, using Tencent's wand-asr-v1 for speech-to-text and GLM-5.3-Flash for sorting notes. WeChat's own review and filing steps remain outside the agent, so the test does not show that every mini program goes live automatically.

  21. PandailyNewsAI score46

    Galbot and Tsinghua's LATENT Wins IROS Award for Humanoid Tennis Forehand

    AIA Galbot, Tsinghua University and collaborators paper won IROS 2026's Best Entertainment and Amusement Paper Award for LATENT, a humanoid tennis-return method trained on imperfect amateur motion-capture clips. In simulation, the full forehand policy succeeded on 96.52 percent of returns, versus 71.85 percent for PULSE. On a real Unitree G1, the paper reports 90.90 percent forehand success across 20 consecutive rallies, with motion capture still used rather than the robot's own cameras.

  22. PandailyNewsAI score38

    Huawei Presents Experimental XMFS Shared-Memory Filesystem at LPC 2026

    AIHuawei engineers presented XMFS, an experimental Linux kernel prototype filesystem, at the Linux Plumbers Conference in Prague on October 5. It aims to let applications reach cross-node shared memory on CXL 3.0 or Huawei unified bus servers through standard POSIX file calls. The code exists only on openEuler, not in the mainline Linux kernel.

  23. South China Morning Post · TechNewsAI score22

    Anthropic adds Simplified and Traditional Chinese language options to Claude chatbot

    AIAnthropic has added simplified and traditional Chinese as on-screen language options in Claude's web interface in supported markets. The company continues to block users in mainland China, and China briefly appeared as a billing country option before disappearing from settings, according to checks by the South China Morning Post.

  24. Latent SpaceBlogAI score73

    Claude Haiku 5.5 launches at GPT-6 Luna pricing with 1M context

    AIAnthropic released Claude Haiku 5.5, priced the same as OpenAI's GPT-6 Luna, with a 1M-token context window. Artificial Analysis scored it 43 on its Intelligence Index, slightly ahead of GPT-6 Luna at 38, but it uses about 3x more output tokens at max effort.

    Why it matters: The roundup pairs Anthropic's launch claims with Artificial Analysis's independent numbers, showing where Haiku 5.5 is cheap and strong and where token use offsets its price.

  25. Hacker News · AI (150+ points)BlogAI score58

    OpenAI withdraws three mathematical results

    AIOpenAI has withdrawn three of its mathematical results, according to a Hacker News post linking to a history file in OpenAI's math GitHub repository. The linked page is the only source here, and the feed supplied no further text describing the withdrawn results or the reasons for the withdrawal.

  26. Ant LingOfficialAI score36

    Ant Ling's Ling-3.1-Flash launches on Novita and OpenRouter

    AILing-3.1-Flash from Ant Ling is now available via Novita on OpenRouter, with Novita launching as a Day-0 partner. The model has 560B total parameters, 25B activated, and is built for hybrid reasoning and tool-using workflows. Novita offers it free until October 13 at 9:00 AM PT.

  27. meng shaoXAI score39

    Claude Haiku 5.5 tops GPT-6 Luna on benchmarks, with 2x faster token output

    AIAnthropic's Claude Haiku 5.5, released alongside Claude Opus 5.5 and Claude Sonnet 5.5, is reported to lead GPT-6 Luna across benchmarks, with OpenRouter measuring roughly twice the token output speed. Anthropic says Haiku 5.5 is its cheapest, fastest, and most capable small model, costing about 75% less to run than Claude Haiku 4.5 on average. The post also notes some CodeX users are reportedly migrating to Claude Code.

  28. IThome · AINewsAI score40

    Microsoft Confirms Copilot+ PC Brand Lives On, Runs 2 Trillion Local AI Inferences Monthly

    AIMicrosoft Windows and devices head Pavan Davuluri confirmed the Copilot+ PC brand has not been discontinued, saying more than 40% of commercial laptops are Copilot+ PCs shipping in tens of millions annually. He said these devices run over 2 trillion local inferences per month across search, image processing, and video calls. Microsoft plans to strengthen them through hybrid intelligence with local context, local actions, and local models.

  29. howie.seriousXAI score22

    Grok bot's X rate limit is 1,000 calls per day

    AIThe Grok bot's rate limit on X is 1,000 calls per day, which the author finds more than sufficient. Previously, using X's official API cost $20 per top-up and was spent quickly, while now it can be used for free.

    Image from @howie_serious's post
  30. Meta NewsroomOfficialAI score36

    Meta Donates 1,000 Ray-Ban Meta AI Glasses to Singapore Disability Groups

    AIMeta is donating 1,000 Ray-Ban Meta AI glasses to four Singapore organisations serving people with disabilities, alongside a US$30,000 grant for accessibility training. The glasses help users who are blind or have low vision read text, identify objects and describe their surroundings. The grant will fund a free curriculum from the Singapore Association of the Visually Handicapped on using the glasses safely in daily life.

  31. MarkTechPostNewsAI score65

    Perplexity releases pplx-embed-v2-late, a 0.6B edge model and 9B model

    AIPerplexity has released pplx-embed-v2-late, a pair of ColBERT-style multimodal embedding models in 0.6B and 9B sizes that retrieve text, images and rendered PDF pages in a shared embedding space. Both are available on Hugging Face under the MIT license, while a hosted API endpoint is planned but not yet live.

  32. howie.seriousXAI score22

    Howie Serious says Grok bot can gather X information for users

    AIHowie Serious (@howie_serious) says his Grok bot has found its first powerful use case: a personal agent that collects Twitter information. He argues this beats humans scrolling phones, manually gathering posts, or getting absorbed in endless feeds.

    Image from @howie_serious's post
  33. Claude Code · GitHub ReleasesOfficialAI score22

    Claude Code v2.1.294 fixes prompt and agent hook judgment

    AIClaude Code v2.1.294 fixes prompt and agent hooks written as instructions, which had allowed actions they should block. It also improves how prompt hooks on Stop and SubagentStop are judged, making Claude less likely to stop early.