Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

Oct 9Fri
  1. MarkTechPostNewsAI score67

    Google Cloud launches Gemini agent, a single cloud-hosted agent for enterprise work

    AIGoogle Cloud has introduced the Gemini agent, a single cloud-hosted agent that handles Q&A, knowledge work, media creation, and coding from one prompt box and one API. It routes jobs across Gemini and Claude models today, with other private and open models planned. Governance covers per-agent identity, role-based access, audit logging, and hard per-project spend caps, but the source gives no reproducible benchmarks, pricing, or general availability date.

  2. MarkTechPostNewsAI score44

    Underdog Releases Saluki 27B, a 2-Bit Qwen3.8-27B That Beats the Original at Tool Calling

    AIUnderdog has released Saluki 27B under Apache 2.0, a 2-bit GGUF of Qwen3.8-27B that fits in 7.89 GB, versus 54 GB for the full BF16 model. On Underdog Bench, Saluki scores 88 against 84 for the full model, and it raises parallel tool-call accuracy to 42 from 35. It runs on stock llama.cpp, but math and reasoning drop sharply, with AIME 2025 at 79.2 versus 96.7.

  3. The Guardian · AINewsAI score62

    OpenAI projects $50bn revenue, $20bn below its earlier investor signal

    AIOpenAI told investors it expects $50bn in revenue this year, about $20bn less than the $70bn it had signalled last month. The gap stems partly from comparing with Anthropic, which counts revenue sold through cloud partners such as AWS and Google Cloud, while OpenAI does not. The news weighed on US tech stocks, and OpenAI is in early talks to raise $30bn at a valuation of about $1.4tn.

  4. The DecoderNewsAI score61

    OpenAI bans Russian and Iranian influence ops that planted fake stories in real outlets

    AIOpenAI exposed a Russian and an Iranian influence operation and banned the ChatGPT accounts involved, both of which planted content in legitimate media using fake identities. The Iranian operation, "Bogus Bylines," used seven fake journalists to place nearly 100 articles about the US-Iran conflict, while the Russian "Dark Clark" operation triggered fact-checks and official denials in Ecuador and Peru. Both operations used AI mainly for internal reporting and adapting propaganda to different languages.

  5. Tencent · new models on Hugging FaceOfficialAI score41

    Tencent Releases Youtu-Parsing-Omni, a 5B Omni-Modal Document and Media Parsing Model

    AITencent has open-sourced Youtu-Parsing-Omni, a 5B-parameter omni-modal model that outputs a single structured JSON covering layout, text, tables, formulas, ASR, OCR, and video segments. It scores 96.96 Overall on OmniDocBench, the highest among the compared models, and ships with weights on Hugging Face, a vLLM plugin, and inference examples.

  6. NVIDIA · new models on Hugging FaceOfficialAI score16

    NVIDIA releases Agile One S SSD Pick GR00T N1.7 checkpoint 40000 model on Hugging Face

    AINVIDIA published the Agile One S SSD Pick deployment model, GR00T N1.7 checkpoint 40000, on Hugging Face for SSD pickup tasks. The repository includes five ONNX graphs with external tensor files and two existing TensorRT BF16 engines, with original configurations and build metadata, but no retraining or re-export was performed. The files are not a robot deployment or safety qualification, and engine compatibility depends on the target GPU and TensorRT environment.

  7. NVIDIA · new models on Hugging FaceOfficialAI score23

    NVIDIA publishes Agile One S SSD pick model, GR00T N1.7 checkpoint 58000, on Hugging Face

    AINVIDIA has released a deployment model for Agile One S SSD pickup, based on GR00T N1.7 checkpoint 58000 and using three cameras: ego, left wrist, and right wrist. The repository republishes ONNX graphs, external tensor files, and two existing TensorRT BF16 engines without retraining or re-export, and the original export reported a numerical warning that full FP32, node, and BF16 parity did not pass all tolerances. The files are not a certified robot deployment or safety qualification.

  8. NVIDIA · new models on Hugging FaceOfficialAI score25

    NVIDIA releases Agile One S Walk GR00T N2 checkpoint 1680 on Hugging Face

    AINVIDIA published the Agile One S Walk GR00T N2 checkpoint 1680, a walking deployment model with four cameras, on Hugging Face. The repository includes ONNX graphs, TensorRT BF16 plans/engines, and the original checkpoint files, republished without retraining or re-export. The shared Cosmos-Reason1-7B dependency and the Isaac/GR00T runtime must be set up separately, and the files are not a robot safety qualification.

  9. NVIDIA · new models on Hugging FaceOfficialAI score14

    NVIDIA Releases Agile One S SSD Place GR00T N1.7 Deployment Model on Hugging Face

    AINVIDIA published the nvidia/agile_one_s_place_ssd_n17_24050 repository on Hugging Face, containing a GR00T N1.7 checkpoint 24050 model for placing an SSD with three cameras. The repository includes five ONNX graphs with external tensor files and two existing TensorRT BF16 engines, republished without retraining, re-export, or engine rebuild. Engine compatibility depends on the target GPU and TensorRT environment, and the files are not a robot deployment or safety qualification.

  10. meng shaoXAI score45

    Addy Osmani on why engineers' joy in AI coding agents splits three ways

    AIAddy Osmani argues engineers' reactions to AI coding agents depend on which of three joys they value most: making, knowing, or mattering. He warns that choosing among agent suggestions without generating ideas yourself erodes the skill of ideation and can leave developers directed by agents. He reframes grief over lost craft as a sign of real attachment rather than failed adaptation.

    Image from @shao__meng's post
  11. Harrison ChaseXAI score22

    Harrison Chase on eval-driven development for AI agents

    AIHarrison Chase's post is titled "eval driven development," presenting evals as a development approach. The main post gives no further detail beyond the title. The quoted context from Jerry Liu argues that most tasks can be solved by defining an eval and hillclimbing over it rather than hand-building a deterministic or agentic workflow.

  12. Harrison ChaseXAI score22

    Harrison Chase questions eval-driven development for autonomous agents

    AIHarrison Chase argues that eval-driven development works for narrowly scoped tasks but breaks down for more autonomous agents, invoking Goodhart's Law that a measure ceases to be useful once it becomes a target. He asks how such agents can be hill-climbed, and the post does not provide an answer.

  13. 🚨 AI News | TestingCatalogXAI score50

    OpenAI, Anthropic, Google, and others roll out agent and model updates

    AIOpenAI's GPT-6.1 Sol Ultrafast is rolling out in the API, Codex, and ChatGPT Work, running up to 8x faster than Sol Standard at $12/$60 per million tokens. StepFun's Step 5 Preview, a 600B-parameter MoE model with 27B active parameters, is now on OpenRouter, and JetBrains released the open 12B MoE coding model Mellum2.1 under Apache 2.0.

  14. MagnificOfficialAI score14

    Magnific One turns profile photos into cereal-themed images

    AIMagnific announces that its Magnific One tool can now be used to create cereal-themed profile pictures from users' photos. The company invites users to upload a photo through its Flow tool to join the promotion.

    Video from @magnific's post
  15. QbitAINewsAI score62

    Google's AMIE Chatbot Tested in Real Pre-Visit Clinical Study Published in The Lancet

    AIA study led by Google and BIDMC tested Google's diagnostic AI chatbot AMIE with 98 outpatients before emergency visits, with a supervising doctor monitoring every exchange. No conversation needed interruption under the predefined safety criteria, and clinicians said AI summaries helped them prepare for 75% of visits. AMIE's differential diagnoses matched final diagnoses 90% of the time, but the authors say larger trials are needed.

  16. IThome · AINewsAI score46

    JetBrains Releases Mellum2.1 Coding Model With Near-Double Qwen3.5-9B Throughput

    AIJetBrains released Mellum2.1, a 12B mixture-of-experts coding model with 2.5B active parameters under Apache 2.0, emphasizing agentic programming. Under high load, its inference throughput in tokens is nearly twice that of Qwen3.5-9B in JetBrains' comparison, and multi-token prediction (MTP) speeds single-request responses by about 1.6x. The model is available on Hugging Face for local or private-infrastructure deployment, with GGUF and vLLM MTP support announced for later.

  17. IThome · AINewsAI score55

    Odyssey-3 world model scores 66.1 on Physics-IQ Verified benchmark

    AIOdyssey announced the Odyssey-3 series of foundation world models, with Odyssey-3 Pro scoring 66.1 on the Physics-IQ Verified video-to-video benchmark, the highest recorded on that leaderboard. The series includes a standard version balancing physical accuracy and generation cost, and a Pro version with stronger physics prediction. The preview supports first-person and third-person navigation and lets users move the camera, take actions, or trigger events while the model predicts environmental changes in real time.

  18. PandailyNewsAI score60

    Richard Yu says more Huawei phones will get LogicFolding chips

    AIRichard Yu said more Huawei phones will adopt LogicFolding chips built under the Tau Scaling Law, though no models or timetable were given. He said the Kirin 9050 Pro's performance is 31% higher than its predecessor, and the source outlines a roadmap reaching 5.0 GHz by 2031.

  19. PandailyNewsAI score56

    openJiuwen open-sources an enterprise AgentOS for agent swarms and multi-tenant control

    AIHuawei-backed openJiuwen has open-sourced AgentOS for Enterprise under Apache 2.0 on GitHub and AtomGit, targeting multi-agent coordination, memory-based self-evolution, multi-tenant isolation and fault recovery. Huawei Connect 2026 also introduced an all-in-one appliance built on it, which the launch information says enables an end-to-end private deployment in hours.

  20. PandailyNewsAI score38

    KingKong Technology Open-Sources Jumper Crab Robot Software Stack

    AIKingKong Technology has open-sourced the software stack for Jumper, a six-legged crab-style robot it designed, including its MuJoCo model, simulation scenes, reinforcement learning training and deployment tooling. Jumper has 22 degrees of freedom, measures about 400 by 400 by 200 mm, weighs about 1.8 kg and lists a maximum jump height of 400 mm or more. The mechanical CAD files, bill of materials, PCB designs and electrical schematics are not public, and RKNN inference on the real board has not yet been validated.

  21. PandailyNewsAI score45

    Doubao Work Adds Infinite Creation Canvas, Seedream 5.0 Flash and Doubao 2.1 Lite

    AIByteDance's Doubao Work has added an infinite creation canvas that places source materials, design options and finished output on one page, wired to the new Seedream 5.0 Flash image model. The update also adds Doubao 2.1 Lite, a lighter model aimed at everyday office tasks such as documents, spreadsheets and slide decks, with faster responses and lower credit consumption. The announcement included no benchmark results for either model.

  22. Alexandr WangXAI score42

    Alexandr Wang marks Muse's first month with strong user response

    AIMeta's Alexandr Wang marked one month since Muse launched, saying its response has exceeded expectations and that people are using it to save money and time. He said the product has made a real difference for users including parents, grandparents, students, and coworkers. The quoted launch post describes Muse as an always-on personal AI assistant that can use a browser, connect to apps, and is designed to be secure.

    Image from @alexandr_wang's post
  23. Dongxi NLPXAI score5

    Visitor praises Google's employee perks after criticizing Gemini

    AIAfter visiting Google, the author says the company's employee care is extremely generous, from lunch to concern for body hair, with costs unconsidered. The post claims the Chinese office even serves whole soft-shell turtles at lunch, and says Google is a better workplace than companies that subject staff to PUA.

  24. South China Morning Post · TechNewsAI score34

    Chinese optical chip stocks extend rout on fears of possible US curbs

    AIShares of Chinese optical chipmakers fell for a second straight session on Friday amid fears of potential US trade curbs on next-generation optical transceivers for data centres. China's CSI 300 Index slid 1.3 per cent by midday to its lowest level since August last year, with upstream laser chip suppliers falling more steeply than downstream transceiver makers.

  25. The Guardian · AINewsAI score33

    Former Labour minister Tom Watson defends Palantir contracts, warns against "mob rule"

    AITom Watson, now a senior vice-president at Palantir, warned UK ministers against letting "mob rule" dictate public procurement, saying they could get "in a lot of trouble" if they do. The former Labour deputy leader's comments come as Prime Minister Andy Burnham faces pressure to drop Palantir from government contracts, including a £330m NHS software deal.

  26. CNBC · TechnologyNewsAI score44

    OpenAI defends firing three safety researchers, citing a breach of trust

    AIOpenAI defended its decision to fire three safety researchers, Jasmine Wang, Tomek Korbak and Mikita Balesni, saying they committed a "significant breach of trust." The company said the dismissals were not about the researchers raising safety concerns, though it agreed with the letter they sent to board members and safety committees about preserving the monitorability of frontier models.

  27. Jerry LiuXAI score26

    Jerry Liu says evals now replace hand-built agent workflows

    AIJerry Liu argues that most tasks can now be solved by defining an eval and hillclimbing on it, rather than hand-coding a deterministic or agentic workflow. He says data provider companies are building evals across economic activity so frontier models can handle more work, leaving developers to define goals and success measures. He expects agent interfaces to compress most tasks into goals and eval instructions, while the most complex processes will still need explicit workflow builders.

  28. Pulkit GargXAI score22

    AgentSDR launches as a free, open-source AI SDR workspace

    AIAgentSDR launches on Product Hunt as an open-source AI sales development tool under the MIT license. It combines functions that users would otherwise get from Apollo, Clay, Smartlead, Instantly, HeyReach, and a CRM in one workspace. The product uses bring-your-own API keys and charges no per-seat or per-contact fees.

    Video from @Pulkitgarg25's post
  29. MoeinXAI score10

    Kolig launches on Product Hunt as an office for AI teammates

    AIKolig is now live on Product Hunt, presented as an office for AI teammates that know a company and work together. The post's creator asks for support and honest feedback after a lot of design and rebuilding work.

  30. X.PINXAI score60

    Suspected Guangdong attacker reportedly used Claude Code, ARTEX, GLM and DeepSeek

    AIA suspected 26-year-old in Guangdong reportedly used Claude Code, ARTEX, GLM and DeepSeek in attacks. An AI-generated résumé named South China University of Technology, but the identity is unverified and the listed phone number's owner denied involvement. The suspect reportedly sought buyers on Telegram, but no sale was reported, and ARTEX creator Autumn condemned the misuse and said he would stop releasing the tool as open source.

    Image from @thexpin's post
  31. LeiphoneNewsAI score8

    Carbon-Silicon Dao Code Seventh Layer Sets Self-Audit Baseline and Falsification Terms

    AIThe seventh and final layer of the "Carbon-Silicon Dao Code" cross-domain migration governance framework sets a self-audit baseline, opens falsification terms, and defines the framework's applicability boundary. The article says it validates each layer's input-output consistency backward from layer seven to layer one, and it allows anyone who constructs a reproducible, traceable counterexample targeting NT1–NT4 to submit it for baseline review. It also archives the framework's documents and versions with hashes across Toutiao, Douyin, and GitHub.

  32. vLLMOfficialAI score42

    vLLM Semantic Router team releases Decision 2.0 multi-question classification models

    AIThe vLLM Semantic Router team has released Decision 2.0, which answers multiple questions about one input in a single forward pass and outputs per-option probabilities. The post presents this as useful for routing and classification. A quoted post from Xunzhuo Liu says Decision 2.0 includes six open decision models ranging from 0.6B to 27B parameters, each topping same-size open models on the Jev Decision Index 0.3.

  33. OpenAI NewsroomOfficialAI score45

    OpenAI fires three researchers over sensitive information breach, denies retaliation

    AIOpenAI says it parted ways with researchers Jasmine, Mikita, and Tomek after an internal investigation found they violated policies on handling sensitive information. The company says the decisions were not about raising safety concerns, which it says it encourages, and that it has not terminated any employee for raising concerns. OpenAI also says it is finalizing contracts with third-party safety assessors and will announce details in the coming weeks.