Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 8

Oct 8Thu
  1. Testing CatalogAI score46

    Google announces a unified Gemini agent for Gemini Enterprise work

    AIGoogle has announced a single, universal Gemini agent for Gemini Enterprise as part of its Gemini at Work updates. The agent answers questions, handles knowledge work, creates images and media, and writes and runs code. It works inline in Gmail, Drive, Docs, Slides, Sheets, Chat, and Calendar, with new data and analytics skills for plain-language insights and industry-specific tools for financial services and legal teams.

  2. SantiagoAI score38

    Teamily AI lets people and agents share one group chat context

    AITeamily AI now lets users add people and AI agents to the same group chat, so everyone works from shared context. The example shows a branding change handled by research, writing, and website-building agents, with a designer's feedback incorporated and the finished page shared in one continuous conversation. The platform's 2.0 release, which the post describes as opening to everyone, adds real-time human–agent collaboration and multi-model routing.

  3. a16z NewsAI score45

    CFOs Are Becoming Builders as AI Reshapes Finance Operations

    AIAI-native tools are removing the data bottleneck that long constrained CFOs, shifting the role toward designing the operating systems that turn data into decisions. Finance teams are adopting AI-native software for ERP, forecasting, procurement, and audit, and "finance engineers" are building custom automations and agents. OpenAI's CFO Sarah Friar describes finance moving toward a zero-day close and continuously updated forecasts.

  4. TransformerAI score53

    Yoshua Bengio urges AI researchers to leave frontier labs for safety work

    AIYoshua Bengio, co-president of LawZero, asks researchers at frontier AI companies to reconsider whether they should keep working there, arguing that safety efforts are not slowing a dangerous race. He cites the recent UN Security Council briefing on AI incidents and says he left his earlier research path after ChatGPT made the risks feel immediate. He urges researchers to join AI Safety Institutes or mission-driven organizations such as LawZero.

  5. GuizangAI score22

    Grok bot builds and publishes daily AI news videos on a foldable phone

    AIGuizang (@op7418) says a Grok bot paired with a foldable phone lets him chat on one screen while the bot publishes content on the other. Background post: he set Grok to produce a daily morning AI news video on a schedule, running content collection, code writing, and video rendering entirely on Grok's cloud virtual machine rather than his local computer. He shares the full prompt so others can run the same workflow with their own Grok bot.

  6. The Guardian · AIAI score42

    One Nation's AI-generated campaign video draws criticism over racist tropes and regulatory gaps

    AIOne Nation's AI-generated campaign video, reportedly played at its Victorian campaign launch, depicts racist stereotypes including a man brandishing a machete and a man in an explosive vest. The Australian Communications and Media Authority cannot act against it because its powers do not cover this content, and the federal Labor government has not yet moved to ban AI-generated content in election periods.

  7. South China Morning Post · TechAI score60

    Can China match Meta's Muse in the race to harness AI agents?

    AIMeta's Muse personal AI agent, launched on September 8, surpassed 2.5 million downloads in its first two weeks and topped free-app rankings on Apple and Google's US app stores. Its popularity has drawn attention to the emerging market for agent harnesses, where China's biggest internet companies are already competing for position. The excerpt does not provide full details on their specific products.

  8. The Verge · AIAI score41

    Meta's Muse and OpenAI's Dots: can consumers trust AI agents with their lives?

    AIMeta's Muse and OpenAI's Dots are always-on AI agents with animated mascots, pitched to consumers and businesses for tasks like restaurant reservations and inbox triage. Muse is free, while Dots is not, and OpenAI also offers "specialist" Dots for marketing, legal work, and accounting. The discussion centers on privacy and security concerns about giving agents access to credit card details and email.

  9. The DecoderAI score46

    Ethereum researchers warn AI math advances could threaten crypto wallet signatures

    AIEthereum researcher Justin Drake warned on X that AI-assisted math could, in the worst case, break the signature system used by crypto wallets within months, and urged a "bunker mode" in which users move funds to addresses that have never signed a transaction. Vitalik Buterin agreed but cautioned against moving too fast, saying he has lost more money to botched migrations than to hacks. No one has yet broken the current ECDSA signature scheme in practice.

  10. OpenRouterAI score62

    StepFun's Step 5 Preview model is now available on OpenRouter at $1.00 per million input tokens

    AIOpenRouter announces that StepFun's Step 5 Preview is live on its platform, priced at $1.00 per million input tokens and $2.70 per million output tokens. Cache hits are 50% off at launch, bringing them to $0.05 per million. A week of free access is rolling out across partners including opencode, Cline, Nous Research, and Kilo Code.

  11. OpenBMBAI score36

    ReJev fine-tunes MiniCPM5-2B to lift decision accuracy to 80.50%

    AIReJev, an independent community project, applied LoRA post-training to OpenBMB's MiniCPM5-2B for bounded agent decisions: state, question, and candidate options yield one choice. On its sealed 1,892-sample holdout, accuracy rose from 51.11% to 80.50% (+29.39 percentage points) with 0% invalid outputs, at about $5.31 in cumulative Modal billing including earlier experimental overhead. The authors describe this as an early, task-specific result, not parity with Jev.

  12. Gergely OroszAI score26

    Developers working more with AI tools, citing more context switching

    AISoftware developer Gergely Orosz questions why he is working more despite AI tools, quoting Sam Newman's view that AI was meant to free developers from drudgery. Newman says most developers are doing more work, with more context switching and a loss of the big picture. The quoted post adds that AI assistants are not human partners and that pairing with them fragments the shared mental model of a program.

  13. QbitAIAI score47

    Vidu Q4 Preview Offers 4K Video Generation at About 0.09 Yuan per Second

    AIShengshu Technology has opened a preview of its Vidu Q4 video generation model, which supports native 4K output and up to 15 reference images and three reference audio clips. Testers generated a one-minute video for about 5.4 yuan, roughly 0.09 yuan per second at 720P, which the article says is a starting price that varies by resolution and mode. The Vidu Q4 preview is available through the Vidu platform, with the MaaS API priced at about 0.6 yuan per second for 720P image-to-video.

  14. The Robot ReportAI score42

    AWS launches open-source Physical AI Toolchain combining its services with NVIDIA's stack

    AIAmazon Web Services launched an open-source Physical AI Toolchain that combines AWS services with NVIDIA's Physical AI software to cover data generation, model training, simulation, edge deployment, and continuous improvement for robots. AWS uses Amazon SageMaker for training and AWS IoT Greengrass for distributing models to edge devices, while NVIDIA contributes Isaac Sim, Isaac Lab, Isaac GR00T, and Cosmos. The toolchain is hardware-neutral and does not directly replace RoboMaker, which was shut down in 2025.

  15. Allie K. MillerAI score38

    Low-leverage AI uses fail once everyone else adopts AI too

    AIAllie K. Miller argues that AI's value is low leverage if it depends on others not using AI, citing inbox triage and social commenting as examples that break at scale. She proposes a test: whether a use case still creates value when everyone adapts, which she frames as finding the Nash equilibrium of AI usage.

  16. Understanding AI (Timothy B. Lee)AI score67

    TypeSafe AI's Jev returns probabilities over fixed answers instead of text

    AITypeSafe AI released Jev, a model that answers yes/no, multiple-choice, or rating questions by outputting the estimated probability of each option. The author notes this design lets the model be served faster and more cheaply than LLMs and fits ordinary if-statement logic, and says he used it to flag spam comments on his blog in place of Gemini 3 Flash.

  17. The DecoderAI score75

    Zenity finds one public AWS AgentCore agent could hijack all agents in its region

    AIZenity Labs reported that a chat prompt to one publicly accessible agent on Amazon Bedrock AgentCore could expose credentials and take over every AgentCore agent in the same AWS account and region. The researchers said the agent queried the internal metadata service and sent its temporary credentials to an external server, while AgentCore's default permissions applied across all agents. AWS reportedly made IMDSv2 the default for new deployments and changed the default execution role around August.

  18. NVIDIA BlogAI score34

    Gears of War: E-Day Launches on GeForce NOW With RTX-Powered Cloud Streaming

    AINVIDIA's GeForce NOW now streams Gears of War: E-Day, released globally on October 6, with Ultimate members getting GeForce RTX 5080-class performance plus NVIDIA DLSS and NVIDIA Reflex. Fire TV users will soon be able to buy GeForce NOW memberships directly through Amazon, with availability expected in the coming weeks. The cloud library also adds several new releases this week, including STAR WARS: Galactic Racer and Clive Barker's Hellraiser: Revival.

  19. Databricks BlogAI score35

    How to build governed enterprise apps on Databricks with Replit and Lakebase

    AIReplit and Databricks integration, now generally available with native Lakebase support, lets enterprise teams build apps from plain-language prompts using Replit Agent and deploy them as Databricks Apps. Deployed apps inherit automatic user authentication and Unity Catalog access controls, and Replit Agent auto-provisions a managed Lakebase Postgres database for operational data. Lakebase keeps app-written data inside the Databricks perimeter instead of a separate external database.

  20. PyTorch BlogAI score46

    IBM Builds Spyre as a Native PyTorch Device via torch-spyre

    AIIBM's torch-spyre integration makes Spyre, its dataflow inference accelerator, a native PyTorch device by mapping PyTorch's device, allocator, stream, and event abstractions onto the Spyre runtime and firmware. Tensors stay resident on device="spyre" between operations, and FX graphs remain in the Inductor compiler path. The approach gives eager and compiled execution one path with lower launch overhead.

  21. QbitAIAI score34

    Physical AI firm Zhengxing Innovation unveils retail 24/7 human-robot collaboration solution

    AIZhengxing Innovation launched a Physical AI solution at APRCE 2026 for retail human-robot collaboration, built on its "embodied brain" and comprising the H1 humanoid and C1 wheeled-arm robots plus the M1 management platform. The company says the solution needs no store renovation, reports 99% autonomous task completion, and plans commercial service in 2027 via direct purchase or RaaS subscription.

  22. OpenRouter · New modelsAI score54

    StepFun releases Step 5 Preview, a 600B-parameter agentic model

    AIStepFun has released Step 5 Preview, its flagship model for agentic work, built on a sparse Mixture-of-Experts architecture with 27B active and 600B total parameters. The source says it performs strongly in software engineering and professional tasks, but the feed supplied only an excerpt, so benchmark details are not available here.

  23. SiliconANGLE · AIAI score62

    Google Cloud launches Gemini agent for enterprise work across devices and apps

    AIGoogle Cloud introduced Gemini agent, a unified AI assistant that acts autonomously, generates code, and completes work across web, mobile, desktop, and third-party apps. It runs jobs on models matched to each task, including Gemini Flash and a flagship frontier model, with Anthropic Claude models also available. Hard spend limits per project let companies enforce budgets and charge AI costs to departments.

  24. ZDNet · AIAI score24

    Google Maps adds Ask Maps food ordering via Square and Uber Eats, plus fan-favorite dining list

    AIGoogle Maps now lets users place restaurant orders through its Gemini-powered Ask Maps chat tool, with Square and Uber Eats joining Toast as partners. Google also released its first fan-favorite dining list, which tracks trending food and drink interest across 10 cities, including a 216% rise in cheeseburger interest in Tokyo over the past year.

  25. ElevenLabs BlogAI score39

    ElevenReader Launches in Brazil With Fábio Porchat Narration and 70,000 Portuguese Books

    AIElevenReader, ElevenLabs' consumer audio platform, launches in Brazil with narration by actor and comedian Fábio Porchat. Through a partnership with Bookwire Brasil, the app offers 70,000 licensed Brazilian Portuguese titles, most of which have no audio edition. The app is free on iOS and Android, with the full catalog available through ElevenReader Ultra.

  26. JetBrains AI BlogAI score62

    JetBrains releases Mellum2.1, an open coding model trained with reinforcement learning

    AIJetBrains released Mellum2.1, a 12B mixture-of-experts model with 2.5B active parameters under the Apache 2.0 license, built for coding agents. Post-training shifted to reinforcement learning across thousands of environments and millions of sandboxed runs, and the model is available on Hugging Face. The source reports gains over Mellum2 on LiveCodeBench, AIME, GPQA Diamond, BFCL v4, IFEval, and SWE-bench Verified, and says it serves almost twice the tokens of Qwen3.5-9B under heavy load.

    Why it matters: The post shows how reinforcement learning in real sandboxed environments changed a compact open model's repository work, with benchmark gains against Mellum2 and two peers.