Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. OpenRouterOfficialAI score52

    StepFun's Step 5 Preview is now available on OpenRouter

    AIOpenRouter says StepFun's Step 5 Preview is now live on its platform, priced at $1.00 per million input tokens and $2.70 per million output tokens. Cache hits are 50% off at launch, bringing them to $0.05 per million, and a week of free access is rolling out across partners including opencode, Cline, Nous Research, and Kilo Code.

  2. OpenRouterOfficialAI score44

    StepFun's Step 5 Preview launches on OpenRouter as agentic flagship

    AIStepFun's Step 5 Preview is now live on OpenRouter as the company's new flagship for agentic work. It uses a sparse MoE design with 27B active and 600B total parameters, a 1M context window, and accepts text, image, and video input. The post highlights strength in coding and professional knowledge work, especially finance.

    Image from @OpenRouter's post
  3. OpenBMBOfficialAI score36

    ReJev fine-tunes MiniCPM5-2B to lift decision accuracy to 80.50%

    AIReJev, an independent community project, applied LoRA post-training to OpenBMB's MiniCPM5-2B for bounded agent decisions: state, question, and candidate options yield one choice. On its sealed 1,892-sample holdout, accuracy rose from 51.11% to 80.50% (+29.39 percentage points) with 0% invalid outputs, at about $5.31 in cumulative Modal billing including earlier experimental overhead. The authors describe this as an early, task-specific result, not parity with Jev.

    Image from @OpenBMB's post
  4. Gergely OroszXAI score26

    Developers working more with AI tools, citing more context switching

    AISoftware developer Gergely Orosz questions why he is working more despite AI tools, quoting Sam Newman's view that AI was meant to free developers from drudgery. Newman says most developers are doing more work, with more context switching and a loss of the big picture. The quoted post adds that AI assistants are not human partners and that pairing with them fragments the shared mental model of a program.

  5. QbitAINewsAI score47

    Vidu Q4 Preview Offers 4K Video Generation at About 0.09 Yuan per Second

    AIShengshu Technology has opened a preview of its Vidu Q4 video generation model, which supports native 4K output and up to 15 reference images and three reference audio clips. Testers generated a one-minute video for about 5.4 yuan, roughly 0.09 yuan per second at 720P, which the article says is a starting price that varies by resolution and mode. The Vidu Q4 preview is available through the Vidu platform, with the MaaS API priced at about 0.6 yuan per second for 720P image-to-video.

  6. The Robot ReportNewsAI score42

    AWS launches open-source Physical AI Toolchain combining its services with NVIDIA's stack

    AIAmazon Web Services launched an open-source Physical AI Toolchain that combines AWS services with NVIDIA's Physical AI software to cover data generation, model training, simulation, edge deployment, and continuous improvement for robots. AWS uses Amazon SageMaker for training and AWS IoT Greengrass for distributing models to edge devices, while NVIDIA contributes Isaac Sim, Isaac Lab, Isaac GR00T, and Cosmos. The toolchain is hardware-neutral and does not directly replace RoboMaker, which was shut down in 2025.

  7. Allie K. MillerXAI score38

    Low-leverage AI uses fail once everyone else adopts AI too

    AIAllie K. Miller argues that AI's value is low leverage if it depends on others not using AI, citing inbox triage and social commenting as examples that break at scale. She proposes a test: whether a use case still creates value when everyone adapts, which she frames as finding the Nash equilibrium of AI usage.

  8. Karl's AI WattsXAI score24

    Tencent adds file preview and AI editing to WorkBuddy

    AITencent has added a new feature to its WorkBuddy product that lets users open files in different formats for preview and edit them with AI chat. For example, a Markdown file can be translated by AI and the translation can overwrite the original content directly.

    Image from @aiwarts's post
  9. Understanding AI (Timothy B. Lee)BlogAI score67

    TypeSafe AI's Jev returns probabilities over fixed answers instead of text

    AITypeSafe AI released Jev, a model that answers yes/no, multiple-choice, or rating questions by outputting the estimated probability of each option. The author notes this design lets the model be served faster and more cheaply than LLMs and fits ordinary if-statement logic, and says he used it to flag spam comments on his blog in place of Gemini 3 Flash.

    Why it matters: The article explains why Jev's fixed-answer design, with probability outputs, is faster and cheaper than LLMs for classification, and shows its use in a real spam-filter setup.

  10. StepFunOfficialAI score18

    StepFun invites readers to explore its Step 5 Preview model

    AIStepFun is promoting Step 5 Preview, pointing readers to a model overview, benchmarks, and demos on its website. The post also links to technical specifications and API documentation for developers.

  11. StepFunOfficialAI score60

    StepFun's Step 5 Preview is live on OpenRouter with a week of free access

    AIStepFun says Step 5 Preview is now available on OpenRouter, with a week of free access rolling out across OpenCode, Cline, Nous Research, Kilo Code, and other tools. The company describes it as flagship-tier intelligence for agentic and professional work at substantially lower task cost, letting users switch models without changing their workflow.

    Image from @StepFun_ai's post
  12. The Wall Street Journal · TechNewsAI score10

    Quantum Research Races Ahead as Venture Dealmaking Soars

    AIQuantum research is advancing rapidly, according to the original headline, while venture dealmaking has soared even as exits remain stalled. The source text provides no further figures, named companies, or technical details.

  13. The DecoderNewsAI score72

    One public AI agent on AWS could take over every other agent in its region

    AIZenity Labs says a single publicly accessible agent on Amazon Bedrock AgentCore could take over all AgentCore agents in the same AWS account and region. A chat prompt let the researchers query the instance metadata service and steal temporary credentials, and AgentCore's default permissions allowed read, write, and delete access across agents. According to Zenity, AWS made IMDSv2 the default for new deployments and changed the default execution role around August.

    Why it matters: The report traces how one public agent's weak isolation exposed credentials and every other agent in the region, showing why default permissions matter for enterprise deployments.

  14. NVIDIA BlogOfficialAI score34

    Gears of War: E-Day Launches on GeForce NOW With RTX-Powered Cloud Streaming

    AINVIDIA's GeForce NOW now streams Gears of War: E-Day, released globally on October 6, with Ultimate members getting GeForce RTX 5080-class performance plus NVIDIA DLSS and NVIDIA Reflex. Fire TV users will soon be able to buy GeForce NOW memberships directly through Amazon, with availability expected in the coming weeks. The cloud library also adds several new releases this week, including STAR WARS: Galactic Racer and Clive Barker's Hellraiser: Revival.

  15. Databricks BlogOfficialAI score35

    How to build governed enterprise apps on Databricks with Replit and Lakebase

    AIReplit and Databricks integration, now generally available with native Lakebase support, lets enterprise teams build apps from plain-language prompts using Replit Agent and deploy them as Databricks Apps. Deployed apps inherit automatic user authentication and Unity Catalog access controls, and Replit Agent auto-provisions a managed Lakebase Postgres database for operational data. Lakebase keeps app-written data inside the Databricks perimeter instead of a separate external database.

  16. Ethan MollickXAI score14

    Mollick argues organizations are narrow superintelligence that AI must integrate with

    AIEthan Mollick argues that organizations such as universities and Walmart already act as narrow superintelligences, doing things no single human can through complex processes no one explicitly designed. He contends that failing to design how AI works alongside these existing organizational systems is a major reason AI's high capability has not yet produced large gains in scientific discovery or economic productivity.

  17. PyTorch BlogOfficialAI score46

    IBM details how Spyre becomes a native PyTorch device

    AIIBM's torch-spyre team connects PyTorch's device, allocator, stream, and event abstractions to the Spyre runtime and firmware, so tensors stay resident on device="spyre". Spyre is IBM's dataflow AI accelerator for inference, with 32 cores, 2 MB of scratchpad per core, and up to 128 GB of LPDDR5. PyTorch's PrivateUse1 backend gives Spyre its own device identity, and FX graphs stay in the Inductor compiler path.

  18. QbitAINewsAI score34

    Physical AI firm Zhengxing Innovation unveils retail 24/7 human-robot collaboration solution

    AIZhengxing Innovation launched a Physical AI solution at APRCE 2026 for retail human-robot collaboration, built on its "embodied brain" and comprising the H1 humanoid and C1 wheeled-arm robots plus the M1 management platform. The company says the solution needs no store renovation, reports 99% autonomous task completion, and plans commercial service in 2027 via direct purchase or RaaS subscription.

  19. OpenRouter · New modelsBlogAI score54

    StepFun releases Step 5 Preview, a 600B-parameter agentic model

    AIStepFun has released Step 5 Preview, its flagship model for agentic work, built on a sparse Mixture-of-Experts architecture with 27B active and 600B total parameters. The source says it performs strongly in software engineering and professional tasks, but the feed supplied only an excerpt, so benchmark details are not available here.

  20. HeyGenXAI score22

    Ryan Serhant launches daily AI avatar video series with HeyGen

    AIReal estate figure Ryan Serhant is launching a daily video series on sales, business, and personal branding, delivered by his official AI avatar rather than filmed by him. The launch is announced in a post linking to a Variety report, and the avatar is produced with HeyGen.

    Video from @HeyGen's post
  21. SiliconANGLE · AINewsAI score62

    Google Cloud launches Gemini agent for enterprise work across devices and apps

    AIGoogle Cloud introduced Gemini agent, a unified AI assistant that acts autonomously, generates code, and completes work across web, mobile, desktop, and third-party apps. It runs jobs on models matched to each task, including Gemini Flash and a flagship frontier model, with Anthropic Claude models also available. Hard spend limits per project let companies enforce budgets and charge AI costs to departments.

  22. ElevenLabs BlogOfficialAI score39

    ElevenReader Launches in Brazil With Fábio Porchat Narration and 70,000 Portuguese Books

    AIElevenReader, ElevenLabs' consumer audio platform, launches in Brazil with narration by actor and comedian Fábio Porchat. Through a partnership with Bookwire Brasil, the app offers 70,000 licensed Brazilian Portuguese titles, most of which have no audio edition. The app is free on iOS and Android, with the full catalog available through ElevenReader Ultra.

  23. JetBrains AI BlogOfficialAI score62

    JetBrains releases Mellum2.1, an open coding model trained with reinforcement learning

    AIJetBrains released Mellum2.1, a 12B mixture-of-experts model with 2.5B active parameters under the Apache 2.0 license, built for coding agents. Post-training shifted to reinforcement learning across thousands of environments and millions of sandboxed runs, and the model is available on Hugging Face. The source reports gains over Mellum2 on LiveCodeBench, AIME, GPQA Diamond, BFCL v4, IFEval, and SWE-bench Verified, and says it serves almost twice the tokens of Qwen3.5-9B under heavy load.

    Why it matters: The post shows how reinforcement learning in real sandboxed environments changed a compact open model's repository work, with benchmark gains against Mellum2 and two peers.

  24. Google Cloud · AI & Machine LearningOfficialAI score19

    Irish Brands Scale AI Operations with Gemini Enterprise, from Ryanair to Startups

    AIIrish organizations including Ryanair, Smyths Toys, the Irish Revenue Commissioners, Virgin Media Ireland and startups Hexis, Kitman Labs, Spryt and IMPT are moving agentic AI from prototypes to production using Gemini Enterprise and Google Cloud. Ryanair is deploying Gemini Enterprise and Google Workspace for about 35,000 employees, while Smyths Toys' AI agent Codie has resolved more than 60% of web inquiries. The article also states that Google operations added an estimated 10 billion euros to Irish GDP in 2025.

  25. Google Cloud · AI & Machine LearningOfficialAI score78

    Google Cloud launches Gemini agent as single universal work agent

    AIGoogle Cloud announced the Gemini agent, a single agent that answers questions, handles knowledge work, creates media, and writes and runs code from one prompt box. It runs in the cloud with persistent memory, uses multi-agent orchestration, and adds Workspace integration, domain skills for data and industries, identity-based governance through Agent Gateway, and spend caps. The source also cites customer deployments and says nearly 80% of Google Cloud customers use its AI products.

    Why it matters: The announcement shows how a single work agent spans chat, Workspace, data analysis, governance, and cost controls, useful for judging enterprise agent deployment scope.

  26. ElevenLabs BlogOfficialAI score26

    How to build a meeting transcription API with Scribe v2 and Scribe v2 Realtime

    AIElevenLabs explains how to build meeting transcription products using its Scribe v2 and Scribe v2 Realtime models through its API. Real-time transcription suits live captions and in-meeting bots, while batch transcription suits post-meeting notes and records, with Scribe v2 Realtime reporting 150 ms latency and supporting up to 50 key terms for prompting.