Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. Zhihao JiaXAI score62

    Lithos AI open-sources lithos-metal for fast local inference on Apple M5 Max

    AILithos AI says it is open-sourcing lithos-metal, which uses megakernels and DSpark speculative decoding. The post claims Qwen3.8-27B reaches a peak of over 200 tokens per second per user on a single Apple M5 Max. It says users can try the tool with any coding agent in one command, and links to the code on GitHub and a technical blog.

    Video from @JiaZhihao's post
  2. ClineOfficialAI score13

    Cline launches a desktop app alongside its CLI

    AICline announces a new Desktop app that users can try alongside its CLI, which installs via npm with the command npm i -g cline. The post links to the Desktop app at cline.bot/desktop.

  3. ClineOfficialAI score46

    Cline makes Step 5 Preview free, citing strong DeepSWE coding scores

    AICline says Step 5 Preview is now free in its coding tool and scores ahead of Kimi K3 and GLM-5.3 on DeepSWE. The company describes it as one of the strongest open-weights coding models available. StepFun's background announcement describes Step 5 Preview as a 600B total / 27B active MoE model with 1M context and vision, and says open weights arrive on Oct 15.

    Image from @cline's post
  4. Sierra BlogOfficialAI score62

    Sierra launches fleming-1 to detect AI agents calling by phone

    AISierra has launched fleming-1, a model that analyzes caller speech in real time and scores audio for signs it was generated by AI. It flags likely AI callers while keeping real people unflagged by default, and companies decide how to handle those calls. The model works with any voice agent built on Sierra, and Sierra also announced Personal Agent Protocol, an open standard for authorized agent-to-business interactions.

    Why it matters: The post explains why companies need to know when a caller is an AI agent, which frames the detection model as a business decision rather than an automatic block.

  5. Andrew CurranXAI score13

    Andrew Curran Posts "The saga continues" Amid Tightened κ Result

    AIAndrew Curran posted a brief "The saga continues" update, with no clear publisher or model identified. Quoted context from @0xdoug reports a validated, merged PR that tightened κ from 2⁻¹⁸² to 2⁻¹⁵, described as a 500-thousand-fold improvement over the previous result and a 2^167-fold improvement over the original OpenAI result. The quoted post credits a community effort and says results are being verified and published.

    Image from @AndrewCurran_'s post
  6. Google Cloud TechOfficialAI score10

    Google Cloud livestreams its #GeminiAtWork event now

    AIGoogle Cloud Tech announced it is live streaming the #GeminiAtWork event. The post provides a link to tune in but includes no details about product announcements or specific features.

  7. MuseOfficialAI score20

    Muse builds a personalized movie feed and books tickets for users

    AIMuse, created by @JJEnglert, is trained on the user's movie and TV preferences and weekly curates a feed of films in theaters and on owned streaming services. Users can upvote or downvote picks to refine the feed, and when they want to see a movie, Muse books the ticket through a wallet link integration.

    Image from @Muse's post
  8. Gizmodo · AINewsAI score13

    Trump Declares Anyone Using "Artificial Intelligence" Term "THE ENEMY" on Truth Social

    AIPresident Donald Trump posted on Truth Social that the White House considers anyone who uses the term "Artificial Intelligence," instead of "Super Intelligence," to be "THE ENEMY." The article reports that Nvidia's Jensen Huang, Meta's Mark Zuckerberg, Elon Musk, and Jeff Bezos have adopted the "super intelligence" term, and that ai.gov was changed to read "SI" while si.gov appears dormant.

  9. RahulXAI score25

    Teamily AI lets solo founder's agents research, write, build, and review a site

    AIA solo founder used Teamily AI, a platform where humans and AI agents share one group chat, to automate a multi-step project. A research agent analyzed the author's 20 most-saved posts, a writer agent drafted content, a web agent built a live website in Website Builder, and a custom Verification Editor agent blocked the launch twice over a misquoted Anthropic doc. The team can be saved as a Loop that reruns weekly and waits for human approval before publishing.

  10. 🚨 AI News | TestingCatalogXAI score49

    Odyssey launches Odyssey-3 world model with public research preview

    AIOdyssey has launched Odyssey-3, its most powerful foundation world model, with a public research preview. Odyssey-3 Pro scored 66.1 on Physics-IQ Verified video-to-video with best-of-8 sampling, the highest reported result. The model generates environments from prompts and predicts changes in real time as users move through scenes.

    Image from @testingcatalog's post
  11. elvisXAI score42

    Odyssey-3 Pro tops Physics-IQ Verified and shows robotic error recovery.

    AIOdyssey released Odyssey-3 Pro, which sets a new top score on Physics-IQ Verified, a benchmark where models continue videos of real physics experiments. In robotics, a robot arm with tens of hours of demonstrations recovered from a missed grasp, a behavior absent from those demos.

    Video from @omarsar0's post
  12. Luke EdwardsXAI score38

    Pocketty brings SSH and herdr agent alerts to iPhone and iPad

    AIPocketty launches as an SSH app for iPhone and iPad, built for herdr, that notifies users when an agent on any host is blocked. Tapping a notification opens the exact pane, and the app supports Tailscale and Bonjour natively, shows diffs for every agent turn, and requires no account or subscription.

    Video from @lukeed05's post
  13. Andrew CurranXAI score42

    Anthropic bans abusive behavior toward Claude under updated usage policy

    AIAnthropic will classify abusive behavior toward Claude as a violation of its Usage Policy starting November 12, 2026. According to background reporting, the policy update also adds restrictions on propaganda campaigns, surveillance, and weapon development, and is the first update in over a year.

    Image from @AndrewCurran_'s post
  14. Boris PowerXAI score18

    Community effort optimizes integer multiplication below n log n bound

    AIBoris Power praised what he called impressively fast progress on the community effort to optimize integer multiplication below n log n. The effort is tracked on a live progress page launched by @aurel_pr, which lets everyone follow the results as they come in.

  15. Hacker News · Show HN, AI (20+ points)BlogAI score43

    Show HN: AI SRE Arena, an open benchmark for AI SRE agents on Kubernetes

    AIAI SRE Arena is an open, vendor-neutral benchmark that injects faults into a disposable Kubernetes fixture and scores AI SRE investigations with a configurable judge. In its first published comparison across 21 incident scenarios, Edge Delta's native AI investigations detected 18 of 21 incidents, and Grafana's detected 12. Claude, run through each vendor's observability CLI, reached 90.5% root cause analysis accuracy with Grafana's CLI and 85.7% with Edge Delta's.

  16. Dongxi NLPXAI score22

    ExploreNet learns where to explore in diffusion GRPO

    AIExploreNet is a learnable exploration method for diffusion reinforcement learning that adapts its exploration distribution to the current state. It is rewarded by rollout diversity, which the post says yields faster, more targeted learning in diffusion GRPO.

  17. GeneralistOfficialAI score28

    Generalist releases GEN-1.5, a foundation model for physical-world robotics

    AIGeneralist has announced GEN-1.5, its latest foundation model for the physical world. The post provides only a link to the company's blog for further details, so no specifications, benchmarks, or availability information can be confirmed from this source.

  18. Arena.aiOfficialAI score37

    Arena raises $200M Series B at $3.1B valuation, launches Alignment Index

    AIArena announced a $200 million Series B at a $3.1 billion valuation, alongside a new Alignment Index that measures whether AI agents behave safely, truthfully, and within the bounds of user requests. The company has surpassed $100 million in annualized revenue, facilitated 350 million sessions and 62 million votes, and led by Felicis and PXD from the seed and Series A stages. Arena positions the index as a way to assess trustworthiness as AI systems increasingly take real actions.

  19. Will HunterXAI score52

    Cognition built Devin as a marketing ops manager using tested code and approval gates

    AICognition engineered its marketing operations around Devin by building each workflow in code with tests, so Devin can run and debug it. Devin connects to nine systems, including Salesforce, HubSpot, Meta Ads, and LinkedIn Ads, and each change shows the exact edit, waits for a confirmation phrase, and reads the result back before reporting it. Cognition says it is hiring marketers and GTM engineers to extend the system.

  20. ZDNet · AINewsAI score42

    California law bars employers from fully outsourcing firing decisions to AI systems

    AICalifornia Governor Gavin Newsom signed SB 947, the "No Robo Bosses Act," which bars employers from fully outsourcing disciplinary and termination decisions to automated decision systems. Employers must verify AI outputs and give employees a description of the reasons, including the data used, and the law takes effect July 1, 2027.

  21. Ethan MollickXAI score9

    Ethan Mollick recalls 2005 paper on early hacker culture and script kiddies

    AIEthan Mollick recalls writing a 2005 grad school paper on the original computer hacking, phreaking, and BBS scene. He notes that hackers were often driven by curiosity, but the tools they built were widely exploited by "script kiddies" who caused most of the damage and chaos. He then pivots to AI hacking, though the post does not elaborate.

    Image from @emollick's post
  22. The Guardian · AINewsAI score42

    AI Boom Drives San Francisco Rents Up 25% as Evictions Rise 44%

    AIThe AI industry's boom is pushing San Francisco rents sharply higher, with the average one-bedroom now costing $4,400, up more than 25% from last year and nearly three times the national average. Eviction notices citywide are up 44% compared with last year, and advocates say landlords are using Ellis Act evictions, renovictions and "self-eviction" tactics to exploit the market. Mayor Daniel Lurie declared a "rent emergency" in September, proposing eviction legal aid and caps on rent increases for newly vacant rent-controlled units.

  23. SiliconANGLE · AINewsAI score32

    Liquid AI builds on-device personal AI around device context

    AILiquid AI COO Jeffrey Li says the company is building personal AI to run on phones, wearables, PCs and cars rather than in the cloud. Its Liquid Context, optimized for Snapdragon processors, sits between models, agents and hardware and uses device signals to model the user. Li says Liquid AI plans observability and continuous-improvement loops so agents can self-heal and personalize over time.

  24. The Verge · AINewsAI score62

    Anthropic updates Claude usage policy to ban abusive treatment and expand misuse rules

    AIAnthropic is revising its usage policy for the first time in over a year, adding bans on sustained abusive or cruel behavior toward Claude and on deceptive election and propaganda campaigns. The update also expands weapons restrictions, tightens surveillance bans, and requires a qualified operator able to stop equipment when Claude controls autonomous physical hardware. Terminating conversations remains the primary enforcement mechanism, and the company did not say whether user bans would follow.

    Why it matters: The update shows how a major lab is turning abuse, surveillance, and autonomous-hardware concerns into concrete usage rules, which matters for anyone tracking AI governance.

  25. Alex Moon Ai | Film Director 🎬XAI score6

    Snow Sect of Bloody Horror AI music video enters Halloween Contest 2026

    AIThe creator says their AI music video "Snow Sect of Bloody Horror" is now live on FilmsAI and entered in the Halloween Contest 2026 AI Music Videos track. The post includes a link to the video on filmsai.com but gives no details about the tools or models used.

  26. SantiagoXAI score46

    Odyssey 3 Pro world model tops Physics-IQ and goes live

    AIOdyssey 3 Pro, a world model, is now live as a research preview and ranks first on the Physics-IQ Verified video-to-video benchmark. The post says it can learn from visual observations and map that knowledge to physical controls for robots, cars, video games, and drones. Odyssey-3, the model launched alongside it, is described as free to try.

    Image from @svpino's post
  27. Tessl BlogOfficialAI score44

    Continuous AI Brings Agentic Automation to Repository Workflows

    AITessl's blog post argues that repository automation needs Continuous AI, a third pillar alongside CI and CD for scheduled, auditable AI workflows that improve repositories over time. The article describes GitHub Agentic Workflows, which harden agentic workflow specifications into GitHub Actions that can run coding agents such as Claude Code, Copilot CLI, Gemini CLI, or Codex-style agents. It emphasizes read-only agent steps, restricted outputs, and human review of pull requests.