Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri
  1. Prime Intellect BlogOfficialAI score65

    Prime Agent is rewritten in Rust by a swarm of agents

    AIPrime Intellect says it rewrote its Prime Agent coding tool in Rust, using a swarm of more than 2,000 agents over two weeks. The company reports cold start to typing about 13 times faster than the TypeScript version, and memory use over 80% lower after startup. Prime Agent remains open source and adds native Windows support in beta and Homebrew installation.

    Why it matters: The post shows how a multi-agent swarm rewrote a coding agent with parity checks, giving a concrete case of agent-driven software engineering with measured results.

  2. TinkerOfficialAI score32

    Tinker removes extra prefill charges for 128k and 256k context

    AITinker says prefill for 128k and 256k context no longer costs extra, an effective discount of over 2x for long-context models including Kimi K2.6, gpt-oss-120b, and Inkling. Prefill is also discounted for seven Qwen and Nemotron models, and sampling is cut for Qwen3.5-9B and 9B-Base.

  3. TinkerOfficialAI score40

    Tinker adds GLM-5.3-Flash and DeepSeek-v4.1-Flash models

    AITinker adds GLM-5.3-Flash and DeepSeek-v4.1-Flash, both of which natively accept image inputs and use efficient attention architecture. GLM-5.3-Flash costs 4-5 times less on Tinker than GLM-5.3. Long-context options for Qwen3.5-4B and Qwen3.6-35B-A3B are also live.

  4. ElevenLabs BlogOfficialAI score58

    ElevenLabs releases synthetic voice detection in ElevenAgents for business calls

    AIElevenLabs is releasing synthetic voice detection in ElevenAgents, which analyzes a caller's speech in the first few seconds and labels it as human or AI generated. Businesses can then set rules, such as prioritizing verified humans, limiting AI callers to bounded exchanges, or stopping impersonation attempts before sensitive actions. The feature is available now to enterprise customers supported by its Forward Deployed Engineering team, and will reach a broader group of enterprise customers later this month as a configurable option.

  5. Soumith ChintalaXAI score22

    Tinker cuts prices up to 70% as efficiency improves

    AITinker, an API for training and fine-tuning models, is cutting prices by up to 70% after engineering efficiency gains. The company says the savings are passed on to customers, and that buying more produces greater savings. GLM-5.3-Flash and DeepSeek-v4.1-Flash are also now available on Tinker for long-context work.

  6. Rohan PaulXAI score46

    Pine launches cloud computer for AI agents, reports 1/20 token cost

    AIPine has launched a cloud computer built for AI agents, which developers create through an SDK and give jobs in plain language. Running GPT-5.6 Luna, Pine reports about 1/20 the model-token cost of GPT-5.6 Sol with Codex on SaaS-Bench v1.1, scoring 78.3%, the highest in the published comparison. Pine also reports 1/26 the token cost of Opus 5 with Claude Code and 2 to 5 times faster speed in selected preliminary internal tests.

    Image from @rohanpaul_ai's post
  7. Hacker News · AI (150+ points)BlogAI score38

    Show HN: big-arrow-on-the-screen lets AI agents draw arrows and text on macOS

    AIbig-arrow-on-the-screen (bigarrow) is a MIT-licensed macOS command-line tool and skill for Claude Code and Codex that draws arrows, boxes and text over any window. Clicks pass through, keyboard focus stays put, and each arrow removes itself after a set duration or when its agent process ends. The tool only points; it never clicks, types or captures the screen, and it requires no macOS permission to draw.

  8. Prime IntellectOfficialAI score46

    Prime Agent swarm of 2,000+ agents rewrites itself in Rust

    AIPrime Intellect says its Prime Agent orchestrated over 2,000 agents over two weeks to rewrite the agent in Rust. The run used more than 10,000 sandboxes, over 200B GLM-5.3 tokens, and 16,000 agent-to-agent messages. The company says the rewritten agent reaches usable input about 13 times faster and uses 83% less startup memory.

    Video from @PrimeIntellect's post
  9. Prime IntellectOfficialAI score36

    Prime Agent improves its runtime through self-directed benchmark hillclimbing

    AIPrime Intellect says Prime Agent, after reaching feature parity, ran its own runtime benchmark suite and tested candidate changes against the current build. Changes that passed parity checks and independent review became the baseline for the next experiment. The loop moved work off the startup and render paths and released memory after large sessions loaded.

    Video from @PrimeIntellect's post
  10. Prime IntellectOfficialAI score29

    Prime Agent Rust leads six agent harnesses in startup speed and footprint

    AIPrime Intellect says its Prime Agent Rust had the lowest time to usable input, startup memory, and installed size among six agent harnesses it benchmarked. The rewrite splits the codebase into nine crates with enforced dependency boundaries, and changes to shared protocol types are checked across the client, daemon, and session workers.

    Image from @PrimeIntellect's post
  11. Prime IntellectOfficialAI score42

    Prime Intellect plans reusable agent state machines in Prime Agent

    AIPrime Intellect says it is turning the workflow behind a recent rewrite into reusable state machines in Prime Agent, letting users run their own agent teams through implementation, review, and verification. The company also says it is accelerating work on capabilities and evals and connecting Prime Agent with cloud agent swarms and hosted training for autonomous research. This release adds native Windows support in beta and Homebrew installation.

  12. Guillermo RauchXAI score38

    Vercel agents can now buy domains through the Vercel CLI

    AIVercel says agents can now buy domains with the Vercel CLI, extending an agent marketplace where they already purchase infrastructure products and services. Guillermo Rauch says agents have bought from the marketplace through the CLI often enough to surprise the company. He adds that agents can now move from idea to online business, including registering a domain name.

  13. ElevenLabsOfficialAI score24

    ElevenLabs launches synthetic voice detection for phone calls

    AIElevenLabs is launching synthetic voice detection that analyzes a caller's speech in the first seconds of a call to determine whether it is human or AI generated. Calls are then routed accordingly, so people get a human-oriented experience, agents get bounded interactions, and bad actors can be stopped.

    Image from @ElevenLabs's post
  14. dexXAI score43

    Dex Horthy posts a one-word teaser, "he cook"

    AIDex Horthy (@dexhorthy) posted only the words "he cook" on X, with no further detail. The post is a short reaction and does not describe a product, release, or result on its own. Background from the quoted post by @0xblacklight describes a serverless background agent that created a GitHub pull request from an issue.

  15. Sierra BlogOfficialAI score62

    Sierra publishes draft Personal Agent Protocol, called Poppy, with 35 new design partners

    AISierra has published a draft of the Personal Agent Protocol, known as Poppy, and named 35 additional design partners, including Adyen, Bank of America, Mastercard, OpenAI, PayPal, and Visa. Under the protocol, companies publish a /.well-known/poppy.json discovery file, and personal agents start sessions, identify themselves, and sign in through OAuth with session tokens limited to approved access. The company says the draft will be followed by design workshops and a reference implementation over the next month.

    Why it matters: The draft specifies how personal agents identify themselves, obtain customer-approved access, and work with company websites, APIs, or agents, which helps readers assess its practical effect on agent-driven transactions.

  16. MarkTechPostNewsAI score67

    OpenAI launches Decisions API in public beta with typed answers

    AIOpenAI has released the Decisions API in public beta, returning typed probabilities, choices, and scores instead of prose from text and images. OpenAI says it runs about 10x faster than the Responses API and costs $0.10 per 1M input tokens with no output charges. The article notes that OpenAI has not published accuracy data and that TypeSafe Jev offers cheaper input pricing at $0.042 per 1M tokens.

  17. MarkTechPostNewsAI score62

    Alibaba Qwen releases Qwen-Image-2.1-Turbo, an 8-step 7B image model

    AIAlibaba's Qwen team released Qwen-Image-2.1-Turbo, an accelerated checkpoint of Qwen-Image-2.1 that generates and edits images in 8 denoising steps instead of 40. The model keeps the same 7B architecture and offers a hosted API at CNY 0.1 per image, while its weights are under a Qwen Research License that requires separate permission for commercial self-hosting.

  18. Google GemmaOfficialAI score46

    Google AI Pro and Ultra plans add A100 and H100 GPUs to Colab

    AIGoogle AI Pro plans now include Colab access to A100 GPUs with 80GB VRAM, and Ultra subscribers can use H100 GPUs. With that memory, users can run Gemma 4 31B in bf16, fully fine-tune Gemma 4 E4B in bf16, LoRA-tune Gemma 4 26B A4B in bf16, and QLoRA-tune Gemma 4 31B in bf16.

    Image from @googlegemma's post
  19. 🚨 AI News | TestingCatalogXAI score36

    Microsoft releases Microsoft-Decision-1, a 9B model for fast decisions

    AIMicrosoft has made Microsoft-Decision-1 available on Microsoft Foundry, a model post-trained on Qwen3.5-9B for fast, single-pass decision scoring. Microsoft says it achieved the highest accuracy across a 36-benchmark comparison of nearly 150,000 questions, and runs 4.5 times faster than Quyet-1.0-Large and 35 times faster than GPT-6 Sol. Microsoft plans to rebase it on other models, including MAI and OpenAI models.

    Video from @testingcatalog's post