Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 2

Oct 2Fri
  1. eric zakariassonXAI score47

    xAI releases experimental TypeScript SDK with Grok models and tools

    AIxAI has released an experimental TypeScript SDK, installable via npm install @xai-official/sdk, that covers text, voice, image, and video in one package. It provides access to the latest Grok models along with server-side tools including real-time X search, web search, code execution, and remote MCP.

    Video from @ericzakariasson's post
  2. DatabricksOfficialAI score44

    Omnigent: open-source meta-harness coordinating Claude Code and Codex agents

    AIDatabricks' new open-source meta-harness, Omnigent, lets multiple coding agents such as Claude Code and Codex share sessions, rules, and security policies in one system. A walkthrough by @leonvz demonstrates forking work across agents, multi-agent review and debate with Debby, and splitting implementation across subagents with Polly.

    Video from @databricks's post
  3. ChatGPTOfficialAI score60

    ChatGPT adds Finances for subscriptions, budgets, credit, and investments

    AIChatGPT now offers Finances, which can find forgotten subscriptions, flag unfamiliar or duplicate charges, and track recurring bills that have increased. It also provides weekly updates, monthly spending breakdowns, budget building, credit score tracking, debt payoff planning, emergency fund estimates, and investment mix and concentration views. Users can access it at

    Why it matters: The post lists concrete finance features across budgeting, debt, and investments, showing how a general assistant is expanding into personal money management.

  4. ChatGPTOfficialAI score60

    Finances in ChatGPT rolls out to Free and Go users in the U.S.

    AIChatGPT's Finances feature is rolling out to Free and Go users in the U.S. Users can securely connect their accounts through Plaid and Experian to get answers based on their own financial information.

    Why it matters: The post names the rollout scope and the account connection method, which helps readers judge how the feature handles personal financial data.

    Video from @ChatGPT's post
  5. Harrison ChaseXAI score38

    LangSmith Custom Apps lets teams build trace review UIs in-workspace

    AILangChain's LangSmith Custom Apps lets agent teams build their own review UI over their traces and publish it directly into the workspace. Developers build the interface on their LangSmith data, while hosting, authentication, and permissions are handled by the platform.

  6. CursorOfficialAI score42

    Cursor's Rollouts detects deployment regressions and launches cloud agent fixes

    AICursor introduced Rollouts, a tool that writes a monitoring plan and watches changes as they deploy to catch regressions before users see them. When Rollouts detects a regression, it identifies the offending PR and opens an issue, and one click starts a cloud agent to fix it. Rollouts usage credits are included through Oct 3.

  7. GitHubOfficialAI score44

    GitHub Copilot adds Project HydraFusion and new models to model picker

    AIGitHub has made the Project HydraFusion research preview available in the GitHub Copilot app and @code, where it orchestrates multiple models rather than acting as a single model. New models from Anthropic (Fable 5.1 and Opus 5.5) and OpenAI (GPT-6.1 Sol) are also now selectable in the Copilot model picker.

  8. Together AIOfficialAI score34

    Together AI shares how its team uses AI to boost collective productivity

    AITogether AI's CPO and product team outlined how they use AI to make the whole team more productive, not just individuals. The approach includes a shared context repo readable by any AI harness, cutting half a day of research to about 5 minutes, and evals that test their product the way agents actually use it.

  9. Ant LingOfficialAI score27

    Ling-3.1-flash now available free on OpenRouter

    AIAnt Ling has made Ling-3.1-flash available on OpenRouter at no cost, inviting users to try it and share feedback. The post provides no details on model size, benchmarks, context length, or pricing beyond the free access.

  10. Harrison ChaseXAI score26

    LangChain improves memory for managed Deep Agents in enterprise settings

    AIHarrison Chase says LangChain is improving memory in managed Deepagents, noting that memory is difficult to get working well in company settings. The linked background post describes user memory in Managed Deep Agents 0.8, which lets an agent remember the people it works with.

  11. ElevenLabsOfficialAI score32

    ElevenLabs earns FedRAMP 20x Class A certification for federal voice AI

    AIElevenLabs has achieved FedRAMP 20x Class A certification, covering ElevenAgents and its Text to Speech and Speech to Text APIs when run in Zero Retention Mode with US data residency. The company says the published evidence package and FedRAMP Marketplace listing should shorten security reviews for public sector and regulated buyers. The certification extends ElevenLabs for Government, its dedicated offering for federal agencies.

    Image from @ElevenLabs's post
  12. Microsoft AIOfficialAI score41

    Microsoft MAI voice models now available on LiveKit for agents

    AIMicrosoft's MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash are now live on LiveKit for building voice agents. LiveKit says MAI-Transcribe-2-Streaming debuts at #1 on the Artificial Analysis accuracy leaderboard, and suggests pairing it with MAI-Voice-2.1-Flash for efficient, expressive voice agents.

  13. Microsoft CopilotOfficialAI score20

    Copilot Code lets more people build apps and workflows

    AIMicrosoft's Copilot Code is designed to help more people turn ideas into apps, workflows, and solutions for their work. Microsoft Copilot EVP Jacob Andreou discusses how the product expands who gets to build.

    Video from @MSFTCopilot's post
  14. Google AIOfficialAI score62

    Google launches Project Suncatcher prototype satellite to test TPUs in orbit

    AIGoogle AI announced that its Project Suncatcher prototype satellite, built with Planet, has launched into orbit on SpaceX's Transporter-18 rideshare mission. The initial mission will gather data on how Google TPUs handle the physical stress and extremes of spaceflight. The post says low Earth orbit systems could generate up to 8x more solar power than on Earth, and that future work may link multiple satellite constellations for scaled machine learning.

    Why it matters: The post explains a space-based machine learning prototype and why orbit's near-constant sunlight matters, which helps readers weigh the idea's practical potential.

    Video from @GoogleAI's post
  15. Kilo (acq. by Anaconda)OfficialAI score36

    Ling 3.1 Flash is free in Kilo Code until October 13

    AIKilo Code is offering Ling 3.1 Flash for free until October 13, with the model served by Novita Labs. Ant Ling's background post describes the model as roughly 560B total parameters with about 25B active per token and a context window of up to 1M tokens. Ant Ling says it scores 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional, and plans to open-source it soon.

  16. Microsoft AIOfficialAI score36

    Microsoft's MAI voice models now available on OpenRouter

    AIMicrosoft AI's MAI voice models are now accessible through OpenRouter, according to the announcement. The quoted OpenRouter post highlights MAI-Voice-2.1, a text-to-speech model that supports 23 languages with native accents and is priced at $22 per 1M characters.

  17. Hugging Face BlogOfficialAI score70

    Ai2 open-sources AstaBrief 8B, a fast model for generating cited research reports

    AIAi2 released AstaBrief 8B, an open-weights model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. The model runs as Fast mode in Asta, averaging 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The post also describes filtering synthetic training data by citation density and building DPO pairs judged by two models that agreed.

    Why it matters: The post explains how supervised fine-tuning, preference data, and citation-density filtering were used to build a cited-report model, which is useful for teams training their own models.

  18. RunwayOfficialAI score25

    Canva, ElevenLabs, and Nebius execs discuss AI tools for creatives

    AICanva Head of AI Research Stefano Corazza, ElevenLabs CRO Ashley Kramer, and Nebius CMO Lindsey Irvine discuss building AI tools for creatives. They emphasize that control and consistency matter most to users. They also address how agents are changing the way marketing teams work.

    Video from @runwayml's post
  19. Mustafa SuleymanXAI score36

    Microsoft's Voice and Transcribe Streaming models now available on Vercel

    AIMicrosoft AI's Voice and Transcribe Streaming model is now available on Vercel for building agents, and the post claims it ranks first for quality and speed. The post says it is cheaper than other hyperscalers and 60% cheaper than Eleven Labs.

  20. Latent.SpaceXAI score31

    Airbnb CTO Ahmad Al-Dahle takes inside-out approach to AI adoption

    AIFormer Meta Llama leader Ahmad Al-Dahle, now Airbnb CTO, is applying an inside-out AI strategy that transforms internal operations before changing customer experiences. The approach includes Everest, an internal tool that helped speed up a product launch.

  21. Liquid AIOfficialAI score64

    Hugging Face guide shows multi-harness RL for coding agents via a capture proxy

    AILiquid AI shared a Hugging Face guide to multi-harness reinforcement learning for coding agents, in which a proxy records the token ids and logprobs vLLM samples so training works without changing the harness. Per the quoted post, LFM2.5-2.6B rose from 42% to 54% after training across four harnesses at once, and imitation fine-tuning on 3,189 rollouts from Qwen3.8-27B plateaued at 47.5%, below both RL runs. The proxy, trainer, tasks, SFT data, training code and seven trained models are described as open.

    Why it matters: The guide explains how to train one model with RL across several coding agent harnesses without modifying the harnesses, using a proxy that records token ids and logprobs.

  22. Google · AI blogOfficialAI score58

    Google recaps September 2026 AI launches, led by Gemini 4 Argon

    AIGoogle's September 2026 roundup highlights Gemini 4 Argon, a frontier model with a 1-million-token output limit aimed at complex tasks such as cybersecurity defense. Argon is rolling out first to trusted cyber defenders through the Fairwind Program, with developer, enterprise, and consumer access to follow after guardrail feedback. The post also covers Gemini 3.8 Flash, Connected Apps in Gemini, and WeatherNext 3.

  23. Google ResearchOfficialAI score60

    Google's TEE-based federated learning system adds verifiable privacy guarantees

    AIGoogle announces a next-generation federated learning system that uses Trusted Execution Environments to provide verifiable, auditable data anonymization. The system publishes access policies to a public transparency log and is deployed in Gboard, which has launched English and Japanese next-word prediction models with stronger privacy guarantees and improved accuracy. Training time has also sped up significantly because computation moved to the server and is parallelized across many machines.

    Why it matters: The post shows how Trusted Execution Environments make federated learning's privacy claims externally verifiable, rather than relying on trust in the server operator.

  24. merveXAI score36

    llama.cpp adds support for decision models on modest hardware

    AIllama.cpp now supports decision models, which route tickets, moderate content, or choose an agent's next step by returning a probability for every option. Five open models from 144M to 27B parameters are supported at launch, and the team says more will follow in the coming days. Because most decision models do not need large GPUs, they are a good fit for llama.cpp, and a Hugging Face blog post explains how to set them up.

  25. Hugging FaceOfficialAI score67

    Hugging Face guide shows how to train agent models across multiple harnesses with RL

    AIHugging Face and collaborators published a guide to multi-harness RL that trains models through a capture proxy without changing the agent harness. The proxy records the token ids and logprobs vLLM samples, and the source reports LFM2.5-2.6B rising from 42% to 54% after training across four harnesses. Fine-tuning on 3,189 successful rollouts from Qwen3.8-27B plateaued at 47.5%, below both RL runs, and the capture proxy, trainer, tasks, SFT data, training code, and seven trained models are released openly.

    Why it matters: The source gives a concrete method for training models across several agent harnesses, with measured gains and a note that imitation learning underperformed RL.

    Image from @huggingface's post
  26. Latent SpaceBlogAI score43

    Airbnb CTO Ahmad Al-Dahle details AI rollout across engineering and support

    AIAirbnb CTO Ahmad Al-Dahle, who joined in January from Meta, says 60% of the company's code is now AI-authored and pull-request throughput per engineer is up about 1.6x. He says roughly half of Airbnb's support tickets are now resolved by AI, in line with a nearly 45% figure from the company's Q2 results. Airbnb's internal context graph, Everest, helped launch its grocery delivery and airport pickup services, which Al-Dahle says took eight to nine months and about six weeks to develop, respectively.

  27. RunwayOfficialAI score23

    Runway AI Summit panel discusses real-world robot evals and deployment

    AIParil Jain of The Bot Company and Quan Vuong of Physical Intelligence discussed real-world robot evaluation, field deployment, and where robotics will be in three years at the Runway AI Summit. The panel was described in the post without specific findings or figures.

    Video from @runwayml's post
  28. GitHub Copilot ChangelogOfficialAI score53

    GitHub Copilot adds new models, dynamic workflows, and desktop app automation

    AIGitHub Copilot's weekly release adds Claude Sonnet 5.5 and GPT-6.1 Sol for specified plan tiers, plus HydraFusion, a research preview that lets Copilot select and coordinate models for a task. It also introduces dynamic workflows in public preview, which let users save and reuse multi-step processes, and computer use in public preview on macOS and Windows for automating desktop apps.