Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 5

Oct 5Mon
  1. Alex HeathXAI score52

    Reflection's founders discuss building a DeepSeek of the West with Beam

    AIReflection is set to release Beam, its first open-weight AI model, aiming to become a Western counterpart to DeepSeek. The source says Beam is trained from scratch for coding, reasoning, and AI agents, with benchmarks placing it alongside the strongest open models and more efficient token economics. Reflection has raised $4.6 billion from investors including Nvidia, Sequoia, and Lightspeed, and the interview covers its monetization plans for open-weight models.

    Video from @alexeheath's post
  2. Gergely OroszXAI score35

    Gergely Orosz says coding agent product strategy feels like "YOLO"

    AIGergely Orosz says many coding agents seem to follow a "YOLO" product strategy, with rapid week-over-week change learned about through random social media posts. He notes this makes some sense given how quickly the industry and capabilities keep changing. Quoted context reports that Anthropic is removing Cowork's local option for Pro/Max users, with new tasks running in the cloud while existing local tasks stay on the computer.

  3. CognitionOfficialAI score58

    Cognition's Devin adds Dreaming, a nightly memory graph across sessions

    AICognition introduces Dreaming, a feature in which Devin builds a memory graph of how a user likes to work across sessions. At night, Devin self-improves this memory by removing stale records and discovering latent information. Cognition also says it is creating an open-source standard called Agent Memory Repo, linked in the post.

    Video from @cognition's post
  4. OpenAIOfficialAI score38

    OpenAI limits text watermark detector access citing watermark weaknesses

    AIOpenAI says text watermarks are often undetectable in short passages and can be fully removed by rewriting or translating text. For now, only approved researchers will get access to its detector so they can help evaluate and improve the technology. OpenAI says it will keep testing and refining text watermarking with feedback from users, developers, policymakers, and researchers.

  5. OpenAIOfficialAI score30

    OpenAI adds invisible text watermarks to detect model-generated content

    AIOpenAI says it may give people a choice about text watermarking, which embeds an invisible statistical signal during generation to help show whether text was likely produced by an OpenAI model. The watermark does not reveal the text's author, owner, or any person, account, conversation, or prompt. In OpenAI's testing, watermarking did not affect model capability, speed, or response quality.

  6. OpenAIOfficialAI score58

    OpenAI expands content provenance to text watermarking for EU AI Act compliance

    AIOpenAI is extending its content provenance approach to text, starting with watermarking eligible text from ChatGPT and Codex in the EU over the coming weeks. The company says this is in response to EU AI Act requirements and acknowledges the significant limitations of current text watermarking technology. API customers can turn on text watermarking for select models worldwide starting today.

  7. Vaibhav (VB) SrivastavXAI score42

    OpenAI speeds up GPT-6 Astra and GPT-6.1 Sol inference by 50%

    AIOpenAI has optimized inference for GPT-6 Astra and GPT-6.1 Sol, making them about 50% faster by default across subscription plans and Sign in with ChatGPT partners. The change requires no action from users and should be noticeable within two hours of rollout.

  8. Liquid AIOfficialAI score37

    Liquid AI's d1 decision model adds vision, rivaling GPT-6.1 Sol at lower cost

    AILiquid AI released d1 with vision support, accepting images, text, or both as inputs. In tests on six real applications, d1 matched or beat GPT-6.1 Sol on four while costing 19x to 200x less than both GPT-6.1 Sol and Claude Opus 5.5. It returns probabilities for yes/no, choice, or score questions in one forward pass, with text decisions in 200 to 300 ms.

    Image from @liquidai's post
  9. Liquid AIOfficialAI score36

    Liquid AI's d1 model inspects parts from camera images with 85-97% accuracy

    AILiquid AI's vision-enabled decision model d1 inspects parts directly from camera images and is described as the best such model currently on the market. It reaches 85% to 97% accuracy across four VisA inspection tasks covering circuit boards, candles, cashews, and chewing gum. It understands each task from a short description without task-specific training.

    Video from @liquidai's post
  10. Liquid AI · new models on Hugging FaceOfficialAI score44

    LiquidAI releases d1-omni-600M, a 600M decision model for text, image and audio

    AILiquidAI has released d1-omni-600M on Hugging Face, a 587M-parameter model that answers named yes/no, choice and score questions over text, images or up to 30 seconds of speech in a single forward pass. It returns typed answers with zero output tokens by reading the model's distribution over options, and is built on LFM2.5-Encoder-350M with a 16,384-token context length. The model is not a chat model and does not generate text.

  11. Google AntigravityOfficialAI score20

    Gemini 3.8 in Antigravity builds a self-generating pencil sketch forest

    AIGoogle Antigravity shared a demo of a mini self-generating forest made entirely of colorful pencil strokes, built with Gemini 3.8 from a sketch reference and a prompt. The quoted post says the result came from turning an art sketch into an exploratory mini game.

  12. Google AntigravityOfficialAI score23

    Google Antigravity adds AlphaGenome Atlas Skill for genomic research workflows

    AIGoogle Antigravity has integrated the AlphaGenome Atlas Skill into its scientific workbench, enabling AI agents to help researchers prioritize genetic variants, generate structural plots, and build testable hypotheses. The company showcases researchers Natasha and Kyle using the tool in a demonstration video.

  13. LiveKitOfficialAI score30

    LiveKit demos Microsoft speech models in a voice support agent

    AILiveKit Agents pairs MAI-Transcribe-2-Streaming for speech-to-text, Gemma 4 on LiveKit Inference for reasoning and tool calls, and MAI-Voice-2.1-Flash for speech output in a demo support call. The post links separate speech-to-text and text-to-speech resources for developers.

    Video from @livekit's post
  14. Google AIOfficialAI score46

    Gemma 4 and BOTANIC-1 pinpoint crop-yield DNA mutations in minutes

    AILiving Models paired Google's Gemma 4 with BOTANIC-1, a plant-DNA model trained on 320 species, to identify causal genetic variants. In a melon yield test, the pipeline ranked the target mutation first out of 2,494 possibilities in under four minutes. The approach aims to speed up breeding of climate-resilient crops that would otherwise take years of field trials.

  15. X.PINXAI score50

    Huawei and Qualcomm reach multiyear cross-licensing patent deal

    AIHuawei's new multiyear patent agreement with Qualcomm would make Qualcomm the net payer for the first time, according to Nikkei Asia. The deal cross-licenses patents in 5G, computing, AI, and networking, and Qualcomm will also buy some of Huawei's U.S. patents outright. Financial terms have not been disclosed, and the transaction still requires regulatory approval.

  16. CursorOfficialAI score39

    Cursor lets users replace its system prompt with their own

    AICursor is enabling an option to replace its built-in system prompt with a custom one, while rules, skills, and tool schemas still load. The feature is being rolled out account by account rather than to all users at once.

    Image from @cursor_ai's post
  17. CursorOfficialAI score38

    Cursor SDK agents can now be steered while running

    AICursor announced that developers can steer Cursor SDK agents while they run using run.steer(), which adds a message to the next turn. If a subagent is mid-task, it moves to the background and continues working.

    Video from @cursor_ai's post
  18. Together AIOfficialAI score34

    Together AI launches Together Link to run open models in coding harnesses

    AITogether AI has announced Together Link, which lets developers run frontier open models inside their favorite coding harness. The product includes spending tracking and an Auto router that selects low-cost models for quick fixes and more capable models for harder tasks.

    Video from @togethercompute's post
  19. Nous ResearchOfficialAI score38

    Upstage's Solar Mini 4 free on Nous Portal for two weeks

    AIUpstage's Solar Mini 4 is free on Nous Portal for the next two weeks. The model has 3B active parameters out of 35B total and a 512K context window. It scores 24 on the Artificial Analysis Intelligence Index, above models with roughly 10x the active parameters.

    Video from @NousResearch's post
  20. GitHub Blog · AI & MLOfficialAI score63

    GitHub releases ReviewBench, an open benchmark for AI code review agents

    AIGitHub has released ReviewBench, an open benchmark for evaluating AI code review agents on 219 public pull requests across 19 languages. The benchmark reports grounded and augmented precision, recall, and F1 metrics, and its dataset, rubric, and judge are publicly available. GitHub says ReviewBench predicted the direction of a Copilot code review ensemble experiment's production results before A/B testing.

    Why it matters: The post explains how ReviewBench was built and validated, and reports an offline-to-production comparison that shows how well a benchmark predicts real experiment outcomes.

  21. FireworksOfficialAI score34

    DeepSeek V4.1 Flash now available for training on Fireworks

    AIFireworks AI has made DeepSeek V4.1 Flash available for training on its Dedicated Training API and Managed Training surfaces. The post positions the model as a strong base for agentic coding, terminal automation, and tool use, and notes it is cost-efficient to serve.

  22. Replit ⠕OfficialAI score22

    Replit weekly changelog adds GPT-6.1 Sol and Claude Sonnet 5.5 models

    AIReplit shipped a weekly update letting users build with GPT-6.1 Sol and Claude Sonnet 5.5, along with an Ask agent integration with Jev. The release also includes an updated Settings UI and enterprise Workplace controls for company-wide rules and controlled exceptions. Full details are in the Replit changelog.

  23. Tibor BlahoXAI score62

    OpenAI adds opt-in text watermarking for API and EU ChatGPT and Codex output

    AIOpenAI is rolling out text watermarking for EU AI Act compliance, with opt-in access for API customers globally on select models starting today. Watermarking stays off by default in the API, while an invisible watermark will be added to eligible ChatGPT and Codex text in the European Union over the coming weeks. Access to the text watermark detector is initially limited to approved researchers and expert organizations, and the image and audio verification tools remain publicly accessible.

    Image from @btibor91's post
  24. clem 🤗XAI score62

    Hugging Face turns 10 coding harnesses into RL environments via a capture proxy

    AIHugging Face says a capture proxy lets reinforcement learning train open models inside unmodified coding harnesses such as Claude Code, Codex, and OpenCode. The proxy records the exact token IDs and logprobs vLLM samples and hands them to TRL for training. On LFM2.5-2.6B, training in four harnesses at once raised OpenCode results from 34% to 58%, while SFT on 3,189 Qwen3.8-27B rollouts plateaued at 47.5%.

    Why it matters: The capture proxy lets models train inside real coding harnesses without reimplementing them, with measured gains and a comparison against SFT on the same data.

    Image from @ClementDelangue's post
  25. DatabricksOfficialAI score31

    Databricks makes IP Functions generally available for network analytics in SQL

    AIDatabricks has made IP Functions generally available, letting users parse, validate, and join IPv4 and IPv6 addresses and CIDR blocks with built-in SQL functions optimized in Photon. In benchmarks versus another leading cloud data warehouse, CIDR joins ran up to 3.1x faster and cost up to 6.4x less. The functions support its Security Lakehouse vision for threat detection, investigation, and network analytics on one governed copy of data.

    Image from @databricks's post
  26. ElevenLabsOfficialAI score22

    ElevenCreative launches $100,000 contest to find the catchiest ad

    AIElevenCreative is launching The Search, a $100,000 competition to find the world's catchiest ad, with a jingle people keep humming the next day as the core criterion. First place wins $50,000, and 11 winners each get a one-on-one session with the ElevenLabs Creative Production Team.

    Video from @ElevenLabs's post
  27. CohereOfficialAI score22

    Cohere's North connects to Slack, Notion, GitHub and more apps

    AICohere's North now links to everyday work apps including Slack, SharePoint, OneDrive, Outlook, Exchange, Jira, Linear, Notion, and GitHub. The post frames the integrations as a way to speed up everyday work.