Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 1

Oct 1Thu
  1. Ali GhodsiXAI score40

    Databricks launches ai_decide() for fast decisions on governed data

    AIDatabricks has released ai_decide(), a function that runs decision models natively across enterprise data, enabling fast typed decisions instead of slow text generation. The approach targets large datasets already governed within Databricks, according to the post and linked blog.

    Image from @alighodsi's post
  2. Higgsfield AI 🧩OfficialAI score28

    Higgsfield launches FLUX 3 Image model on its platform

    AIHiggsfield AI has made FLUX 3 Image available to try on its platform, linking directly to its image generation page with the flux-3-image model selected. The post provides no details on capabilities, pricing, or benchmarks.

  3. Higgsfield AI 🧩OfficialAI score28

    Higgsfield adds Black Forest Labs' FLUX 3 Image model for editing and 4K generation

    AIHiggsfield has launched FLUX 3 Image, a new Black Forest Labs image model, now available on its platform. The model lets users edit a single detail without altering the rest of the image, place objects using bounding boxes, combine up to 10 reference images, and generate output at up to 4K resolution.

    Video from @higgsfield's post
  4. Boris ChernyXAI score54

    Claude Code adds mods that customize behavior and UI via plugins

    AIClaude Code now supports mods that change its behavior, customize the UI, or add features, written in a few lines of TypeScript or generated by Claude. Mods ship inside plugins and are installed with /plugin in the CLI or desktop app, and the author says mods can be shared so others can try them.

  5. Mike KnoopXAI score44

    Qwen3.8-27B verified on ARC-AGI, nearly fitting Kaggle runtime limits

    AIMike Knoop notes that a verification of the roughly 27B-parameter model is notable because it is about as large as fits within the official ARC Prize Kaggle competition runtime. ARC Prize reports Qwen3.8-27B from Alibaba's Qwen team scored 42.4% on ARC-AGI v2 at $0.45 per task and 87.5% on v1 at $0.22 per task.

    Image from @mikeknoop's post
  6. Alex HeathXAI score46

    OpenAI's Dots lead ChatGPT's always-on personal agent plans

    AIAlex Heath's podcast with OpenAI's @embirico covers Dots, a new always-on personal agent in ChatGPT that asks permission before acting by default. The episode also discusses Space, a workspace where people and agents collaborate on documents, data, and projects, along with pricing and possible access for free users.

    Video from @alexeheath's post
  7. Latent.SpaceXAI score55

    Anthropic's Thariq explains what Claude Code mods can access and control

    AIAnthropic's Thariq describes Claude Code mods, which can read conversation scope such as turn count and token usage. Mods run in process, so they can spawn subagents, parse their results with structured output, and modify the UI, which hooks cannot do. Mods ship inside plugins and are installed with /plugin in the CLI or desktop app.

    Video from @latentspacepod's post
  8. SpaceXAIOfficialAI score46

    Grok 4.7 now available on Gemini Enterprise Agent Platform

    AIGrok 4.7 is now available on the Gemini Enterprise Agent Platform, according to the post from SpaceXAI, the account owned by xAI and Grok. The post gives no further details on pricing, context length, or capabilities.

    Video from @SpaceXAI's post
  9. ReplicateOfficialAI score47

    FLUX 3 Image launches with native 4K generation and bounding-box layout control

    AIBlack Forest Labs has released FLUX 3 Image, which generates native 4K images and supports hyper-specific layouts using bounding boxes. It can make multiple targeted edits at once while staying consistent across them, and up to 10 references can be used to compose an image. The first week is 50% off, and an open-weights version is coming in the coming weeks.

    Image from @replicate's post
  10. Sam AltmanXAI score47

    Sam Altman says ChatGPT subscriptions should work across third-party apps

    AISam Altman says users should be able to use their AI subscription wherever they need it, with an image shown rather than detailed text. The main post provides no further specifics on which services or features are covered. Related context indicates OpenAI recently shipped Sign in with ChatGPT and plans to answer questions about it.

  11. Lydia Hallie ✨XAI score62

    Claude Code adds mods that customize behavior and UI via TypeScript plugins

    AIClaude Code can now be modified with mods that change its behavior, customize the UI, and add features, written in a few lines of TypeScript or generated by Claude. Mods ship inside plugins and are installed with /plugin in the CLI or desktop app. A TypeScript function can intercept internal events such as tool calls, prompts, model requests, and renders, and add custom UI and commands.

    Why it matters: The post details how mods hook into Claude Code's tool calls, prompts, and rendering, which shows what extending the coding agent actually involves.

  12. Black Forest LabsOfficialAI score54

    Black Forest Labs introduces FLUX 3 Image with precise editing controls

    AIBlack Forest Labs announces FLUX 3 Image, which supports multi-turn edits that leave other pixels unchanged, layout control via bounding boxes, generation up to 4K, and up to 10 reference images. Commercial weights are available for companies running image generation at scale, and an open weights version is launching in the coming weeks.

    Video from @bfl_ai's post
  13. TypeSafe AIOfficialAI score34

    Jev outperforms LLM judge for research agent risk monitoring at 250x lower cost

    AIIn a business risk monitoring test by @edwardirby, the Jev judge matched the report quality of an ordinary LLM judge while missing no investigations, versus the LLM missing 5 of 11. The LLM's threat scores also flip-flopped from 0.35 to 0.68 to 0.50 on the same threat, while Jev was 250x cheaper and 3-6x faster.

  14. AnthropicOfficialAI score38

    Harvard physicist builds toolkit to match Claude with science calculations

    AIHarvard physicist Matthew Schwartz argues that LLMs are poorly matched to science when used as human-style collaborators, so he built a toolkit for exact quantitative calculations. Working with Claude, the approach surfaced connections to ecology, population genetics, and a dozen other fields, with domain experts steering it toward interesting questions.

  15. Sophia YangXAI score20

    Ember-1 Shows Lower Cost and Faster Reasoning Than Kimi K3

    AIEmber-1, built on the Kimi K3 base model, is reported by @kickingkeys to cost about 20% less and use about 36% fewer reasoning tokens across 124 test prompts. The same tester says it runs about 2x faster, while the source's painting examples show its own expressive style.

  16. Perplexity DevelopersOfficialAI score41

    Perplexity launches Decisions API powered by pplx-decider-v1-27b

    AIPerplexity introduced its Decisions API, powered by pplx-decider-v1-27b, a multimodal model that outputs a probability distribution over a fixed set of answers rather than text. The company says the API costs $0.04 per million input tokens and scores 85.71% across benchmarks.

    Image from @perplexitydevs's post
  17. Microsoft CopilotOfficialAI score34

    Microsoft Copilot adds GPT-6.1 Sol and Claude Sonnet 5.5 models

    AIMicrosoft Copilot begins rolling out OpenAI's GPT-6.1 Sol and Anthropic's Claude Sonnet 5.5 today, joining Claude Opus 5.5 and GPT-6 Sol added earlier this month. Users can pick the model suited to each task, with Work IQ grounding responses in their files, meetings, and chats within existing permissions. The rollout starts today in Copilot Cowork and Copilot Studio, with Word, Excel, PowerPoint, and Chat following in phases over the coming week.

  18. Grok BotOfficialAI score36

    Grok Bot can now suggest ways to help proactively

    AIGrok Bot can now offer suggestions for ways to help without the user needing to ask first. The post does not provide further details about how these proactive suggestions work.

    Video from @bot's post
  19. Google GemmaOfficialAI score54

    Google Gemma credits StudentBench study comparing AI and expert human GRE tutors

    AIGoogle Gemma relays a StudentBench study reporting that AI tutors matched expert human tutors on immediate GRE learning gains. The author reports 2,383 students and a cost of 7 cents per AI tutor hour versus $75 for an expert human hour. The post also says the top AI tutor beat expert human tutors on average in 5 of 7 academic topics, and that the data and paper are publicly available.