Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 7

Oct 7Wed
  1. Leandro von WerraXAI score36

    Snorkel expands Open Benchmarks Grants to $30M for AI evaluation

    AISnorkel AI is expanding its Open Benchmarks Grants tenfold to a $30M commitment to fund more diverse, robust, and continuously updated open AI benchmarks. The program adds an Open Benchmarks Red Team to test and strengthen those benchmarks, plus a Snorkel Research Fellowship for independent researchers developing new evaluation methods. The source says OBG-funded benchmarks have appeared on model cards from every major frontier lab.

  2. 🚨 AI News | TestingCatalogXAI score34

    Microsoft brings hybrid local-cloud intelligence to Copilot for Windows

    AIMicrosoft is adding hybrid intelligence to Copilot for Windows, letting it use local PC context and local models for tasks. Per Satya Nadella's quoted post, Windows will route each task to local or cloud models, and Copilot will act on the user's behalf only with permission.

    Video from @testingcatalog's post
  3. Ethan MollickXAI score58

    Ethan Mollick Tries Intelligent UI in ChatGPT, Finds It Beats Text Walls

    AIEthan Mollick had early access to Intelligent UI and found it a welcome change from long blocks of text. He suggests interfaces will increasingly be built on demand for each user's problem. The quoted OpenAI post says GPT-6 and Intelligent UI are rolling out in ChatGPT for everyone, delivering fast, interactive answers with visual explanations and task tools.

  4. OpenRouterOfficialAI score46

    Perplexity Decider v1.1 arrives on OpenRouter with free output

    AIPerplexity's open-weights multimodal decision model, Decider v1.1, is now available on OpenRouter. It accepts text, JSON, or images and returns typed answers with probabilities, priced at $0.02 per million input tokens with free output. Perplexity says it scores highest on Hugging Face's new Decision Index 0.3 benchmark and costs half as much as v1.

  5. Simon WillisonBlogAI score62

    Anthropic releases Claude Haiku 5.5, priced like GPT-6 Luna up to 100,000 tokens

    AIAnthropic has released Claude Haiku 5.5, priced at $0.10 input and $0.50 output per million tokens up to 100,000 tokens, matching GPT-6 Luna. Beyond 100,000 tokens the price rises to $0.50 and $2.50, and the author found the new tokenizer uses about 1.25x as many tokens as Haiku 4.5 on the same long prompt. The model cannot disable reasoning and defaults to medium effort.

  6. elvisXAI score22

    Envato launches Burst Mode for AI image direction exploration

    AIEnvato has launched Burst Mode, which starts from a rough idea or reference image and generates up to six image directions side by side. Users can steer those directions with moodboards and a creativity control ranging from focused to wild. Envato says it can produce up to six directions at once, up to 10× faster, for one AI credit.

  7. GitHubOfficialAI score40

    Claude Haiku 5.5 is now generally available in GitHub Copilot.

    AIAnthropic's Claude Haiku 5.5 is now generally available in GitHub Copilot, a lightweight model built for fast, high-volume work such as subagents, quick edits, and terminal tasks. GitHub's early testing found it matched Claude Sonnet 5 on many coding tasks while using significantly fewer tokens and steps. It can be used in the GitHub Copilot app, CLI, or @code.

  8. MarkTechPostNewsAI score67

    Anthropic releases Claude Haiku 5.5, a small model with 1M context

    AIAnthropic has released Claude Haiku 5.5, its cheapest and fastest small model, priced at $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100K tokens. It keeps a 1M token context window, up to 128K output tokens, and is generally available on the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Anthropic reports 72.4% on OSWorld 2.1 (offline subset) versus 15.7% for Haiku 4.5, and the article notes that non-default temperature, top_p or top_k values return a 400 error.

  9. ChatGPTOfficialAI score13

    ChatGPT's Intelligent UI turns requests into visual, interactive outputs

    AIOpenAI's ChatGPT Intelligent UI can produce visual outputs for requests such as breaking down a bicycle's design, planning a dinner, building a wardrobe, or splitting a dinner bill from a receipt. The post says these visual responses are easier to understand and more enjoyable to use.

    Image from @ChatGPT's post
  10. ChatGPTOfficialAI score85

    GPT-6 with Intelligent UI rolls out globally in ChatGPT, Free and Go tiers next day

    AIGPT-6 with Intelligent UI begins rolling out globally in the ChatGPT Chat tab for Plus, Pro, Business, and Enterprise users today. The rollout expands to Free and Go tiers starting tomorrow. Plus, Pro, Business, and Enterprise get GPT-6 Sol, while Free and Go get GPT-6 Luna.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  11. ChatGPTOfficialAI score72

    ChatGPT introduces Intelligent UI for interactive visual answers

    AIOpenAI's official ChatGPT account announced Intelligent UI, which lets ChatGPT answer with fully interactive user interfaces. The post says this makes complex topics easier to learn and lets ChatGPT quickly create tools for tasks in the moment. It also states that GPT-6 is coming to ChatGPT for everyone, without giving a date.

    Video from @ChatGPT's post
  12. falOfficialAI score13

    Fal launches Vidu Q4 image-to-video and reference-to-video models

    AIFal announced two Vidu Q4 video models, one for image-to-video and one for reference-to-video, now available to try on its platform. The post provides links to both model pages but no pricing, specs, or benchmark details.

  13. falOfficialAI score34

    Vidu Q4 video model now live on fal with native audio

    AIfal has launched Vidu Q4, offering image-to-video generation from a single first frame with native audio. Reference-to-video supports up to 12 reference images and 3 voice clips for consistent characters and voices. Clips run 3 to 16 seconds at resolutions from 540p up to 4K.

    Video from @fal's post
  14. OpenRouterOfficialAI score46

    GPT-6 Luna Decisions API now available on OpenRouter

    AIOpenAI's GPT-6 Luna Decisions model is now available on OpenRouter, letting apps choose the right model, tool, or action from text, JSON, or images. It returns typed answers with probabilities, costs $0.10 per million input tokens, has a 1M-token context window, and output is free.

  15. Google ResearchOfficialAI score62

    Google Research finds AI boosts patent drafting but junior lawyers' gains vanish without it

    AIA Google Research field experiment with 133 patent lawyers found AI tool access raised drafting scores by 0.34 to 0.38 standard deviations over three months. When the tool was removed for a redlining task, only senior lawyers kept an advantage of 0.45 SD, while junior lawyers showed no discernible improvement. The authors argue that tools which boost current output must not stop junior professionals from building the judgment that senior experts rely on.

    Why it matters: The field experiment separates AI's short-term productivity gains from skill retained after the tool is removed, which matters for training junior professionals.

  16. SantiagoXAI score32

    Envato's Burst mode returns six image directions from one idea

    AIEnvato's new Burst mode turns a single idea into six image directions at once, for one AI credit, which the company says is up to 10× faster. Users can then refine a direction with "more like this" or pick a specific style with "polish."

  17. OpenAI DevelopersOfficialAI score10

    OpenAI launches a rebuilt plugin submission flow for reviews

    AIOpenAI has rebuilt its plugin submission flow to provide clearer review feedback to developers. Developers can submit their plugins for review through the submission page on the OpenAI developer site.

  18. OpenAI DevelopersOfficialAI score22

    OpenAI showcases companies building with plugin extensions for ChatGPT

    AIOpenAI highlighted eight companies building with plugin extensions: Adobe, Canva, Figma, Shopify, Instacart, Atlassian, tldraw, and MagicPath AI. The post lists the partners but gives no details on the plugins' features or capabilities.

    Image from @OpenAIDevs's post
  19. OpenAI DevelopersOfficialAI score38

    OpenAI adds richer interactive plugin extensions to ChatGPT

    AIOpenAI says developers can now build richer, more interactive ChatGPT plugins using plugin extensions. The extensions support launching plugins from the sidebar, @ mentions, and file viewers, with a walkthrough shared alongside the post.

    Video from @OpenAIDevs's post
  20. GitHub Copilot ChangelogOfficialAI score38

    Claude Haiku 5.5 is now generally available in GitHub Copilot

    AIAnthropic's lightweight Claude Haiku 5.5 is now generally available in GitHub Copilot for fast, high-volume tasks such as subagents, quick edits, and terminal work. In early testing, it matched Claude Sonnet 5 on many coding tasks while using significantly fewer tokens and steps. The model is billed at provider list pricing under usage-based billing and is available to Copilot Pro, Pro+, Max, Business, and Enterprise users.

  21. Amjad MasadXAI score40

    Replit building desktop app with Microsoft and Nvidia OpenShell

    AIReplit is building a powerful desktop app with a focus on security and reliability, citing supply-chain attacks and catastrophic agent mistakes as risks of desktop AI apps. The company is partnering with Microsoft and will be an early adopter of Nvidia's OpenShell. A quoted Replit post says the desktop preview runs builds locally on Windows, with each build in its own sandbox powered by Microsoft Execution Containers and OpenShell, and offers a waitlist.

  22. ClaudeDevsOfficialAI score43

    Anthropic adds computer and browser use toolsets to Claude SDKs

    AIAnthropic's Python and TypeScript SDKs now include built-in computer use and browser use toolsets for Claude. The SDKs run the agent loop and send actions to drivers, replacing the custom loop developers previously had to write to map clicks and keystrokes to commands.

    Video from @ClaudeDevs's post
  23. 👩‍💻 Paige BaileyXAI score31

    Ben Affleck says he has seen Google and OpenAI video models

    AIBen Affleck says he writes Python, understands convolutional neural networks, and has worked extensively with GPUs. He also says he used his celebrity status to get private looks at Google and OpenAI's video models. The post asks whether he uses Veo, Omni, or Nano Banana, but the source does not confirm any of these.