Skip to content
TodayOct 8Thu65 items
  1. Pandaily57

    Shanghai AI Lab Open-Sources Intern-Decision Small Models for Structured Decisions

    Shanghai AI Lab has open-sourced Intern-Decision, a family of 0.8B, 2B and 4B parameter models that return structured decisions with probabilities instead of free text. The developers self-report that the 4B model averages 90.02% accuracy across seven test suites, ahead of a commercial reference model at 88.74%, with about 44 milliseconds of local latency on a single RTX 4090. Weights are on Hugging Face, and MetaX says the models run on its hardware from launch.

  2. 雷峰网 Leiphone58

    Alibaba's Qwen Roadmap Targets 5T to 10T Parameters Amid Self-Improving Model Work

    At the Apsara Conference, Alibaba's Qwen team outlined a roadmap of Qwen4 followed by Qwen4.5 and Qwen5, aiming for 5T to 10T parameters. The article notes that Qwen3.8 reached 2.4T parameters and that Qwen3.8-Flash activates 6B parameters per inference while cutting training cost to one-ninth. It also describes Qwen3.8-Max running model-driven experiments in chip design and inference optimization, and multimodal updates including a video model slated for November.

  3. 量子位 QbitAI52

    Claude Haiku 5.5 launches with higher benchmark scores and new migration requirements

    Anthropic released Claude Haiku 5.5, which the article says outperforms DeepSeek V4.1 Flash and GLM-5.3-Flash on official benchmarks and matches GPT-6 Luna on price. On OSWorld 2.1, its Low effort tier scores 42.0% at $0.07 per task, versus 15.7% at $1.45 for Haiku 4.5 at Max. Migrating from Haiku 4.5 requires changes to thinking configuration, sampling parameters, assistant prefill, and the computer-use tool version.

  4. 量子位 QbitAI80

    GPT-6 rolls out to free ChatGPT users with interactive answer interfaces

    OpenAI began rolling out GPT-6 to free and Go ChatGPT users on October 8, replacing GPT-5.6 Luna with GPT-6 Luna, while paid users receive GPT-6 Sol. The update adds Intelligent UI, which generates charts, buttons, and interactive tools inside chat answers. OpenAI's safety report shows gains on jailbreak and instruction-hierarchy tests but also regressions in some self-harm, sexual, and emotional-dependence evaluations, including for under-18 users.

  5. SiliconANGLE · AI60

    ChatGPT's GPT-6 Intelligent UI replaces text walls with charts and tappable elements

    OpenAI says its GPT-6 models can generate visual, interactive answers in ChatGPT, such as charts, forms, and tappable buttons, when the model judges they help. Basic questions stay text-only, while users can request an interactive response at any time. The Intelligent UI is available now to Plus, Pro, Business, and Enterprise subscribers, with Free and Go users getting access the next day.

  6. The Verge · AI72

    ChatGPT's Intelligent UI adds interactive charts, diagrams, and tools to answers

    OpenAI is rolling out an Intelligent UI feature in ChatGPT that lets answers combine text with diagrams, charts, forms, and tappable buttons. It is available starting today to Plus, Pro, Business, and Enterprise users, and will expand to Go and free tiers on Thursday. Higher-tier users get the mid-range GPT-6 Sol model, while Go and free users get GPT-6 Luna.

  7. Xiaomi MiMo63

    Xiaomi releases MiMo-V2.5-TTS series of speech synthesis models

    Xiaomi released the MiMo-V2.5-TTS Series, three speech synthesis models for stock voices, voice design, and voice cloning. The models accept natural-language style instructions and inline audio tags, and the source says the three models are free of charge for a limited time on the Xiaomi MiMo API platform. Xiaomi also open-sourced integration Skills for agent applications on GitHub.

    Why it matters: The release shows how a TTS family adds style instructions, inline audio tags, and voice design or cloning to speech synthesis, which matters for agent and creative workflows.

  8. Xiaomi MiMo44

    Xiaomi releases open-source MiMo-V2.5-ASR speech recognition model with dialect support

    Xiaomi MiMo has released MiMo-V2.5-ASR, an open-source speech recognition model that the company says achieves state-of-the-art results across multiple benchmarks. The model supports bilingual Chinese–English recognition, Chinese dialects such as Wu, Cantonese, Hokkien, and Sichuanese, code-switching, and lyrics transcription. It is also designed to handle noisy environments and multi-speaker conversations.

  9. Testing Catalog36

    Gemini Agent for Business may add Claude Opus 5 and Sonnet 5.5

    Google's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.

  10. Google Research22

    Today at 12:00pm, don't miss our live demonstration on EmbeddingGemma 2 at the @COLM_conf Google booth (#107). Connect with Sahil Dua and explore an open multimodal model that unifies text, images, audio, and video representations. #COLM2026 @GoogleDeepMind

    Today at 12:00pm, don't miss our live demonstration on EmbeddingGemma 2 at the @COLM_conf Google booth (#107). Connect with Sahil Dua and explore an open multimodal model that unifies text, images, audio, and video representations. #COLM2026 @GoogleDeepMind

  11. Testing Catalog50

    OPENAI 🔥: GPT-6.1 Sol Ultrafast is rolling out on ChatGPT Work, Codex, and the API. > GPT-6.1 Sol in Ultrafast mode is $12 per million input tokens and $60 per million output tokens. > Ultrafast mode delivers 8x faster speeds than Sol Standard. Ultrafast testing time 👀 https://x.com/OpenAIDevs/status/2108262812489531498/video/1

    OPENAI 🔥: GPT-6.1 Sol Ultrafast is rolling out on ChatGPT Work, Codex, and the API. > GPT-6.1 Sol in Ultrafast mode is $12 per million input tokens and $60 per million output tokens. > Ultrafast mode delivers 8x faster speeds than Sol Standard. Ultrafast testing time 👀 https://x.com/OpenAIDevs/status/2108262812489531498/video/1

  12. OpenAI Developers38

    API pricing for GPT-6.1 Sol in Ultrafast mode is $12 per million input tokens and $60 per million output tokens. Built for work where speed and intelligence make a difference: debugging an outage, agents navigating apps, and live experiences where every second counts.

    API pricing for GPT-6.1 Sol in Ultrafast mode is $12 per million input tokens and $60 per million output tokens. Built for work where speed and intelligence make a difference: debugging an outage, agents navigating apps, and live experiences where every second counts.

  13. OpenAI · YouTube70

    OpenAI adds Intelligent UI to GPT-6 in ChatGPT Chat tab

    OpenAI introduced Intelligent UI for GPT-6 in ChatGPT, which lets the chatbot answer with fully interactive interfaces and quickly build tools for a task. The feature rolled out globally to Plus, Pro, Business, and Enterprise tiers in the Chat tab and expands to Free and Go tiers, with Enterprise access depending on workplace admin settings.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  14. OpenAI · YouTube72

    OpenAI rolls out GPT-6 with Intelligent UI in ChatGPT Chat

    OpenAI says GPT-6 in ChatGPT adds Intelligent UI, which lets ChatGPT answer with interactive interfaces and build quick tools for a task. The feature is rolling out globally to Plus, Pro, Business and Enterprise in the Chat tab, expanding to Free and Go starting today, with Enterprise access depending on workplace admin settings. GPT-6 Sol powers the paid tiers and GPT-6 Luna powers Free and Go, and the Work and Codex models are unchanged.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  15. OpenAI · YouTube67

    OpenAI launches GPT-6 Intelligent UI for interactive ChatGPT answers

    OpenAI's GPT-6 in ChatGPT adds Intelligent UI, which lets ChatGPT answer with interactive interfaces and build quick tools for a task. The feature is rolling out to Plus, Pro, Business, and Enterprise first, with Free and Go tiers following, and Enterprise access depends on workplace admin settings. The update covers only the Chat experience, and the models powering Work and Codex are not changing.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  16. Elvis Saravia42

    Huge release from @odysseyml. Odyssey-3 Pro sets a new top score on Physics-IQ Verified, a benchmark that asks models to continue videos of real physics experiments. The robotics results stood out to me. With tens of hours of demos, the robot arm recovered from a missed grasp, a behavior that never appeared in those demos.

    Huge release from @odysseyml. Odyssey-3 Pro sets a new top score on Physics-IQ Verified, a benchmark that asks models to continue videos of real physics experiments. The robotics results stood out to me. With tens of hours of demos, the robot arm recovered from a missed grasp, a behavior that never appeared in those demos.

  17. Santiago46

    The #1 video-to-video model in the Physics-IQ Verified benchmark is finally live! (Their research preview is) Odyssey 3 Pro is a world model, and nobody beats it for physical accuracy. You can use this model to control a robot, drive a car, play a video game, or pilot a drone. • It takes visual observations from the world • Uses these observations to learn how things work • Then maps that knowledge to the system's physical controls

    The #1 video-to-video model in the Physics-IQ Verified benchmark is finally live! (Their research preview is) Odyssey 3 Pro is a world model, and nobody beats it for physical accuracy. You can use this model to control a robot, drive a car, play a video game, or pilot a drone. • It takes visual observations from the world • Uses these observations to learn how things work • Then maps that knowledge to the system's physical controls

  18. Odyssey31

    We believe world models will power increasingly capable physical AI, generate environments to train intelligences, and enable new kinds of human experiences, and we believe Odyssey-3 is a big leap towards this. Experience Odyssey-3 today! https://odyssey.systems/meet-odyssey-3

    We believe world models will power increasingly capable physical AI, generate environments to train intelligences, and enable new kinds of human experiences, and we believe Odyssey-3 is a big leap towards this. Experience Odyssey-3 today! https://odyssey.systems/meet-odyssey-3

  19. Odyssey34

    Odyssey-3’s learned world knowledge can also be applied to physical systems, enabling physical AI developers to adapt Odyssey-3 to control robots, power humanoids, drive cars, fly drones, and other autonomous machines.

    Odyssey-3’s learned world knowledge can also be applied to physical systems, enabling physical AI developers to adapt Odyssey-3 to control robots, power humanoids, drive cars, fly drones, and other autonomous machines.

  20. Odyssey38

    Odyssey-3 is a foundation world model, enabling many applications in physical AI, human experiences, and even how we train intelligences. We're particularly excited by agents learning from experience inside Odyssey-3, working to accomplish objectives.

    Odyssey-3 is a foundation world model, enabling many applications in physical AI, human experiences, and even how we train intelligences. We're particularly excited by agents learning from experience inside Odyssey-3, working to accomplish objectives.

  21. Odyssey40

    Today we're launching Odyssey-3, the most powerful foundation world model yet. It sets a new state of the art on Physics-IQ, and powers robots, trains AIs, and generates interactive experiences. It's really cool. Experience the model today, all for free!

    Today we're launching Odyssey-3, the most powerful foundation world model yet. It sets a new state of the art on Physics-IQ, and powers robots, trains AIs, and generates interactive experiences. It's really cool. Experience the model today, all for free!

  22. Odyssey42

    Odyssey-3 can generate interactive environments from a prompt, all in real time. It's a new kind of world simulator, and learns representations of physics, dynamics, and cause-and-effect from a broad dataset of visual observations.

    Odyssey-3 can generate interactive environments from a prompt, all in real time. It's a new kind of world simulator, and learns representations of physics, dynamics, and cause-and-effect from a broad dataset of visual observations.

  23. MarkTechPost58

    JetBrains releases Mellum2.1, a 12B MoE open model for coding agents

    JetBrains has released Mellum2.1, a 12B mixture-of-experts thinking model with 2.5B active parameters, under Apache 2.0 on Hugging Face. Post-training reinforcement learning in real software repositories raised SWE-bench Verified from 2.0 to 47.0, according to JetBrains' self-reported results. Qwen3.5-9B still leads on SWE-bench Pro, GPQA Diamond and AIME, and GGUF builds start at 7.0 GB for local use.

  24. Leandro von Werra70

    Carbon-A open model and database predict 566 million gene candidates across 22,617 species

    Carbon-A is an open model that predicts gene locations directly from DNA, and it has been used to annotate genomes from over 22,000 species. The release includes a database of 566 million gene candidates, about 16 times the gene annotations in the RefSeq dataset. Wet-lab RNA experiments supported 239 candidates missing from RefSeq across cats, Syrian hamsters, chickens, and Arabidopsis.

    Why it matters: The source ties an open gene-annotation model to specific wet-lab checks and gene counts, helping readers judge how far its predictions extend beyond well-studied genomes.

  25. Thomas Wolf62

    Carbon-A open model finds 566 million candidate genes across 22,617 species

    The team released Carbon-A, an open model that finds genes directly in DNA, along with a database of 566.34 million candidate genes across 22,617 species. The model reads genomes without needing a close relative, and wet-lab validation in cats, chickens, and arabidopsis is cited, with 239 genes found missing from reference annotations of common species. The authors say the model marks gene locations but does not design DNA or predict gene function.