Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 4

Oct 4Sun
  1. Teknium 🪽XAI score20

    Teknium posts two eyes emojis, teasing an unexplained Hermes-related announcement.

    AITeknium, a researcher associated with the Hermes model family, posted only two eye emojis with no explanation of the main post's content. The post appears to be a teaser linked to a quoted post from @alexhvnsen describing a Hermes "Alan's way" companion app with Telegram-based control, a macOS and VM hybrid setup, and a proactive lead bot.

  2. Jerry LiuXAI score34

    LlamaIndex launches Extract v2.5 document extraction agents, cutting errors on scanned forms

    AILlamaIndex introduced Extract v2.5, a series of agents tuned for document extraction, including cost-effective, agentic, and agentic plus tiers, available in LlamaParse. The company says the agents reduce error rates by 2x or more compared with frontier models at a small fraction of the price, and they handle handwritten and drawn annotations on scanned documents while grounding values in the source text.

    Video from @jerryjliu0's post
  3. Teknium 🪽XAI score29

    Teknium says ESP32 hardware is now set up at home

    AITeknium announced that an ESP32 is now running at home, with no further technical details given in the post. The post is a brief update that quotes an @adolandev post about Hermes Gadget, an open SDK for a small device that speaks to a user's own Hermes model and can be tested with a desktop simulator.

Oct 3

Oct 3Sat
  1. Harrison ChaseXAI score28

    LangChain's ModelRouterMiddleware routes runs to models using Jev

    AILangChain's ModelRouterMiddleware uses Jev to read the first message and select a model that handles the entire run. Because Jev is cheap, developers can also re-select a model after each tool result using a custom hook. The router was demonstrated in a quick project by @dbreunig built with DSPy and Jev.

  2. Aidan GomezXAI score46

    AlephAlpha releases Kolibri, a German-English model with a technical report

    AIAlephAlpha has released Kolibri, a German-English model with 78B total parameters, 3.46B active, and context up to 1M tokens. The weights are available under Apache 2.0 for running on users' own hardware. Cohere's Aidan Gomez congratulated the team on the model and its detailed technical report.

  3. TechRadar · AINewsAI score25

    Emergn study says up to 23% of UK senior leaders may overstate their AI knowledge

    AIEmergn research claims as many as one in four (23%) UK senior leaders may be bluffing about their AI knowledge. The study says 38% believe their career prospects could suffer unless they significantly improve their AI skills over the next year. Only 43% say they have received substantial AI training in the past 12 months, and Emergn calls for more relevant, human-centric training.

Oct 2

Oct 2Fri
  1. Jerry LiuXAI score34

    LlamaIndex's Extract v2.5 agents reason over tables spanning multiple pages

    AILlamaIndex introduced Extract v2.5, a set of document extraction agents that can reconstruct records split across pages and assemble them with thousands of other cells into structured tabular output. The post says the agents handle real-world documents like insurance claims, regulatory filings, and legal schedules, where a record may start on one page and finish on the next. The accompanying background post claims record-spanning-page accuracy rose from 85.5% to 96.5%, and that the agentic tier outperforms Opus 5.5 and GPT-6 Sol at 30% to 4x lower cost.

    Video from @jerryjliu0's post
  2. PikaOfficialAI score31

    Pika redesigns video creation with Video Studio and direct model access

    AIPika has redesigned its video creation experience, offering Video Studio for creators who want to focus on craft and direct generation through specific models for those with model preferences. A creator notes that Seedance on the platform costs about half of what they pay elsewhere, with no locked top-tier model behind a higher plan.

  3. Guillermo RauchXAI score34

    Muse Ships Open-Source ESP32 Firmware and Linux SDK for Gadgets

    AIMuse has released Muse Gadgets, an open-source ESP32 firmware and Linux SDK for building hardware devices that work with Muse. Developers can obtain an API token from gadgets.muse.ai and use a coding agent with the GitHub repo to build peripherals. Guillermo Rauch praised the team's rapid shipping.

  4. SGLangOfficialAI score38

    SGLang v0.5.20 adds Intel XPU support and faster RL rollouts

    AISGLang has released v0.5.20, bringing Intel XPU into standard releases alongside RL sampling masks that make rollouts more reliable with up to 52% faster decode. The update also adds Unified Radix Tree SWA branching-point caching, which the project says lifts cache hit rate about 20 points and cuts TTFT by roughly one-third, plus up to 12.5× faster ROCm model loading. New models named in the release include GLM-5.3-Flash, Qwen3.8-Flash-Next, K2 Horizon, Hy4-Preview, FastH3, and VDN-H3.

  5. SGLangOfficialAI score39

    SGLang adds a scoring API and multi-item scoring for decision models

    AISGLang's update adds a /v1/score endpoint that returns scores for requested labels such as Yes/No or A/B/C, avoiding the label loss of generate with top-k logprobs. Its multi-item scoring computes shared context once and keeps each candidate isolated, with 16-candidate p95 on Qwen3-8B dropping from 54.1 ms (Generate) to 20.6 ms.

  6. SGLangOfficialAI score28

    SGLang's /v1/decisions API turns Qwen3.8-27B into a decision model

    AISGLang demonstrated Qwen3.8-27B as a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. The company says its native /v1/decisions API lets LLMs and VLMs be used for classification and scoring. It also announced /v1/systemone for running Jev-like open models with the TypeSafe SDK.

  7. SGLangOfficialAI score58

    SGLang v0.5.21 adds native decisions API and new model support

    AISGLang has released v0.5.21 with a native Decisions API that turns an LLM or VLM into a low-latency classifier and scorer. The release also lets /v1/score rerank search or RAG results in one call, lets PD instances switch between prefill and decode without restarting, and adds support for models including DeepSeek-V4.1 Flash, Kimi K3, and GLM-5.3-Flash on AMD MI355X. The announcement reports a 22% faster first token on long prompts for DeepSeek-V4.1 Flash and 20.6% higher prefill throughput for Kimi K3 in PD serving.

    Image from @sgl_project's post
  8. LiveKitOfficialAI score23

    AssemblyAI Universal 3.6 Pro now live in LiveKit Inference

    AIAssemblyAI's Universal 3.6 Pro speech-to-text model is now available in LiveKit Inference, with 45% fewer wrong yes/no confirmations and about 30% less background speech transcribed. It supports 32 languages plus code-switching and endpointing that waits out phone numbers and emails, at the same $0.45/hr price, accessible by switching to universal-3-6-pro.

    Image from @livekit's post
  9. Sara HookerXAI score26

    Adaption Labs makes its Invent dataset tool available via API

    AIAdaption Labs has made Invent, its tool for generating AI training datasets from a plain-language description, available through an API. Developers can reportedly produce AI-ready training datasets in minutes with a few lines of code, according to the quoted post. Documentation is available at docs.adaptionlabs.ai.

    Image from @sarahookr's post
  10. ClineOfficialAI score38

    Cline Desktop adds beta Connectors for Gmail, Slack, and more

    AICline Desktop now offers Connectors in beta, letting users link Gmail, Slack, Google Calendar, Linear, Sentry, Notion, and other apps in one click. Once connected, Cline can use these apps' tools to retrieve context and take actions on the user's behalf.

    Video from @cline's post
  11. François CholletXAI score28

    Keras community call outlines pluggable backends and KerasHub updates

    AIKeras is moving to a pluggable backend design, with MLX and PaddlePaddle backends upcoming as add-on libraries. The team is reducing the operations needed to ship new backends and streamlining unit testing so a single harness can test all ops, such as casting consistency. KerasHub also gains many new models and is shifting its preprocessing from tf-text to PyGrain.

  12. Jerry LiuXAI score44

    LlamaIndex Extract v2.5 hits 93–96% on dense table extraction benchmarks

    AILlamaIndex released Extract v2.5, a set of document extraction agents that it says reach 93%–96%+ accuracy on long-list extraction, including records spanning pages. The post claims the agents outperform frontier VLMs, which it says stop early, miss repeated records, and struggle to attribute values to sources, while LlamaIndex attributes every extracted value to its source. The agents are available through LlamaParse.

    Video from @jerryjliu0's post

Oct 1

Oct 1Thu
  1. ComfyUIOfficialAI score34

    YUI, the first animated short film made with Comfy Agent

    AIComfyUI released YUI, described as the first animated short film made with Comfy Agent, a 4:20 film its creator made in three days. Comfy Agent handled shot regeneration, side-by-side model comparisons, and character consistency checks that previously required weeks of manual testing and generation.

  2. Vercel DevelopersOfficialAI score24

    Laya decision model free on Vercel AI Gateway through October 31

    AIVercel says the Laya decision model is available free on AI Gateway through October 31, in partnership with Boundless. The post suggests using it to route agent work, triage support requests, and check guardrails.

  3. Fidji SimoXAI score20

    Fidji Simo welcomes ChronicleBio AI, backed by Morgan Cheatham and Jimi Hendrix

    AIFidji Simo says she is excited to work with Morgan Cheatham and Jimi Hendrix, two early investors who recognized the importance of data in AI biotech. Morgan Cheatham's linked post says the partnership supports ChronicleBio, which is building AI models to identify biologically distinct patient subgroups within complex chronic diseases. The team reportedly combines clinical phenotyping, molecular measurement, and AI, and its early work has surfaced disease subtypes linked to potentially relevant therapies.

  4. Fidji SimoXAI score22

    Fidji Simo's ChronicleBio uses AI to find treatments for chronic diseases

    AIFidji Simo's company ChronicleBio aims to develop treatments for POTS, Long COVID, ME/CFS, and hEDS, which currently have zero FDA-approved therapies. The company plans to build the deepest dataset on these conditions and use AI to identify the biology behind their symptoms. The goal is to match the right drugs to the right patients.

  5. PikaOfficialAI score34

    Pika's Camera Director app reframes shots from new angles

    AIPika has launched Camera Director, an app that reframes any shot to show it from different camera angles while preserving the original content. A quoted post praises its quality and says it is useful for real-world work.

  6. PikaOfficialAI score31

    Pika unveils new AI creative platform with over 40 apps

    AIPika has launched a new AI creative platform offering more than 40 apps covering needs from finishing edits to full video ads. The company describes the product as built for creatives and designed to meet exacting output standards.

  7. ReplicateOfficialAI score46

    Replicate powers Tavus's Griffin, a video Turing test-passing model

    AIReplicate says it is powering Griffin from Tavus on its platform. Tavus describes Griffin as the first model to pass the video Turing test, with 48% of live conversation participants believing it was a real human. The model ranks first on NVIDIA's full-duplex AI video benchmark.

  8. OpenRouter · New modelsBlogAI score36

    Apodex 1.1 Mini Released as Free Reasoning Model for Research Tasks

    AIApodex has released Apodex 1.1 Mini, a free reasoning-first model designed for complex, long-horizon research and forecasting tasks. According to the source, it works directly with files, data, code, and tools to produce verifiable results.

  9. Prime IntellectOfficialAI score32

    Extropic uses Prime Intellect to post-train Qwen3.6-35B-A3B for thermodynamic ML

    AIExtropic post-trained Qwen3.6-35B-A3B with Prime Intellect for thermodynamic ML research, nearly tripling its held-out eval results in about 100 GRPO steps. The team built a custom RL environment with verifiers and trained on Hosted Training, Prime Sandboxes, and Prime Inference. This let Extropic avoid managing multi-node GPU infrastructure and focus on research.

    Image from @PrimeIntellect's post
  10. NewcomerBlogAI score38

    Benchmark Leads Funding Round for Chip Startup Tendrils Compute

    AIBenchmark has won a hot competition to lead a new funding round for early-stage chip startup Tendrils Compute, according to multiple people familiar with the matter. Sources say the company is already discussing a fast follow-up raise that could value it at more than $1 billion. The deal is part of a wave of VC investment in specialized chips, including inference-focused startups such as Etched, which doubled its valuation to $21 billion in August.