Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 4

Oct 4Sun
  1. Teknium 🪽XAI score20

    Teknium posts two eyes emojis, teasing an unexplained Hermes-related announcement.

    AITeknium, a researcher associated with the Hermes model family, posted only two eye emojis with no explanation of the main post's content. The post appears to be a teaser linked to a quoted post from @alexhvnsen describing a Hermes "Alan's way" companion app with Telegram-based control, a macOS and VM hybrid setup, and a proactive lead bot.

  2. Jerry LiuXAI score34

    LlamaIndex launches Extract v2.5 document extraction agents, cutting errors on scanned forms

    AILlamaIndex introduced Extract v2.5, a series of agents tuned for document extraction, including cost-effective, agentic, and agentic plus tiers, available in LlamaParse. The company says the agents reduce error rates by 2x or more compared with frontier models at a small fraction of the price, and they handle handwritten and drawn annotations on scanned documents while grounding values in the source text.

    Video from @jerryjliu0's post
  3. Teknium 🪽XAI score29

    Teknium says ESP32 hardware is now set up at home

    AITeknium announced that an ESP32 is now running at home, with no further technical details given in the post. The post is a brief update that quotes an @adolandev post about Hermes Gadget, an open SDK for a small device that speaks to a user's own Hermes model and can be tested with a desktop simulator.

Oct 3

Oct 3Sat
  1. Harrison ChaseXAI score28

    LangChain's ModelRouterMiddleware routes runs to models using Jev

    AILangChain's ModelRouterMiddleware uses Jev to read the first message and select a model that handles the entire run. Because Jev is cheap, developers can also re-select a model after each tool result using a custom hook. The router was demonstrated in a quick project by @dbreunig built with DSPy and Jev.

  2. ClineOfficialAI score14

    Cline launches a new desktop app

    AICline announces a new desktop app, available at cline.bot/desktop. The post gives no further details about features, platforms, or pricing.

  3. Aidan GomezXAI score46

    AlephAlpha releases Kolibri, a German-English model with a technical report

    AIAlephAlpha has released Kolibri, a German-English model with 78B total parameters, 3.46B active, and context up to 1M tokens. The weights are available under Apache 2.0 for running on users' own hardware. Cohere's Aidan Gomez congratulated the team on the model and its detailed technical report.

  4. TechRadar · AINewsAI score25

    Emergn study says up to 23% of UK senior leaders may overstate their AI knowledge

    AIEmergn research claims as many as one in four (23%) UK senior leaders may be bluffing about their AI knowledge. The study says 38% believe their career prospects could suffer unless they significantly improve their AI skills over the next year. Only 43% say they have received substantial AI training in the past 12 months, and Emergn calls for more relevant, human-centric training.

  5. whXAI score3

    Proximal Hiring for Post-Training, Data Research, and Large-Scale Infra

    AIProximal is recruiting for roles in post-training, data research, and infrastructure at tremendous scale. The post invites people interested in these areas, or in learning more about the company's work, to reach out. A related post notes that a researcher has joined Proximal in San Francisco and sees many open problems in data and post-training.

Oct 2

Oct 2Fri
  1. Jerry LiuXAI score34

    LlamaIndex's Extract v2.5 agents reason over tables spanning multiple pages

    AILlamaIndex introduced Extract v2.5, a set of document extraction agents that can reconstruct records split across pages and assemble them with thousands of other cells into structured tabular output. The post says the agents handle real-world documents like insurance claims, regulatory filings, and legal schedules, where a record may start on one page and finish on the next. The accompanying background post claims record-spanning-page accuracy rose from 85.5% to 96.5%, and that the agentic tier outperforms Opus 5.5 and GPT-6 Sol at 30% to 4x lower cost.

    Video from @jerryjliu0's post
  2. InferactOfficialAI score12

    vLLM Toronto meetup to cover project's next direction

    AIInferact invites developers to a free vLLM Meetup in Toronto, co-hosted with Cohere, where co-founder and lead maintainer Roger Wang will discuss where the vLLM project is heading next. Engineers from Cohere and NVIDIA are also scheduled to speak, and spots are limited, with registration through a link in the thread.

  3. MuseOfficialAI score18

    Muse launches developer portal as connector submissions pass 3,000

    AIMuse has released a developer portal to simplify connector intake, letting developers submit, manage, and track their connectors. The company says it crossed 3,000 connector submissions on its Connector Platform, and it hopes the portal will help drive the next 3,000.

  4. PikaOfficialAI score31

    Pika redesigns video creation with Video Studio and direct model access

    AIPika has redesigned its video creation experience, offering Video Studio for creators who want to focus on craft and direct generation through specific models for those with model preferences. A creator notes that Seedance on the platform costs about half of what they pay elsewhere, with no locked top-tier model behind a higher plan.

  5. Guillermo RauchXAI score34

    Muse Ships Open-Source ESP32 Firmware and Linux SDK for Gadgets

    AIMuse has released Muse Gadgets, an open-source ESP32 firmware and Linux SDK for building hardware devices that work with Muse. Developers can obtain an API token from gadgets.muse.ai and use a coding agent with the GitHub repo to build peripherals. Guillermo Rauch praised the team's rapid shipping.

  6. SGLangOfficialAI score38

    SGLang v0.5.20 adds Intel XPU support and faster RL rollouts

    AISGLang has released v0.5.20, bringing Intel XPU into standard releases alongside RL sampling masks that make rollouts more reliable with up to 52% faster decode. The update also adds Unified Radix Tree SWA branching-point caching, which the project says lifts cache hit rate about 20 points and cuts TTFT by roughly one-third, plus up to 12.5× faster ROCm model loading. New models named in the release include GLM-5.3-Flash, Qwen3.8-Flash-Next, K2 Horizon, Hy4-Preview, FastH3, and VDN-H3.

  7. SGLangOfficialAI score39

    SGLang adds a scoring API and multi-item scoring for decision models

    AISGLang's update adds a /v1/score endpoint that returns scores for requested labels such as Yes/No or A/B/C, avoiding the label loss of generate with top-k logprobs. Its multi-item scoring computes shared context once and keeps each candidate isolated, with 16-candidate p95 on Qwen3-8B dropping from 54.1 ms (Generate) to 20.6 ms.

  8. SGLangOfficialAI score28

    SGLang's /v1/decisions API turns Qwen3.8-27B into a decision model

    AISGLang demonstrated Qwen3.8-27B as a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. The company says its native /v1/decisions API lets LLMs and VLMs be used for classification and scoring. It also announced /v1/systemone for running Jev-like open models with the TypeSafe SDK.

  9. SGLangOfficialAI score58

    SGLang v0.5.21 adds native decisions API and new model support

    AISGLang has released v0.5.21 with a native Decisions API that turns an LLM or VLM into a low-latency classifier and scorer. The release also lets /v1/score rerank search or RAG results in one call, lets PD instances switch between prefill and decode without restarting, and adds support for models including DeepSeek-V4.1 Flash, Kimi K3, and GLM-5.3-Flash on AMD MI355X. The announcement reports a 22% faster first token on long prompts for DeepSeek-V4.1 Flash and 20.6% higher prefill throughput for Kimi K3 in PD serving.

    Image from @sgl_project's post
  10. LiveKitOfficialAI score23

    AssemblyAI Universal 3.6 Pro now live in LiveKit Inference

    AIAssemblyAI's Universal 3.6 Pro speech-to-text model is now available in LiveKit Inference, with 45% fewer wrong yes/no confirmations and about 30% less background speech transcribed. It supports 32 languages plus code-switching and endpointing that waits out phone numbers and emails, at the same $0.45/hr price, accessible by switching to universal-3-6-pro.

    Image from @livekit's post
  11. Sara HookerXAI score26

    Adaption Labs makes its Invent dataset tool available via API

    AIAdaption Labs has made Invent, its tool for generating AI training datasets from a plain-language description, available through an API. Developers can reportedly produce AI-ready training datasets in minutes with a few lines of code, according to the quoted post. Documentation is available at docs.adaptionlabs.ai.

    Image from @sarahookr's post