Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 3

Oct 3Sat
  1. DatabricksOfficialAI score27

    Databricks Genie One adds ontology, uploads, and scheduled tasks

    AIDatabricks has rolled out a set of updates to Genie One spanning context, data access, collaboration, and automation. Genie Ontology is enabled by default to provide business-aware context, and workspace instructions can apply organizational data conventions to every prompt. Users can also upload Word documents, images, CSVs, spreadsheets, and PDFs, query Unity Catalog tables with schema preview and one-click access requests, and automate recurring work with scheduled tasks that reference past runs.

    Video from @databricks's post
  2. PixVerseOfficialAI score22

    PixVerse launches Short Drama plugin for AI agent video creation

    AIPixVerse has introduced Short Drama, a plugin that lets an AI agent turn a user's scene description into a finished video. The plugin takes text covering characters, setting, action, and mood, and produces video through the agent workflow, giving teams concrete material to review and develop.

    Video from @PixVerse's post

Oct 2

Oct 2Fri
  1. Jerry LiuXAI score34

    LlamaIndex's Extract v2.5 agents reason over tables spanning multiple pages

    AILlamaIndex introduced Extract v2.5, a set of document extraction agents that can reconstruct records split across pages and assemble them with thousands of other cells into structured tabular output. The post says the agents handle real-world documents like insurance claims, regulatory filings, and legal schedules, where a record may start on one page and finish on the next. The accompanying background post claims record-spanning-page accuracy rose from 85.5% to 96.5%, and that the agentic tier outperforms Opus 5.5 and GPT-6 Sol at 30% to 4x lower cost.

    Video from @jerryjliu0's post
  2. NVIDIA AIOfficialAI score33

    Nemotron 3 Diarization tracks overlapping speakers on Hugging Face

    AINVIDIA's Nemotron 3 Diarization model identifies who spoke when, including during overlapping speech, and is now available on Hugging Face. It supports up to eight speakers and has 100M parameters. The post thanks users for downloads and trending activity and shares a follow-up answering community questions.

    Video from @NVIDIAAI's post
  3. Prime IntellectOfficialAI score43

    CMU's SMDD-Bench adds 502 drug design tasks for RL training

    AICMU researchers released SMDD-Bench, a benchmark of 502 small-molecule drug design tasks that use RDKit, ADMET-AI, and Boltz-2 as feedback loops. The authors argue that long-horizon planning, exploration, and learning from imperfect feedback remain open problems beyond math and coding, and the benchmark is available in Prime Intellect's Environments Hub for training with prime-rl.

  4. Replit ⠕OfficialAI score40

    Replit adds interactive charts, new models, and Jev integration

    AIReplit chat now generates interactive charts when users ask Replit Agent to visualize data. Users can also choose GPT-6.1 Sol from OpenAI or Claude Sonnet 5.5 from Anthropic when building with Agent, or stay in auto mode. Jev is available through Replit AI Integrations for classifying content, routing requests, and scoring leads without managing API keys.

    Video from @Replit's post
  5. Prime IntellectOfficialAI score38

    Prime Intellect stores MLA KV cache in NVFP4 for more cached tokens

    AIPrime Intellect compresses the MLA latent KV cache to NVFP4, reducing each row from 576 to 352 bytes. This fits about 50% more cached tokens per decoder compared with FP8. Its native sparse-MLA kernel unpacks the format on-chip, and the company is contributing that kernel to FlashInfer as an experimental operation.

    Image from @PrimeIntellect's post
  6. Prime IntellectOfficialAI score20

    Prime Intellect launches Prime Inference for serving AI model tokens

    AIPrime Intellect has introduced Prime Inference, an inference service it says has served trillions of tokens for reinforcement learning and dedicated customer deployments. The company argues that owning your intelligence requires owning your inference, and the post promises to unpack its inference stack.

    Video from @PrimeIntellect's post
  7. PikaOfficialAI score31

    Pika redesigns video creation with Video Studio and direct model access

    AIPika has redesigned its video creation experience, offering Video Studio for creators who want to focus on craft and direct generation through specific models for those with model preferences. A creator notes that Seedance on the platform costs about half of what they pay elsewhere, with no locked top-tier model behind a higher plan.

  8. Aravind SrinivasXAI score44

    Perplexity Computer builds a 3D map of NYC restaurants

    AIPerplexity's Computer built a 3D map of nearly 26,000 restaurants and cafes across New York City's five boroughs. Users can search by dish or neighborhood and step inside places such as Peter Luger and Grand Central Oyster Bar. The post frames such projects as ones an agent can run for hours to produce something substantial.

  9. Aravind SrinivasXAI score62

    Perplexity open-sources models, an inference engine, and security tools

    AIPerplexity has released several open source projects, including the pplx-decider-v1-27b multimodal decision model, the pplx-embed-v2-context-9b-preview contextual embeddings model, and the Lily local inference engine for Apple silicon. The post also lists the 0.6B on-device PII-Tracer classifier with its PII-TRACE benchmark, the WANDR research agent benchmark, and the Numbat and Bumblebee security tools, and says more open source releases are coming soon.

  10. Claude Code · GitHub ReleasesOfficialAI score38

    Claude Code v2.1.288 is released with fixes and new controls

    AIAnthropic released Claude Code v2.1.288, adding $.ui.selection() for mods, a built-in gh api for cloud sessions without the GitHub CLI, and --max-findings for /code-review. The release also fixes many issues, including mid-response API timeouts, resume and compaction bugs, and auto mode denials and model switching on Bedrock and Mantle.

  11. SGLangOfficialAI score38

    SGLang v0.5.20 adds Intel XPU support and faster RL rollouts

    AISGLang has released v0.5.20, bringing Intel XPU into standard releases alongside RL sampling masks that make rollouts more reliable with up to 52% faster decode. The update also adds Unified Radix Tree SWA branching-point caching, which the project says lifts cache hit rate about 20 points and cuts TTFT by roughly one-third, plus up to 12.5× faster ROCm model loading. New models named in the release include GLM-5.3-Flash, Qwen3.8-Flash-Next, K2 Horizon, Hy4-Preview, FastH3, and VDN-H3.

  12. SGLangOfficialAI score39

    SGLang adds a scoring API and multi-item scoring for decision models

    AISGLang's update adds a /v1/score endpoint that returns scores for requested labels such as Yes/No or A/B/C, avoiding the label loss of generate with top-k logprobs. Its multi-item scoring computes shared context once and keeps each candidate isolated, with 16-candidate p95 on Qwen3-8B dropping from 54.1 ms (Generate) to 20.6 ms.

  13. SGLangOfficialAI score28

    SGLang's /v1/decisions API turns Qwen3.8-27B into a decision model

    AISGLang demonstrated Qwen3.8-27B as a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. The company says its native /v1/decisions API lets LLMs and VLMs be used for classification and scoring. It also announced /v1/systemone for running Jev-like open models with the TypeSafe SDK.

  14. SGLangOfficialAI score58

    SGLang v0.5.21 adds native decisions API and new model support

    AISGLang has released v0.5.21 with a native Decisions API that turns an LLM or VLM into a low-latency classifier and scorer. The release also lets /v1/score rerank search or RAG results in one call, lets PD instances switch between prefill and decode without restarting, and adds support for models including DeepSeek-V4.1 Flash, Kimi K3, and GLM-5.3-Flash on AMD MI355X. The announcement reports a 22% faster first token on long prompts for DeepSeek-V4.1 Flash and 20.6% higher prefill throughput for Kimi K3 in PD serving.

    Image from @sgl_project's post
  15. LiveKitOfficialAI score23

    AssemblyAI Universal 3.6 Pro now live in LiveKit Inference

    AIAssemblyAI's Universal 3.6 Pro speech-to-text model is now available in LiveKit Inference, with 45% fewer wrong yes/no confirmations and about 30% less background speech transcribed. It supports 32 languages plus code-switching and endpointing that waits out phone numbers and emails, at the same $0.45/hr price, accessible by switching to universal-3-6-pro.

    Image from @livekit's post
  16. PyTorch BlogOfficialAI score24

    PyTorch Certified Associate Gets New Four-Module Certification Pathway

    AIThe Linux Foundation Education has launched a PyTorch Certified Associate (PTCA) Certification Pathway that combines four self-paced learning modules with the PTCA exam. The pathway includes 15–17 hours of self-paced learning and hands-on labs covering tensors, data handling, model development, and performance optimization. The source recommends additional hands-on practice before taking the exam.

  17. GitHub Copilot ChangelogOfficialAI score34

    Copilot code review gains API access and Balanced default effort level

    AIGitHub Copilot code review can now be requested through the REST and GraphQL APIs, with an optional review effort level set per request. Balanced became the default review effort level for new and existing repositories and organizations as of September 28, 2026, while users who explicitly selected Lite keep that setting. The changes are generally available to Copilot Pro, Pro+, Max, Business, and Enterprise plans.

  18. Sara HookerXAI score26

    Adaption Labs makes its Invent dataset tool available via API

    AIAdaption Labs has made Invent, its tool for generating AI training datasets from a plain-language description, available through an API. Developers can reportedly produce AI-ready training datasets in minutes with a few lines of code, according to the quoted post. Documentation is available at docs.adaptionlabs.ai.

    Image from @sarahookr's post