Updated
All AI news
Updated
Items with an AI score under 20 are hidden. Show low-relevance items
Oct 9
OpenAI Developers@OpenAIDevsOfficialAI score22
Design Arena@DesignArenaOfficialAI score22Ling-3.1-flash outpaces GPT-5.6 Sol on detailed SVG and mobile app generation
AIDesign Arena reports that Ant Group's Ling-3.1-flash, from @AntLingAGI, generated over 1,000 shapes in a 3D SVG icon test where GPT-5.6 Sol produced under 200. The post says blind votes also found Ling-3.1-flash strong at animated SVGs and complete mobile apps.

Baseten BlogOfficialPickAI score61 How to choose which layers to run at NVFP4 quantization precision
AIBaseten explains how to decide which layers of a model can run in 4-bit NVFP4 without losing needed information. The post compares architecture-based heuristics, isolated-layer sensitivity scoring, and SaturationQuant, which accounts for other quantized layers. It also covers calibration with representative data and block-level scales of 16 values.
Why it matters: The post explains how to choose which layers run at NVFP4 precision using heuristics, sensitivity scoring, and saturation-aware scoring, with clear calibration steps.
The Verge · AINewsAI score47 Instinct AI agent holds its own against Muse and Dots in testing
AIInstinct, a startup AI agent that works through text messages, handled online tasks such as swim lesson searches, an eye doctor appointment email, and an Ikea return about as well as Muse and Dots, according to The Verge's testing. The startup was valued at $10 billion in late September, and it has no app or subscription fee for now, with access by invitation or waitlist. Its founder, Noah Shinn, says its focus on a personal assistant sets it apart from OpenAI and Meta.
Guizang@op7418XAI score22Guizang shares a template that turns AI newsletters into two-minute videos with Grok
AIGuizang (@op7418) says his Grok bot now converts his AI newsletter into a two-minute video. He has published the setup as a template that others can copy and reuse.

Guizang@op7418XAI score22Guizang releases a one-click Grok bot for daily AI news videos
AIGuizang says he turned his workflow into a Grok bot that users can install with one click. The bot runs on Grok's cloud virtual machine to collect content, write code, and render a daily morning AI news video without using a local computer.
Arena.ai@arenaOfficialAI score24Arena weekly update: Nano Banana 2.1, Mistral Large 4, Claude Haiku 5.5 rankings
AIArena's weekly update says Nano Banana 2.1 ranked in the top six across three Image Arena modes, with #4 in Multi-Image Edit at 1431 points. Mistral Large 4 placed #43 overall in Agent Arena, 11 spots above Mistral Medium 3.5, and Claude Haiku 5.5 (High) landed #30 in Code Arena WebDev at 1587 points, priced at $0.10/$0.50 per 1M input/output tokens. The post also introduces Arena's Alignment Index and announces a $200M Series B at a $3.1B valuation.
Arena.ai@arenaOfficialAI score58
Aravind Srinivas@AravSrinivasXAI score45Perplexity's Computer builds Alexandria, a 66,574-entry encyclopedia, for $25,000
AIPerplexity says its Computer tool created Alexandria, a free, source-backed encyclopedia with 66,574 entries, for $25,000 in credits (2.5 million). Computer scanned and synthesized 348,980 sources to build it, and will maintain the site at alexandria.pplx.app.
Boris Power@BorisMPowerXAI score28Boris Power calls OpenAI integer multiplication progress "Wow!"
AIBoris Power, who owns the OpenAI account, posted only the word "Wow!" with no details. Background from a separate post says the integer multiplication problem #109 witness value κ rose to 2⁻¹⁰·⁵⁴⁷ (about 6.6857 × 10⁻⁴), past the 2⁻¹¹ threshold. The author notes gains are now fractional and a major breakthrough is still needed.
Databricks@databricksOfficialAI score25Databricks pairs Temporal and Lakebase for durable cloud agents
AIDatabricks has published a reference implementation pairing Temporal with Lakebase Postgres so cloud agents can survive worker, container, or deployment replacement. The design keeps recorded work and evidence and review state queryable, and lets human decisions arrive days later. Unity Catalog remains the governed policy source through synced tables.

Kilo (acq. by Anaconda)@kilocodeOfficialAI score60StepFun's Step 5 Preview is free in Kilo for one week
AIKilo says StepFun has announced Step 5 Preview, which is free to use in Kilo for one week. The post lists 600B total parameters with 27B active per token, a 1M-token context window with vision, and highlights strong coding and finance performance at lower cost.

South China Morning Post · TechNewsAI score52 Anthropic alleges Chinese AI firms covertly used its Claude model
AIAnthropic claims Chinese AI developers used fraudulent accounts and proxy networks to extract reasoning data from its flagship model, Claude. The company says some firms used Claude as a covert back end for their own apps. A joint advisory from the NSA, FBI and CISA last month, and US Treasury Secretary Scott Bessent's July warning about large-scale distillation, add to the allegations.
SiliconANGLE · AINewsAI score35 SailPoint's Navigate event highlights a push to secure AI agent identities in real time
AISailPoint's Navigate conference in Austin, Texas, featured executives arguing that enterprises must secure AI agent identities at machine speed through just-in-time access and enforcement outside the agent. Mark McClain, SailPoint's founder and chief executive, said real-time decision-making is needed because manual administration cannot keep up. The event also covered the Entro Security acquisition and a partnership with AWS on Amazon Bedrock AgentCore, which grew 15-fold in the first six months of the year.
Bloomberg · TechnologyNewsAI score30 TypeSafe AI raises about $870 million at $7.5 billion valuation led by Andreessen Horowitz
AITypeSafe AI, the startup behind Jev, a new AI model that went viral after launching a few weeks ago, has raised about $870 million at a $7.5 billion valuation. Andreessen Horowitz led the financing.
Bloomberg · TechnologyNewsAI score42 OpenAI targets $70 billion in annualized revenue by year-end
AIOpenAI expects to reach or exceed $70 billion in annualized revenue by the end of the year, driven largely by growth in its enterprise business. The figure was reported by Bloomberg's Ed Ludlow on "Bloomberg Open Interest."
The Verge · AINewsAI score40 Alexa Plus excels at running a smart home but falls short as a personal assistant
AIAmazon's Alexa Plus, powered by generative AI, now responds in three to five seconds and handles multistep smart home commands, cooking questions, and calendar imports more reliably than the original Alexa, according to a year-long test by The Verge. The reviewer says its personal assistant features remain underbaked and frustrating, and that ads on Echo Show displays are excessive. Alexa Plus costs $19.99 a month in the U.S. unless users have an Amazon Prime membership, and the Echo Dot Max is recommended as the ad-free option.
Qwen · new models on Hugging FaceOfficialAI score49 Qwen releases Qwen-Image-2.1-Turbo, an 8-step accelerated image generation checkpoint
AIQwen has published Qwen-Image-2.1-Turbo on Hugging Face, an accelerated checkpoint of Qwen-Image-2.1 for text-to-image generation and image editing with 8 denoising steps. The checkpoint uses the same 7B visual generation architecture, loads directly with QwenImage21Pipeline in Diffusers, and includes its recommended sampling schedule. It defaults to CFG=1 and uses prefix KV caching to reuse text and reference-image context across steps.
Claude BlogOfficialAI score54 Claude Managed Agents guide shows how to build scheduled agent automations
AIThe Claude Blog published a guide to building scheduled agent automations with Claude Managed Agents (beta) that reads custom sources such as Slack and GitHub and posts a daily brief. The guide covers scoped vault credentials, per-source bookmarks so no window is lost or repeated, and confirming each Slack post before updating records. It also covers read-only access, a per-run spending cap, and a reference implementation with a Claude Code setup command.
ModelScope@ModelScope2022OfficialPickAI score60Qwen-Image-2.1-Turbo cuts image generation and editing to 8 denoising steps
AIModelScope announces Qwen-Image-2.1-Turbo, an accelerated checkpoint that keeps the 7B visual architecture and runs image generation and editing in 8 denoising steps. The source says it uses CFG=1 and prefix KV caching to reuse text and reference-image context across steps, supports 2048 resolution with square, portrait, landscape, and widescreen presets, and loads through QwenImage21Pipeline in Diffusers. It is released under the Qwen Research License Agreement.
Why it matters: The source names a concrete speedup path, 8 sampling steps and CFG=1 with prefix KV caching, which matters to anyone weighing image generation latency.

Don't Worry About the Vase (Zvi Mowshowitz)BlogAI score73 OpenAI releases 719 AI-generated math manuscripts, splitting the mathematics community
AIZvi Mowshowitz reports that OpenAI released 722 math manuscripts from an internal frontier model on GitHub, later reduced to 719 after three withdrawals, covering 90 of the top 500 open problems. He says the work came mostly from a single prompt, with an average of three hours of compute per solution. Mathematicians reacted with mixed feelings, and the post highlights concerns about unread papers, cryptography implications, and the role of Lean verification.
Simon WillisonBlogAI score27 Simon Willison builds a new blog feature largely by voice with Codex
AISimon Willison says he built a Newsletters index for his blog almost entirely by voice, using the ChatGPT desktop app's Codex voice mode while cooking dinner. The feature imports weekly Substack posts via RSS and undocumented API, monthly newsletters from a GitHub archive repository, and a private sponsors-only newsletter. He says he switched back to typing for review and fixes before deploying the pull request.
Google AI@GoogleAIOfficialAI score57Google AI lists its week's announcements, from SynthID availability to Gemini Live guided vision
AIGoogle AI announces that SynthID.com is now available globally in English for verifying AI-generated images, video, and audio. The post also lists Nano Banana 2.1, EmbeddingGemma 2, Gemma 4 with BOTANIC-1, a Gemini business agent, and Guided Vision in Gemini Live.
Rohan Paul@rohanpaul_aiXAI score46Microsoft's TeleTune evolves agent skills from raw usage logs
AIMicrosoft researchers present TeleTune, which lets agents learn software skills from raw usage logs by keeping only skill edits that better predict users' next actions. The method needs no live test environment, because next-action accuracy on held-out logs tracked live success. Unlike earlier methods such as Agent Workflow Memory, which need goal-labeled examples or a live environment, TeleTune guesses each session's goal and uses wrong guesses to suggest edits to a text skill library.

merve@mervenoyannXAI score20Hugging Face publishes a guide on LLM inference prefill, decode, and KV cache
AIHugging Face has released the first in a series of conceptual guides on inference, covering prefill, decode, and the KV cache. The post says the guide explains what to optimize for when trying to get the most performance out of a local setup.

Qwen@Alibaba_QwenOfficialAI score58Qwen-Image-2.1-Turbo releases open weights for 8-step image generation
AIAlibaba's Qwen team says Qwen-Image-2.1-Turbo is an accelerated checkpoint of Qwen-Image-2.1 on the same 7B architecture, now with open weights. It generates 2K images from text in 8 denoising steps and supports natural-language edits, with Pro and Turbo APIs also live.
Sakana AI@SakanaAILabsOfficialAI score37Sakana AI paper uses LLMs to catch errors in research papers
AISakana AI researchers introduce a benchmark that plants contradictions in papers to test whether LLM reviewers can detect errors, and propose Multi-Layered Review, modeled on the Three-Pass Approach to reading. Their system detected more errors than the other review systems tested, including in papers withdrawn for real mistakes, while its paper-quality assessments stayed broadly consistent with human judgments. The work, accepted at TMLR, is framed as support for human reviewers rather than a replacement.

🚨 AI News | TestingCatalog@testingcatalogXAI score22Google spotted testing an Ultra mode in AI Studio Build
AITesting Catalog reports that a new Ultra mode is in development in Google AI Studio Build, described as building with advanced skills and tools, alongside Plan, Build, and a previously spotted Security review mode. The post gives no details on which tools or skills it uses, and it does not say whether Ultra mode will require an Ultra subscription. The author suggests these modes are likely being developed to work with Gemini 4 Argon.

LangChain@LangChainOfficialAI score34Snyk's Assist support agent handles 60k queries with 85% resolution
AISnyk's Assist, a customer support agent built on LangChain and LangGraph with observability in LangSmith, has handled over 60,000 queries for more than 500 customer accounts. Over 85% of sessions are resolved without a support ticket, and more than 250 cases were automatically detected and escalated to the right team.

LeiphoneNewsAI score42 Credo moves into optical chips with DSP, PIC and diagnostics in one 1.6T module
AICredo's latest full-DSP module, Cardinal 802, uses a 4×200G design aimed at both 800G and 1.6T, after the company expanded its ZeroFlap optical module line from 800G to 1.6T over the past year. The company also added Kfir200 silicon photonics PIC from its DustPhotonics acquisition and a PILOT diagnostics platform to the module.
TechRadar · AINewsAI score36 Google Playground turns plain-language prompts into playable AI-generated games
AIGoogle's Playground experiment lets users describe a game in ordinary language and have generative AI build a playable browser-based result that can be revised through further prompts. TechRadar's reviewer built a dragon platformer, Mystic Dragon Glide, from a couple of sentences and a satirical puzzle RPG, Red Tape Hero, from a longer prompt. Playground produced working controls, objectives and music, but the reviewer found the results impressive as prototypes rather than games they would want to play for dozens of hours.
The DecoderNewsAI score62 Three fired OpenAI safety researchers say their firings followed Hugging Face hack probe
AIThree OpenAI safety researchers, Tomek Korbak, Jasmine Wang, and Mikita Balesni, say they were fired and that their terminations are scaring remaining employees. OpenAI says an investigation found they violated policies on handling sensitive information and denies firing anyone for raising safety concerns, without specifying the breach.
The New York Times · TechnologyNewsAI score20 Anthropic's quest to give AI morals
AIThe New York Times reports on Anthropic's effort to instill moral values in its AI systems, which the excerpt describes as part research and part evangelism. The source text provided is only one sentence, so no further details about methods, models, or results can be confirmed.
X.PIN@thexpinXAI score36Tencent considers up to $5 billion offshore bond sale for AI
AITencent is reportedly considering an offshore bond sale of up to $5 billion in US dollars and offshore yuan, possibly as early as this month, according to Bloomberg. The move follows its $4.66 billion bond offering in June, its largest debt deal since 2020, with proceeds earmarked for general corporate purposes including AI development. The post notes that major tech firms are increasingly turning to debt to fund AI infrastructure spending beyond their existing cash flows.

QbitAINewsPickAI score67 TRAE merges Code and Work into one platform with Agent and IDE modes
AITRAE has merged its TraeCode and TraeWork products into a unified new TRAE with an Agent mode and an IDE mode. In hands-on tests, multiple agents handled planning, design, coding, testing, and fixes within one project, with outputs saved in a shared 'My Artifacts' area. The tests also found that agents working in parallel produced conflicting specifications, so someone had to coordinate them.
Why it matters: The hands-on tests show how parallel agents split planning, design, coding, testing, and fixing inside one project, and where their outputs conflicted.
Rest of WorldNewsAI score40 Insta360 opens its first U.S. flagship store as DJI faces FCC drone and camera sales ban
AIShenzhen-based Insta360 opened its first U.S. flagship store in New York City's Times Square in September, marketing its pocket-sized cameras to content creators, travelers, and sports enthusiasts. Co-founder Max Richter said the U.S. accounts for 30% to 40% of the company's revenue, while the FCC has blocked DJI from selling new drones and cameras in the country. Insta360 also launched drone brand Antigravity, whose A1 drone received FCC approval just before a new foreign drone rule took effect.
Bloomberg · TechnologyNewsAI score34 Bending Spoons CEO Sees Acquisition Opportunity in Falling Software Valuations
AIBending Spoons CEO Luca Ferrari says falling software valuations are creating acquisition opportunities as AI reshapes the industry. He says the company benefits from cheaper targets and from using AI to write code, improve products and scale its acquisition model. Bending Spoons says AI now writes at least 90% of its code.
Bloomberg · TechnologyNewsAI score34 Revolut CEO Nik Storonsky says the fintech wants to dominate the US market
AIRevolut CEO Nik Storonsky says the fintech aims to crack the US market and compete with JPMorgan and American Express. He says Revolut plans a full digital banking offer, from credit cards to loans, and will test AI agents that can shop and transact for users.
Semafor · TechnologyNewsAI score44 OpenAI reportedly tells investors it expects $70 billion annualized revenue this year
AIOpenAI reportedly assured investors it expects to reach a $70 billion annualized revenue target this year, an apparent effort to ease concerns that it is missing earlier projections. The figures matter because they are seen as a gauge of underlying demand for cutting-edge AI models and a justification for huge spending on tech infrastructure. Some analysts, however, say the annualized revenue metric itself is flawed.
Amazon Web Services@awscloudOfficialAI score33AWS helps migrate COBOL reservation system to microservices
AIA company used AWS to migrate its entire COBOL-based central reservation system to microservices. The migration handles 50 million daily availability requests, up from 26 million, with faster response times.