Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Apr 7

Apr 7Tue
  1. Dario AmodeiXAI score72

    Dario Amodei backs Project Glasswing to counter AI-driven cyber threats

    AIDario Amodei said many of the world's leading companies have joined Project Glasswing, an effort to address cyber threats posed by increasingly capable AI systems. The initiative was introduced by Anthropic and is powered by its newest frontier model, Claude Mythos Preview, which the quoted post says can find software vulnerabilities better than all but the most skilled humans.

    Why it matters: The post gives a concrete example of how a frontier AI lab is organizing industry partners around AI-driven software vulnerability discovery.

  2. Andy JassyXAI score36

    Uber uses AWS Graviton4 and Trainium3 chips for rides and AI

    AIUber is running its ride and delivery matching on AWS Graviton4 chips and training its AI models on Trainium3, according to Amazon CEO Andy Jassy. Jassy says the Graviton4 setup matches riders with drivers in fractions of a second at lower cost, while Trainium3 helps make rides smarter over time.

    Image from @ajassy's post

Apr 6

Apr 6Mon
  1. OpenAI Alignment Research BlogOfficialAI score31

    OpenAI opens applications for Safety Fellowship on AI safety and alignment research

    AIOpenAI announced applications for its Safety Fellowship, a pilot program supporting external researchers, engineers, and practitioners in safety and alignment research on advanced AI systems. The program runs from September 14, 2026 through February 5, 2027, with a monthly stipend, compute support, API credits, and mentorship, and fellows are expected to produce a substantial output such as a paper, benchmark, or dataset. Applications close May 3, and successful applicants will be notified by July 25.

Mar 30

Mar 30Mon

Mar 27

Mar 27Fri

Mar 20

Mar 20Fri
  1. Aman SangerXAI score55

    Cursor's Composer 2 is built on Kimi k2.5 base model with added training

    AIAman Sanger says Cursor's team evaluated many base models on perplexity-based evals and found Kimi k2.5 the strongest. Composer 2 was then built with continued pretraining and a 4x scale-up of high-compute RL, with Fireworks providing inference and RL samplers. The author admits Cursor should have named the Kimi base in its launch blog and says it will do so for the next model.

  2. Aman SangerXAI score22

    Cursor's Composer 2 model praised, built on an open-source base

    AIAman Sanger of Cursor says Composer 2 is a really good model and he is excited for more people to try it. The quoted reply from Lee Robinson says Composer 2 started from an open-source base, with only about one-quarter of the final model's compute coming from that base. Cursor plans full pretraining in the future and says it is following the license through its inference partner terms.

Mar 16

Mar 16Mon

Mar 10

Mar 10Tue

Mar 2

Mar 2Mon

Mar 1

Mar 1Sun
  1. Chris OlahXAI score25

    Law professor predicts APA challenge over government treatment of Anthropic

    AIGeorgetown law professor Mark Zhia predicts a court would likely find the government's treatment of Anthropic arbitrary and capricious under the APA. He argues OpenAI and Anthropic pose similar security risks, yet the government immediately contracted with OpenAI while treating Anthropic worse than Chinese AI firms.

Feb 27

Feb 27Fri
  1. Mckay WrigleyXAI score80

    Pentagon Secretary moves to label Anthropic a supply-chain risk

    AIMckay Wrigley reposted a statement from @SecWar accusing Anthropic of refusing the Department of War unrestricted access to its models for lawful purposes. The quoted statement directs the Department of War to designate Anthropic a Supply-Chain Risk to National Security, bars contractors from commercial activity with Anthropic, and allows Anthropic services for no more than six months. The author's own added text says only that he finds the situation horrifying and supports Anthropic.

    Why it matters: The quoted statement is a direct government action against a named AI lab, giving readers a primary-source view of a dispute over military access to AI models.

  2. Nick TurleyXAI score36

    ChatGPT surpasses 900 million weekly users and 50 million paying subscribers

    AIChatGPT has crossed 900 million weekly users and 50 million paying subscribers, according to OpenAI's Nick Turley. He says people use it for writing, building, research, trip planning, shopping, and everyday tasks, and that growing usage is helping improve response speed, reliability, and naturalness.

Feb 25

Feb 25Wed
  1. Yi TayXAI score62

    Aletheia math research agent solves 6 of 10 FirstProof problems

    AIAletheia, a math research agent, autonomously solved 6 of 10 FirstProof problems without modification, the best result in the inaugural challenge. The author says this is bigger than the IMO-gold achievement from last year, and the results were evaluated by experts with best-of-2 scoring.

  2. Quoc LeXAI score53

    Google's Aletheia math agent solves 6 of 10 FirstProof problems

    AIQuoc Le announced that Aletheia, a math research agent, autonomously solved 6 of 10 FirstProof problems, the best result in the inaugural challenge. The post says this exceeds last year's IMO-gold achievement and points to a paper and thread for full details. The accompanying figure shows 10 unmodified problems, 6 candidate solutions per agent, and expert evaluation yielding 6 solved problems on a best-of-2 basis.

Feb 24

Feb 24Tue

Feb 14

Feb 14Sat
  1. Jakub PachockiXAI score22

    OpenAI now believes its #1stProof problem 2 solution is likely incorrect

    AIJakub Pachocki, OpenAI's account holder, says that after official #1stProof commentary, community analysis, and external expert review, the team now believes its solution to problem 2 is likely incorrect. He thanked reviewers for their engagement and said the team looks forward to continued review.

Feb 13

Feb 13Fri
  1. Jakub PachockiXAI score62

    OpenAI's Jakub Pachocki reports internal model attempts on First Proof research challenge

    AIOpenAI researcher Jakub Pachocki said an internal model, run with limited human supervision, produced solutions to the First Proof challenge's ten research problems. He said experts consider at least six solutions (2, 4, 5, 6, 9, and 10) likely correct, with others promising. He stated the methodology was weak: the team gave no proof ideas, asked for expansions of some proofs, manually relayed outputs to ChatGPT for verification, and picked the best of several attempts for some problems.

Feb 12

Feb 12Thu

Feb 7

Feb 7Sat

Jan 27

Jan 27Tue
  1. Cognition Blog (Devin, Windsurf)OfficialAI score32

    Cognition opens London office to expand Devin autonomous coding for European businesses

    AICognition is opening a London office to expand rollout of Devin, its autonomous software engineering agent, to leading European businesses. The company says finance has emerged as a clear use case, with Goldman Sachs, Santander, Citi, and BNY among partners using Devin for modernization, migration, security remediation, and codebase documentation.

  2. Cognition Blog (Devin, Windsurf)OfficialAI score38

    Cognizant Partners with Cognition to Scale Devin and Windsurf Across Its Engineering Teams

    AICognizant is deploying Cognition's Devin autonomous software engineer and Windsurf agentic IDE across its engineering organization and global client base. Engineers already use Windsurf for agent-assisted coding and are exploring Devin for end-to-end tasks such as code migration, refactoring, testing, and maintenance. Cognition will embed forward-deployed AI engineers to support project selection, engineer enablement, and ROI measurement.

Jan 14

Jan 14Wed
  1. Ahmad Al-DahleXAI score38

    Ahmad Al-Dahle joins Airbnb as CTO after Llama open-source work

    AIFormer Meta AI leader Ahmad Al-Dahle announced he is joining Airbnb as CTO, citing Meta's open-sourcing of Llama, which has reached over 1.2 billion downloads and 60,000+ derivatives. He said the next challenge is applying advancing model capabilities to products that connect people with real places, working with Airbnb CEO Brian Chesky.

Jan 6

Jan 6Tue
  1. Cognition Blog (Devin, Windsurf)OfficialAI score42

    Infosys partners with Cognition to deploy Devin AI software engineer across its enterprise

    AIInfosys will deploy Cognition's Devin, an autonomous AI software engineer, across its own teams and global client base to expand delivery capacity. The rollout begins in its Financial Services practice, covering banking, payments, capital markets, insurance, and wealth management, and is planned to extend to retail, energy, and healthcare. Over the past six months, Infosys reports material productivity gains, including COBOL and JCP servlet migrations completed in record time.

Dec 19, 2025

Dec 19, 2025Fri

Dec 11, 2025

Dec 11, 2025Thu

Dec 4, 2025

Dec 4, 2025Thu
  1. ARC PrizeOfficialAI score62

    ARC Prize 2025 results point to refinement loops as the central AI reasoning trend

    AIARC Prize reports that the top Kaggle entry reached 24% on the ARC-AGI-2 private dataset at $0.20 per task, and that all winning solutions and papers are open source. The top verified commercial model, Opus 4.5 (Thinking, 64k), scored 37.6% at $2.20 per task, while a Poetiq refinement on Gemini 3 Pro reached 54% at $30 per task. The author argues that refinement loops are the main driver of 2025 progress, and says ARC-AGI-3 is planned for early 2026.

    Why it matters: The post links 2025 competition results to a broader argument about refinement loops, showing how benchmark outcomes are being read as evidence of AI reasoning progress.

  2. Quoc LeXAI score38

    Google DeepMind launches new Gemini reasoning research team in Singapore

    AIQuoc Le announced that Google's Gemini team is recruiting for a new Singapore team focused on advanced reasoning and LLM reinforcement learning. The post invites exceptional engineers and researchers to join the team, which is led by Yi Tay and reports into Quoc Le's Mountain View group.

  3. Yi TayXAI score38

    Google DeepMind's Gemini team launches new reasoning research group in Singapore

    AIYi Tay announced that Google DeepMind's Gemini team is starting a new research team in Singapore focused on advanced reasoning, LLM/RL, and improving frontier models such as Gemini and Gemini Deep Think. The team is led by Tay and reports to Quoc Le's broader team in Mountain View, which recently contributed to IMO and ICPC gold medal results with Gemini Deep Think. The team is starting small and is recruiting exceptionally capable engineers and researchers from the region and beyond.

    Image from @YiTayML's post

Nov 19, 2025

Nov 19, 2025Wed
  1. Stability AIOfficialAI score36

    Warner Music Group and Stability AI Partner on Responsible AI Music Creation Tools

    AIWarner Music Group and Stability AI announced a collaboration to build professional-grade, ethically trained AI tools for artists, songwriters, and producers. The companies will work directly with artists to shape tools that enhance the creative process while protecting creators' rights and revenue opportunities. Stability AI's Stable Audio models, trained exclusively on licensed data, are cited as the basis for its commercially safe generative audio offerings.

Nov 13, 2025

Nov 13, 2025Thu

Nov 3, 2025

Nov 3, 2025Mon
  1. ARC PrizeOfficialAI score38

    ARC Prize Launches Verified Program to Certify ARC-AGI Benchmark Scores

    AIARC Prize Foundation announced ARC Prize Verified, a program that certifies frontier model scores on the ARC-AGI benchmark using hidden test sets and adds a third-party academic panel to audit and open-source its testing process. Five AI labs, including Google and xAI, are sponsoring ARC-AGI-3 development, and the foundation says donations do not influence verification scoring. Models that pass verification will appear on the official leaderboard with a verification badge.

Oct 30, 2025

Oct 30, 2025Thu
  1. Stability AIOfficialAI score38

    Universal Music Group and Stability AI Form Alliance to Co-Develop Professional AI Music Tools

    AIUniversal Music Group and Stability AI announced a strategic alliance to co-develop professional music creation tools built on responsibly trained generative AI. The collaboration will work with UMG artists to research their needs and prioritize their feedback in building fully licensed, commercially safe AI music tools. Stability AI's Stable Audio models, which the company says were trained exclusively on licensed data, are part of its audio offering.

Oct 26, 2025

Oct 26, 2025Sun
  1. Factory NewsOfficialAI score36

    AWS and Factory Announce Partnership, Factory Available on AWS Marketplace

    AIFactory has announced a partnership with Amazon Web Services and made its Droids agent platform available on the AWS Marketplace. Enterprise teams can use existing AWS Enterprise Discount Program commitments to buy Factory, with Droids accessible from CLI, Terminal UI, Web, Slack, Linear, and an IDE overlay. The source cites 31× faster feature development, 96.1%+ reduction in migration times, and 95.8% reduction in incident resolution times.

Oct 23, 2025

Oct 23, 2025Thu
  1. Stability AIOfficialAI score38

    Stability AI and EA Partner to Build Generative AI Tools for Game Development

    AIStability AI and Electronic Arts have formed a strategic partnership to co-develop generative AI models, tools, and workflows for EA's artists, designers, and developers. Among the first initiatives is accelerating the creation of Physically Based Rendering (PBR) materials, such as 2D textures that preserve color and light accuracy across environments. The partners also plan to build AI systems that pre-visualize entire 3D environments from a series of prompts.

Sep 7, 2025

Sep 7, 2025Sun
  1. Cognition Blog (Devin, Windsurf)OfficialAI score53

    Cognition raises over $400M at $10.2B valuation after Windsurf acquisition

    AICognition, maker of the AI software engineer Devin, raised over $400M at a $10.2B post-money valuation led by Founders Fund. The company says its acquisition of Windsurf more than doubled its ARR, with combined enterprise ARR up over 30% in the seven weeks after the deal. It also reports Devin ARR grew from $1M in September 2024 to $73M in June 2025, with total net burn under $20M.