Skip to content

#Other

May 26

May 26Tue
  1. Max WoolfAI score23

    New (short!) blog post up: on @OpenRouter 's AI Model Rankings, I noticed a peculiar new LLM topping the rankings by a large margin: Hy3. I looked into the data and only became more confused. https://minimaxir.com/2026/05/openrouter-hy3/

    New (short!) blog post up: on @OpenRouter 's AI Model Rankings, I noticed a peculiar new LLM topping the rankings by a large margin: Hy3. I looked into the data and only became more confused. https://minimaxir.com/2026/05/openrouter-hy3/

May 24

May 24Sun

May 22

May 22Fri
  1. Google LabsAI score12

    Take a quick break from scrolling and check out http://labs.google. 👀 We had a bit of a refresh! Our goal was simply to make sure you can easily find the latest and greatest innovations from the Lab, including those we just announced at I/O. Explore our portfolio, test a few experiments out, learn more about Google Labs, and help us shape what’s next. 🚀

    Take a quick break from scrolling and check out http://labs.google. 👀 We had a bit of a refresh! Our goal was simply to make sure you can easily find the latest and greatest innovations from the Lab, including those we just announced at I/O. Explore our portfolio, test a few experiments out, learn more about Google Labs, and help us shape what’s next. 🚀

May 21

May 21Thu

May 17

May 17Sun
  1. Cognition Blog (Devin, Windsurf)AI score60

    Cognition launches Auto-Triage, letting Devin investigate alerts and open fixes

    Cognition has released Auto-Triage in Devin Automations, which lets Devin respond to Slack messages, Linear events, GitHub activity, schedules, and webhooks. Devin can investigate with connected observability tools and the codebase, then post a summary, tag an owner, or open a PR. Devin runs in network-sandboxed environments with added protections against prompt injection and data exfiltration, and a limited-time offer gives $200 in credits for a first automation.

    AIWhy it matters: The post shows how an agent handles alerts and bug reports from existing team channels, a practical pattern for teams weighing automated incident response.

May 11

May 11Mon
  1. Mira MuratiAI score16

    The current “AI experience” often feels like a conversation that only begins after we stop talking. We have to batch our thoughts. We can’t point at things. We phrase questions like emails. The interface doesn't leave room for us so we adapt to the models.

    The current “AI experience” often feels like a conversation that only begins after we stop talking. We have to batch our thoughts. We can’t point at things. We phrase questions like emails. The interface doesn't leave room for us so we adapt to the models.

  2. Fidji SimoAI score13

    @CellularIntelHQ is such a great example of the hope we can have for AI + bio. I still remember when that company was nothing more than an idea. Proud of @MichaBreakstone and proud to be an advisor to this mission-driven team.

    @CellularIntelHQ is such a great example of the hope we can have for AI + bio. I still remember when that company was nothing more than an idea. Proud of @MichaBreakstone and proud to be an advisor to this mission-driven team.

Apr 23

Apr 23Thu
  1. Chip HuyenAI score22

    Congrats to @AlecRad @Luke_Metz and @soumithchintala! It's really cool to see work produced by a group of 20-somethings without a single PhD between them winning this award. I hope to see these three collaborate again one day :)

    Congrats to @AlecRad @Luke_Metz and @soumithchintala! It's really cool to see work produced by a group of 20-somethings without a single PhD between them winning this award. I hope to see these three collaborate again one day :)

Apr 17

Apr 17Fri

Apr 7

Apr 7Tue

Apr 4

Apr 4Sat
  1. Andrej KarpathyAI score62

    Andrej Karpathy outlines an LLM-maintained markdown wiki workflow for personal research

    Karpathy describes using LLMs to compile raw source documents into a markdown wiki that he views in Obsidian, with the LLM writing and maintaining most of the wiki. He reports that at about 100 articles and 400K words, the LLM agent can answer complex questions directly from the wiki, and he also runs LLM health checks to find inconsistencies and gaps. He shares the underlying idea as an "idea file" that users can give to their own agents to build a customized version.

Apr 2

Apr 2Thu

Mar 25

Mar 25Wed
  1. Andrej KarpathyAI score18

    One common issue with personalization in all LLMs is how distracting memory seems to be for the models. A single question from 2 months ago about some topic can keep coming up as some kind of a deep interest of mine with undue mentions in perpetuity. Some kind of trying too hard.

    One common issue with personalization in all LLMs is how distracting memory seems to be for the models. A single question from 2 months ago about some topic can keep coming up as some kind of a deep interest of mine with undue mentions in perpetuity. Some kind of trying too hard.

Mar 17

Mar 17Tue
  1. Tri DaoAI score49

    The frontier has increasingly shifted to hybrid models - from Qwen to Kimi-Linear and now with NVIDIA's Nemotron-3 Super - that rely on a strong linear sequence model. Today we release Mamba-3, the most powerful linear model to date. https://x.com/_albertgu/status/2033948415139451045

    The frontier has increasingly shifted to hybrid models - from Qwen to Kimi-Linear and now with NVIDIA's Nemotron-3 Super - that rely on a strong linear sequence model. Today we release Mamba-3, the most powerful linear model to date. https://x.com/_albertgu/status/2033948415139451045

Feb 23

Feb 23Mon

Feb 14

Feb 14Sat
  1. Oriol VinyalsAI score13

    Personal update: After an amazing 10 years in London, it's time for a major change. One-way ticket back to California 🌞! I'm incredibly excited to return to the Bay Area to continue building Gemini and pushing us toward the age of AGI 🚀

    Personal update: After an amazing 10 years in London, it's time for a major change. One-way ticket back to California 🌞! I'm incredibly excited to return to the Bay Area to continue building Gemini and pushing us toward the age of AGI 🚀

Feb 4

Feb 4Wed
  1. Guillaume LampleAI score62

    Mistral releases Voxtral 2 transcription models with real-time option

    Mistral announces Voxtral 2 with two transcription models: Voxtral Realtime, released under an Apache 2 license with latency configurable to sub-200 ms, and Voxtral Mini Transcribe 2, which adds speaker diarization, word-level timestamps, and context biasing. The models support 13 languages and are available through the Mistral API, which the post describes as one of the most cost-effective transcription APIs on the market. The attached chart shows word error rates on FLEURS across Italian, Spanish, English, German, Portuguese, French, Russian, Dutch, and Chinese at several latency settings.

Feb 3

Feb 3Tue

Jan 14

Jan 14Wed
  1. Lilian WengAI score3

    I’ve been telling people this a lot today: I enjoy so much working with people who care about what they are building and craftsmanship. It is a privilege to have a chance to work on something I’m passionate about, beyond making a living. I cherish it and don’t take it for granted.

    I’ve been telling people this a lot today: I enjoy so much working with people who care about what they are building and craftsmanship. It is a privilege to have a chance to work on something I’m passionate about, beyond making a living. I cherish it and don’t take it for granted.

Dec 18, 2025

Dec 18, 2025Thu

Nov 29, 2025

Nov 29, 2025Sat
  1. Andrej KarpathyAI score62

    Karpathy argues LLMs are a new kind of intelligence shaped by commercial, not evolutionary, pressure

    Karpathy argues animal intelligence is only one point in a large space of possible minds, and LLMs arise from a fundamentally different optimization process. He contrasts survival-driven animal drives with LLM training shaped by imitation of human text, RL on task distributions, and user engagement metrics, which he says leaves LLMs jagged and prone to sycophancy. He calls LLMs humanity's first contact with non-animal intelligence and says people who build accurate internal models of them will reason about them better.

Nov 26, 2025

Nov 26, 2025Wed

Nov 25, 2025

Nov 25, 2025Tue
  1. Oriol VinyalsAI score14

    On my way to Barcelona to receive a Doctor Honoris Causa from my alma mater, @la_UPC. Truly honored! 🎓 Join Thursday for my Master Class, "From AI to AGI: The Quest for True Intelligence." Hope to see you there! https://telecos.upc.edu/ca/esdeveniments/master-class-del-dr-oriol-vinyals-from-ai-to-agi-the-quest-for-true-intelligence "Create an image at 41.4036° N, 2.1744° E, January 1st, 1983, 15:00 hours."

    On my way to Barcelona to receive a Doctor Honoris Causa from my alma mater, @la_UPC. Truly honored! 🎓 Join Thursday for my Master Class, "From AI to AGI: The Quest for True Intelligence." Hope to see you there! https://telecos.upc.edu/ca/esdeveniments/master-class-del-dr-oriol-vinyals-from-ai-to-agi-the-quest-for-true-intelligence "Create an image at 41.4036° N, 2.1744° E, January 1st, 1983, 15:00 hours."

Oct 14, 2025

Oct 14, 2025Tue