Really proud of how many family-run businesses are powered by Replit.
Really proud of how many family-run businesses are powered by Replit.
Really proud of how many family-run businesses are powered by Replit.
Building Hogwarts live in @spawn Come help me build it: https://www.spawn.co/@mattshumer/hogwarts
I hope we keep Opus 4.6 around for a long time
The author says Beam, a 501B-parameter open model from Reflection AI, comes close to GLM 5.2 in capability but trails GLM 5.3, Kimi K3, and DeepSeek V4.1 Flash in several areas. The author attributes Beam's competitiveness mainly to inference efficiency, with inference compute at roughly one-third to one-quarter of GLM 5.2's.
Google Research released a workshop report, "Open and Emergent Problems in Agentic Privacy and Security: A Contextual Angle," compiled by more than 50 academic and industry leaders from the Google Contextual Agent Privacy and Security (CAPS) Workshop held in late 2025 in New York City.
Reflection joins the list of Nvidia & Thinking Machines who have released their strongest models and come up behind Chinese counterparts. There's a lot of ways people will overthink this, but the clearest takeaway should be that the Chinese are very very good at building LLMs.
Sophia Yang congratulated Reflection AI on Beam, a 501B-parameter open model with 23B active per token. She attributes its efficiency to an RL length penalty that discourages unnecessary tokens and a sparse MoE architecture. Reflection says full weights will be released this month, and the quoted post reports training over 100 million rollouts on 10.5K NVIDIA GB300 GPUs over four weeks.
If you're using Cowork on your local machine, just a heads up that from tomorrow your tasks run in the cloud instead. They can still use the files and tools on your computer though! Felix explains the reasoning here 👇
Press Play ▶️ This week, @amasad talks vibe coding, the lost joy of building with computers, and why AI doomerism misses the accountability question. If AI agents can act on our behalf, who’s responsible when they go too far? Listen to the new Times Tech episode: https://open.spotify.com/episode/6l1hmeZ1FGzbC6fYWSHJSi?si=QtW1jsqQQTGK_KJsbPt8yA&nd=1&dlsi=7225ec4f531942e8
Dex Horthy advises writing all decisions and context into documents in the artifacts, such as design or research files, so sessions can resume after compaction or be handed to another person. He suggests loading them in a new session with a skill like `/rpi:iterate-design-discussion`, or simply @-mentioning the relevant artifacts. His core principle is that nothing important should live only in the context window.
PyTorch has consolidated all media decoding and encoding for images, video, and audio into TorchCodec, which now runs on CPU and CUDA. TorchVision and TorchAudio are narrowed to focus on their transforms, with models, datasets, and pipelines no longer under active development. All three libraries are now ABI stable and no longer need rebuilding for each PyTorch release.
Nous Research says users should control their agent's models, data, memory, compute location, prompts, tools, and code. The post lists choices such as switching models mid-conversation, running fully offline, and exporting the agent. It frames these freedoms as the standard an agent should meet, calling it "yours."
Microsoft & Meta, two of Anthropic's biggest customers, make progress in cutting their Claude bills...
Films & music 🎵 DJs & creators 🧑🎨 Sunset & pizza 🌊 And cool MiniMax caps 🧢 Lisbon, it was so good to meet you. 💜 🇵🇹
cool to see cognition doing "dreaming" - memory needs an offline cleanup loop, not just better retrieval, otherwise stale records pile up open q is how inferred memories get validated before the agent uses them glad it's an open standard too - memory should be open https://x.com/cognition/status/2107165034463867001
Anthropic Subscriptions Offer 5x+ More Value Than OpenAI Limit testing every AI subscription plan from Anthropic, OpenAI, Meta, SpaceXAI, MiniMax, Moonshot, Zdotai, Cursor, and Cognition https://newsletter.semianalysis.com/p/anthropic-subscriptions-offer-5x
The model behind this result is now on @huggingface 🤗 Nemotron-Labs-3-Competitive-Coding is a competitive-programming specialist model based on Nemotron-3-Ultra. https://huggingface.co/nvidia/NVIDIA-Nemotron-Labs-3-Competitive-Coding-550B-A55B-NVFP4
Join us in New York 🗽 Come say hi if you’re around! 👋
AI policy right now, especially from the Labs, makes me think of this poem. If you believe superintelligence is near, you don't need to make hard decisions about how to build AI to make the world better. Just wait for ASI & it will decide for us. But if that doesn't happen...
SemiAnalysis measured usage meters on Anthropic and OpenAI subscription plans to estimate each plan's API-equivalent value. At mid-tier models, it found Anthropic offers roughly 5x the value of OpenAI, after OpenAI halved its $200 plan limits and introduced a $500 tier. The analysis also argues that subscriptions take a large share of inference compute while providing a small share of revenue, so their limits materially affect lab margins.
I particularly like the call stacks (got this from @dillon_mulroy) and ability to annotate code snippets
At the beginning of 2026, we had ~400 employees. Now, we're over 800. Hear from Ana Cismaru, one of our senior product managers, about why she joined Cohere and how her projects & the people keep her motivated. We're still growing. Come join us today: https://cohere.com/careers
Reflection AI introduced Beam, an agentic open model with 501B total parameters and 23B active parameters, trained end-to-end from scratch. The quoted announcement says it targets frontier reasoning efficiency and coding and agentic tasks, with full weights due this month. Clément Delangue, Hugging Face's CEO, reposted it with a welcome to the Reflection organization on Hugging Face.
AIWhy it matters: The quoted announcement names Beam's parameter scale, active-parameter count, and coding and agentic focus, which helps readers gauge where it fits among open models.
Excited about the new open model from @reflection_ai! More US models coming to Ollama!
sandboxes are the new tokens
install it and give me feedback with: claude plugin marketplace add anthropics/claude-plugins-community claude plugin install html-plan@claude-community here's an example: https://claude.ai/artifact/CUF7bzFny3XUFVLAVt9rfc
I've been working on a skill that makes better HTML plans in Claude Code. It uses simple language, shows code snippets, surfaces questions & makes mockups. Linting reduces the normal failure cases that Claude runs into. Would love your feedback before shipping more broadly!
Or build with it through the Pika API Club https://dev.pika.art/models/ideogram/ideogram-4.5?utm_source=x&utm_medium=caption&utm_campaign=260930-ideogram_4_5&utm_content=pika_labs&utm_term=post&utm_id=293eca38-fff5-4367-a98d-247d56609862
Try Ideogram 4.5 on Pika https://create.pika.art/apps/ideogram?tab=image-tools&collection=models&utm_source=x&utm_medium=caption&utm_campaign=260930-ideogram_4_5&utm_content=pika_labs&utm_term=post&utm_id=293eca38-fff5-4367-a98d-247d56609862
The new v0.6.0 release packs a lot of good stuff: - Clef support (text + vision) - High-quality support for Qwen3.8-Flash-Next - Massive performance improvement with Metal - New `llama_batch_ext` API Also the website got a nice refresh: https://llama.app
Make edit after edit with Ideogram 4.5, without impacting the original image quality. It's up to 59% less expensive on Pika and the Pika API Club. Always.
A classifier assigns an input to a fixed label set, covering binary, multiclass, and multilabel variants, such as spam versus not spam or movie genres. It encodes the input into a vector using hand-built features like logistic regression or a learned encoder such as a CNN or BERT. A linear layer then projects that vector into K logit scores, which softmax or sigmoid turns into probabilities.
Congrats to Reflection folks -- it's hard to get the first model out, and we hope you can rapidly accelerate contributions to the ecosystem from here :)
every screen of Claude and ChatGPT response should say somewhere “ceterum censeo we must pace the frontier of global machine intelligence progress” the way the uber app lobbied about taxi cartel protectionism
Listen to "naenia", the new album from Oakwood///diskrot. Blending echoes of these sliding synths from an 80s pop song with ear-wormy hooks. "naenia" is certifiably unskippable. https://suno.com/album/392d26d8-1b25-4cad-9b55-0b782fb7e654
Understanding AI has added Dan Kagan-Kans, a former managing editor of Mosaic and freelance AI journalist, as a writer, with his first post published now. The newsletter says he will be supported by a Tarbell Fellowship in AI journalism, and readers can book video calls with him on Tuesday or Wednesday.
Tue 11am (Franciscan A) Manifold Steering Reveals the Shared Geometry of Neural Network Representation and Behavior Tue 4:30pm (Imperial Ballroom) Do SAEs Capture Concept Manifolds? Wed 4:30pm (Franciscan B) Arithmetic in the Wild: Llama uses Base-10 Addition to Reason About Cyclic Concepts
ECHLO's new project is as cinematic as it is haunting. The album features beautifully realized piano moments that dance around a driving propulsive rhythm. Everything is supported by ethereal vocals that stick with you long after your first listen. https://suno.com/album/bceee842-943a-4948-a359-2ff01d5d2c87
Old Soul is an album about the growing pains of getting older. Over warm acoustic guitar, Isaiah Wallace looks back on evergreens, sweet tea, and late summer nights. https://suno.com/album/751ad762-f9c8-474f-a3b8-a284f3dc68c1
Inspired by the iconic sounds from the 80s, KakerMix uses vintage synths, punchy electronic drums and spacey reverb to make us feel like we’re entering another dimension. https://suno.com/album/4faa3913-892f-47ef-b2a7-63fd1848b5c9