Codex Cloud re-shipped and now connects to tailnets via Tailscale
AIOpenAI quietly re-shipped Codex Cloud, which its Tibo says is now "pretty good." Codex Cloud can now securely connect to resources on a user's tailnet through Tailscale.
Updated
Updated
AIOpenAI quietly re-shipped Codex Cloud, which its Tibo says is now "pretty good." Codex Cloud can now securely connect to resources on a user's tailnet through Tailscale.
AIArtificial Analysis is adding trusted-access models to its Cyber Index, starting with GPT-6 Sol (Daybreak Blue, max), which is available only through OpenAI's Daybreak program. The model hits no safety blocks across the Index and scores 32 points higher overall than the publicly available GPT-6 Sol (max), with its largest gains on CyberGym-E2E.
Why it matters: The source shows how safety refusals shape cyber benchmark scores, with the trusted-access model's gains concentrated on CyberGym-E2E, useful for comparing guarded and unguarded models.
AIHarvey LAB-AA v1.1 adds hallucination checks that audit every model deliverable against task source documents, with material hallucinations zeroing a task's score. GPT-6 Astra averaged 0.03 material hallucinations per task across 120 tasks, while Gemini 3.8 Flash averaged 13.96. Harvey uses GPT-6 Sol (high) as the hallucination checker, separate from its three-judge rubric panel.
AILangChain built Restock, a sample office-supply agent that runs in Slack on Managed Deep Agents and pays through Stripe's Link wallet. The agent searches products, builds a cart, and pays over the Machine Payments Protocol, with the user approving the purchase in Slack and the payment in Link. The post uses a pens order at $22.18 to show the flow from request to confirmed order.
Why it matters: The post walks through how an agent handles search, budget limits, Slack review, and Link approval, showing where each control sits outside the model.
AIOpenAI's Tibo confirmed that the GPT-6 rollout has landed across all accounts and asked how things are going so far. In a previous post on Day 3, he said GPT-6 is coming to Chat, and that Codex and ChatGPT Work reached 40M active users, a new high.
AIOpenAI published 722 math manuscripts covering 372 result groups in a new GitHub repository, openai/math, all produced by an unreleased internal model. The author describes the results as including a near-Riemann hypothesis claim pushed to 0.875, and notes that 25 Fields Medal winners criticized the company's approach to AI math research.
Why it matters: The piece traces how AI math results moved from benchmarks to open problems, offering context on verification and the mathematicians' pushback.
AIScott Aaronson reports, based on his sources, that some AI companies have begun discreetly investigating whether their latest internal models can break important cryptographic protocols and primitives. He notes that cryptography is conspicuously absent from OpenAI's list of 376 papers, and the quoted post adds that the US government has censored academic quantum cryptanalysis results.
AIMax Zeff highlights new data showing OpenAI's growing momentum with business customers, though the post itself provides no specific figures. The post refers to a quoted post from Angel Au-Yeung, which supplies the underlying context but is not reproduced here.
AIAndrew Curran posted a brief "Updated again." with no further details. The post links to background from @0xdoug on OpenAI problem #109 (integer multiplication), which reports tightening the bound from κ = 2⁻¹⁸² to κ = 2⁻³⁴, a roughly 48-million-fold improvement.
AICodex is sending reset cards today. The quoted post says GPT-6 is coming to Chat, Codex and ChatGPT Work reached 40M active users, and a banked reset is being loaded into paid accounts.
AIMicrosoft has repositioned Windows as the home for hybrid intelligence, where AI agents can run locally or in the cloud. The announcement covers four layers: MXC reaching general availability for agent isolation, local models such as MAI Code 1.1 Flash, Copilot on Copilot+ PCs gaining local context and actions in coming months, and new hardware including RTX Spark PCs and DGX Station for Windows.
AIMemory has overtaken compute as the central control point in IT, according to The Next Platform, as generative and agentic AI drive demand for DRAM, HBM, and flash. Server DDR5 memory now sells for roughly 9X to 13X its November 2022 street price, while a 30 TB enterprise SSD costs 6X to 7X more. HBM pricing has risen only about 1.6X since the GenAI boom began, the article says.
AIOpenAI is rolling out GPT-6 to ChatGPT's over 1.2 billion weekly users, adding Intelligent UI, which lets replies include charts, buttons, forms, and interactive tools. The post's image cites tiered access, with Free/Go and Plus/Pro/Business/Enterprise sharing the Sol and Luna model splits, and says the feature is progressively rendered as the model generates it.
AIAccording to the source, OpenAI released a set of results produced by an internal frontier model on 722 math problems, without advance warning or peer review. The excerpt provided does not include full details of the methods or verified outcomes.
AIOpenAI announced on X that it is releasing a series of new mathematical results produced by its internal frontier model. The material is reported as 722 manuscripts posted to GitHub, and the excerpt provides no further detail on the specific results.
AIA company shows goodwill: Claude Haiku 5.5 price cut 90%, priced in line with GPT 6 Luna. Capabilities are very strong... This may be the best cost-performance model.
AIOpenAI disrupted two AI-enabled influence operations that used false-front journalists and a think tank to spread geopolitical messaging. The source does not provide further details on the operations' scale, targets, or attribution.
AIEpoch gave six models 11 real work tasks from its own operations, including graphic design, data insights, and research design, and graded outputs against employee standards. Fable 5.1 and GPT-6 Astra led on average task performance, reliably handling well-defined work such as coding and computational analysis. The report finds that all models still fail on open-ended judgment, including matching Epoch's standards, designing informative experiments, and generating diverse ideas, so the authors conclude AI cannot yet replace workers at Epoch.
Why it matters: The report separates well-defined task reliability from open-ended judgment failures, which benchmark scores on easily verifiable tasks would miss.
AIOpenAI is gaining ground on Anthropic as the AI price war intensifies. Both companies need to show Wall Street they have sustainable businesses ahead of planned IPOs.
AIIn a Wednesday letter to OpenAI board members, three recently fired employees urged the company and its competitors not to advance work that would decrease AI monitorability. They also claimed their firings are chilling those who remain at OpenAI.
AIScott Aaronson describes an AI-generated proof of the UGC conjecture that invents an entirely new, bizarre code with a noise test and a crazy recursive construction. He notes that mathematicians increasingly encounter this kind of "alien craziness" and predicts it will become routine.
AIEthan Mollick shares early first-hand accounts from mathematicians grappling with hundreds of AI proofs released by OpenAI. He highlights problems solved in ways no human has yet understood, raising questions about what it means to know something. The linked Scott Aaronson post quotes a researcher, Dana, describing the proofs as unclear and hard to read without AI help, with some possibly verified by a Lean certificate.
AIOpenAI says Codex and ChatGPT Work together reached 40,000,000 users, a new high for the company. To celebrate, paid accounts will receive a banked reset starting end of day Pacific time.
AIThe main post is a short reply saying reports of a company's death have been greatly exaggerated, with no details about products or figures. The background post from @h_nilforoshan reports that OpenAI's Decisions API, billed as a "Jev-killer," was benchmarked against Jev for HiringCafe, which serves 2.5 million users. On the task of scoring job-description relevance from 1 to 10, the author reports OpenAI costing 2x more and performing 5-10% worse.
AIOpenAI announced on X that GPT-6 and Intelligent UI are now rolling out to all ChatGPT users, after GPT-6 Astra, Sol and Luna were previously limited to ChatGPT Work and Codex. Intelligent UI lets GPT-6 combine text, images and interactive elements such as charts, clickable buttons and forms, with a mahjong learning example shown.
AIOpenAI says ChatGPT now offers Intelligent UI, which returns answers with fully interactive user interfaces. GPT-6 is rolling out in ChatGPT, powered by GPT-6 Sol for Plus, Pro, Business, and Enterprise tiers and GPT-6 Luna for Free and Go tiers. The rollout begins globally today in the Chat tab for paid tiers, with Free and Go tiers following tomorrow.
AIA $200 Claude subscription can deliver about $12,000 of Opus 5.5 tokens at API pricing, according to comments quoted in the post. The post argues that when comparing against ChatGPT's subscription, the value gap is not close.
AIMark Chen recommends trying Intelligent UI, a ChatGPT feature where every completion can create a custom interface. He calls it useful for learning new things. The quoted OpenAI post says GPT-6 and Intelligent UI are rolling out in ChatGPT for everyone, providing interactive answers and on-the-spot tools.
AINoam Brown says LLMs have crossed a threshold by surpassing top human experts on some research problems, a jump that makes the recent surge in math results feel sudden. He expects breakthroughs in other domains to follow as models keep improving, though capabilities remain jagged and often still weaker than humans.
AIguys. what the fuck. i have auto reload on and am just trying to pay you all my dollars. yet me agents keep getting interrupted doing multi-hour long tasks for usage limits. a product im paying thousands a week for should not do this!!! oai and ant plz fix @bcherny @romainhuet
AIPresident Trump announced a White House-led "Super Intelligence Force" that will spend 120 days examining AI risks and federal responses, and signed an executive order directing agencies to use "Super Intelligence" and "SI" instead of "Artificial Intelligence" and "AI." The order asks officials to develop a possible new federal definition within 60 days, but the source says the change is linguistic rather than architectural. Critics quoted in the article argue that renaming does not change the technology itself.
AIEthan Mollick had early access to Intelligent UI and found it a welcome change from long blocks of text. He suggests interfaces will increasingly be built on demand for each user's problem. The quoted OpenAI post says GPT-6 and Intelligent UI are rolling out in ChatGPT for everyone, delivering fast, interactive answers with visual explanations and task tools.
AIAnthropic has released Claude Haiku 5.5, its cheapest and fastest small model, priced at $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100K tokens. It keeps a 1M token context window, up to 128K output tokens, and is generally available on the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Anthropic reports 72.4% on OSWorld 2.1 (offline subset) versus 15.7% for Haiku 4.5, and the article notes that non-default temperature, top_p or top_k values return a 400 error.
AIOpenAI's ChatGPT Intelligent UI can produce visual outputs for requests such as breaking down a bicycle's design, planning a dinner, building a wardrobe, or splitting a dinner bill from a receipt. The post says these visual responses are easier to understand and more enjoyable to use.
AIGPT-6 with Intelligent UI begins rolling out globally in the ChatGPT Chat tab for Plus, Pro, Business, and Enterprise users today. The rollout expands to Free and Go tiers starting tomorrow. Plus, Pro, Business, and Enterprise get GPT-6 Sol, while Free and Go get GPT-6 Luna.
This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”
AIOpenAI's official ChatGPT account announced Intelligent UI, which lets ChatGPT answer with fully interactive user interfaces. The post says this makes complex topics easier to learn and lets ChatGPT quickly create tools for tasks in the moment. It also states that GPT-6 is coming to ChatGPT for everyone, without giving a date.
AIZvi Mowshowitz argues that relying on current AI to automate its own alignment work is a near-suicidal plan, noting that no one has a better one. He reports that labs fear they cannot even execute the first step, and that the third annual The Curve conference was held under Chatham House rules.
AIOpenAI's GPT-6 Luna Decisions model is now available on OpenRouter, letting apps choose the right model, tool, or action from text, JSON, or images. It returns typed answers with probabilities, costs $0.10 per million input tokens, has a 1M-token context window, and output is free.
AIOpenAI has rebuilt its plugin submission flow to provide clearer review feedback to developers. Developers can submit their plugins for review through the submission page on the OpenAI developer site.