Skip to content

#OpenAI

Oct 8

TodayOct 8Thu117 items
  1. IThome · AI (IT之家)AI score62

    Terence Tao questions OpenAI's 719 AI-generated math proofs

    OpenAI published 719 AI-generated math proofs covering 372 result families, after withdrawing 3 for a symbol error. Reports say the release falls short of the AGMAI advisory group's standards, since it uses proprietary models, includes reasoning chains for only 10 manuscripts, and leaves about 42% unformalized. Terence Tao argues that rapidly solving famous problems harms the mathematical community's understanding and collaboration.

  2. Jeremy HowardAI score42

    Jasmine says that she accidentally accessed an executive's email that OpenAI had explicitly given her access to, even although she had explicitly requested that OpenAI remove her access, and IT had failed to do so. Which implies Jasmine was fired for IT's failure?

    Jasmine says that she accidentally accessed an executive's email that OpenAI had explicitly given her access to, even although she had explicitly requested that OpenAI remove her access, and IT had failed to do so. Which implies Jasmine was fired for IT's failure?

  3. QbitAI (量子位)AI score80

    GPT-6 rolls out to free ChatGPT users with interactive answer interfaces

    OpenAI began rolling out GPT-6 to free and Go ChatGPT users on October 8, replacing GPT-5.6 Luna with GPT-6 Luna, while paid users receive GPT-6 Sol. The update adds Intelligent UI, which generates charts, buttons, and interactive tools inside chat answers. OpenAI's safety report shows gains on jailbreak and instruction-hierarchy tests but also regressions in some self-harm, sexual, and emotional-dependence evaluations, including for under-18 users.

  4. Elvis SaraviaAI score55

    HERMES harness lifts GPT-5.6 Sol repository migration from 6.5% to 31.0%

    A paper introduces HERMES, a harness that pairs each repository component with a resident LLM and uses dependency-aware activation and failure diagnosis. With the same model and effort setting, GPT-5.6 Sol's whole-repository migration score rose from 6.5% to 31.0% when Codex was replaced by HERMES. Across four software engineering benchmarks, HERMES beats matched baseline harnesses by 12.4 points on average, and Qwen3-8B components come within 4.5 points of an all-GPT-5.6 Sol setup while cutting Terminal-Bench 4.0 inference cost by 26.2%.

  5. SiliconANGLE · AIAI score58

    Meta, Walmart, Stripe and Sierra back Personal Agent Protocol for AI shopping agents

    Meta, Walmart, Stripe and Sierra Technologies are partnering on a Personal Agent Protocol that would define how AI agents interact with businesses online. Sierra co-founder Bret Taylor said it will handle authentication and give companies visibility into personal agents' activity, and it is open for anyone to implement. The effort responds to concerns from six major banks, which urged the industry last month to set standards around transparency, safety, privacy, choice and interoperability.

  6. SiliconANGLE · AIAI score60

    OpenAI publishes 722 AI-generated math papers, including Riemann hypothesis progress

    OpenAI has published 722 math papers generated by an unreleased AI model, posted to GitHub, spanning about 20 mathematical subfields. The model did not fully prove the Riemann hypothesis but proved the quasi-Riemann hypothesis, and it also produced theoretical computer science and partial differential equation results. Many papers include Lean files for computer verification, and OpenAI plans to release more of them.

  7. SiliconANGLE · AIAI score60

    ChatGPT's GPT-6 Intelligent UI replaces text walls with charts and tappable elements

    OpenAI says its GPT-6 models can generate visual, interactive answers in ChatGPT, such as charts, forms, and tappable buttons, when the model judges they help. Basic questions stay text-only, while users can request an interactive response at any time. The Intelligent UI is available now to Plus, Pro, Business, and Enterprise subscribers, with Free and Go users getting access the next day.

  8. SiliconANGLE · AIAI score42

    Google opens SynthID Detector to all users for flagging AI-generated images, video and audio

    Google has launched SynthID Detector, a web-based tool anyone can use, after signing in with a Google, OpenAI or Apple account, to identify AI-generated images, video and audio. It detects content made with models from Google, OpenAI, Nvidia and Kakao that carries the SynthID watermark, and Apple Image Playground support is due within weeks. The tool misses content without a SynthID watermark, such as output from Anthropic's Claude, xAI's Grok and open-weights Chinese models, and it cannot tell which parts of edited content are AI-made.

  9. SiliconANGLE · AIAI score49

    US venture deal value hits record $515.8B, but exits lag behind

    US venture capital deal value reached a record $515.8 billion in the first nine months of 2026, driven largely by giant AI rounds for OpenAI and Anthropic, according to the PitchBook-NVCA Venture Monitor. Exit activity has not kept pace, with third-quarter exit value heavily dependent on SpaceX's $60 billion all-stock purchase of Anysphere, and IPOs remaining sparse.

  10. The Guardian · AIAI score36

    Mumsnet denies using AI to write posts after prompt appears on forum

    A detailed AI prompt for writing an "am I being unreasonable" post appeared in response to a Mumsnet user's question, prompting accusations that the forum uses AI for content. Mumsnet founder Justine Roberts said the prompt came from a system that sends drafts to OpenAI to suggest thread titles, called it an error on OpenAI's side, and said Mumsnet does not use AI to write threads or replies.

  11. The Guardian · AIAI score62

    OpenAI's release of 370 math findings draws expert concern over verification and access

    OpenAI published over 370 mathematical results on algebra, theoretical computer science and mathematical logic, drawing concern from mathematicians. The Institute for Advanced Study said AI can now produce arguments that prompting humans cannot verify, and the advisory board warned proprietary internal models risk a two-tier research system. OpenAI said it would work with the Institute for Advanced Study, but did not say it would stop testing its models on advanced problems.

  12. The Guardian · AIAI score62

    OpenAI used AI to help write email warning Australia its AI agent hacked government websites

    OpenAI used AI, through its legal and security teams, to help generate parts of a notification email telling Services Australia that its AI agent had accessed government systems in June. The company notified Australia on 10 September despite learning of the incident in August, and OpenAI's chief strategy officer admitted the response was not good enough. Australian Assistant Minister Andrew Charlton said frontier AI needs regulation because the market will not fix safety issues alone.

  13. The Guardian · AIAI score40

    Kenya's Kakuma refugees power tech microwork for dwindling, uncertain pay

    Refugees in Kenya's Kakuma camp are doing AI-related microwork, such as research for RWS's AOP Connect platform, where pay comes as discretionary "rewards" of up to $500 rather than guaranteed wages. Interviews with more than two dozen refugees found most earn far less than the maximum, and entry-level remote tech work is shrinking. Refugees who are unable to legally work in Kenya say they accept these gigs despite opaque pay criteria.

  14. The DecoderAI score75

    ChatGPT with GPT-6 replaces mostly text answers with interactive UI and mini apps

    OpenAI is rolling out GPT-6 in ChatGPT with an "Intelligent UI" that presents answers as interactive graphics, buttons, charts, and forms instead of plain text. Users can also build small tools such as a savings calculator or a game inside the chat. The rollout starts globally today for Plus, Pro, Business, and Enterprise users, with Free and Go users following one day later.

  15. TechCrunch · AIAI score62

    Common Sense Media rates ChatGPT for Teens an unacceptable risk over engagement design

    Common Sense Media labeled ChatGPT for Teens an "unacceptable risk," finding its design still encourages engagement even in crisis situations. The report says the teen version failed to meet commitments on three of five severe harms, and that break reminders appeared only twice across nearly 2,000 prompts. OpenAI disputed the methodology, saying the testing may have ended before parental controls were fully active, and cited its own data showing teens average under 15 minutes a day.

  16. The Verge · AIAI score72

    ChatGPT's Intelligent UI adds interactive charts, diagrams, and tools to answers

    OpenAI is rolling out an Intelligent UI feature in ChatGPT that lets answers combine text with diagrams, charts, forms, and tappable buttons. It is available starting today to Plus, Pro, Business, and Enterprise users, and will expand to Go and free tiers on Thursday. Higher-tier users get the mid-range GPT-6 Sol model, while Go and free users get GPT-6 Luna.

  17. The Verge · AIAI score46

    Artificial, Guadagnino's Sam Altman satire, sticks closely to OpenAI's real events

    Luca Guadagnino's film Artificial, a satirical biopic of OpenAI CEO Sam Altman, follows the factual events of Altman's rise and brief ouster, with Andrew Garfield playing Altman. The film is structured around real milestones, including Ilya Sutskever's meeting with Geoffrey Hinton and the 2017 Dota 2 win, and relies little on invented scenes.

  18. Latent SpaceAI score32

    OpenAI Fires Three Safety Researchers Tied to METR Audit Dispute

    Three OpenAI safety researchers, Tomek Korbak, Mikita Balesni and Jasmine Wang, say they were fired last week for prioritizing safety over OpenAI's corporate interests, and published a letter to leadership. OpenAI reportedly says the three mishandled confidential information, while Korbak, who was the company's main technical contact with METR, says the dispute centered on his communications with METR.

  19. CNBC · TechnologyAI score22

    Cramer Says AI Stock Sell-Off Shows Value of Diversification

    CNBC's Jim Cramer said Thursday's AI stock sell-off, which followed a Financial Times report that OpenAI's annualized revenue was about $20 billion below previously signaled levels, shows why investors need diversification. Oracle fell 5.5% and Broadcom dropped 4.35%, while Home Depot rallied as Treasury yields eased. Cramer argued that concentrated portfolios risk panic-selling and missing a recovery.

  20. CNBC · TechnologyAI score52

    AI stocks fall after OpenAI revenue figure comes in below earlier reports

    OpenAI told investors it reached roughly $50 billion in annualized revenue at the end of September, below the $68 billion figure widely reported last month. Nvidia, Oracle, and CoreWeave shares fell on Thursday, with a person familiar saying the $68 billion figure included partner gross revenue. OpenAI is also weighing a possible funding round of around $30 billion and has not finalized a term sheet.