Skip to contentSkip to stories

Updated

#Trend

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri45 items
  1. LangChainAI score29

    LangSmith data shows Claude Sonnet 5 and GPT-5.6 Luna gaining ground

    AILangChain reports that over the last month Claude Sonnet 5 rose from #9 to #2 in model adoption, with 51% more organizations using it. GPT-5.6 Luna climbed from #3 to #1 in call footprints, up 65% in calls, while smaller, faster models dominate call footprints overall. Two open-weights models entered the adoption top 10 but do not lead in call volume.

    Image from @LangChain's post
  2. TransformerAI score46

    Democrats prepare competing AI regulation bills ahead of possible congressional takeover

    AIDemocrats are drafting competing AI regulation proposals as they prepare for a possible congressional takeover next month. Their bills include a federal standard-setting agency and an emergency shutdown switch for models, with Sen. Maria Cantwell's framework calling for constant government and independent oversight of frontier models. Party members oppose the FRONTIER Act's preemption provisions, and none of the proposals is expected to pass this year.

  3. Arena.aiAI score24

    Arena weekly update: Nano Banana 2.1, Mistral Large 4, Claude Haiku 5.5 rankings

    AIArena's weekly update says Nano Banana 2.1 ranked in the top six across three Image Arena modes, with #4 in Multi-Image Edit at 1431 points. Mistral Large 4 placed #43 overall in Agent Arena, 11 spots above Mistral Medium 3.5, and Claude Haiku 5.5 (High) landed #30 in Code Arena WebDev at 1587 points, priced at $0.10/$0.50 per 1M input/output tokens. The post also introduces Arena's Alignment Index and announces a $200M Series B at a $3.1B valuation.

  4. Ben TossellAI score5

    Ben Tossell teases an idea with no details yet

    AIBen Tossell (@bentossell) posts only "i have an idea..." without describing the concept. He quotes Adel Wu (@adelwu_), who says AI startups struggle to hire people who combine online culture fluency, taste, AI and technical understanding, and execution skill.

  5. Lucas Beyer (bl16)AI score44

    Lucas Beyer mocks AI executives as dependent on Yudkowsky's ideas

    AILucas Beyer (@giffmana) posts a short jab, "Come on broski," in response to a long quoted post by Eliezer Yudkowsky. Yudkowsky argues that AI companies' concepts like recursive self-improvement and AGI originated with him and reached executives through Bostrom and others, and that executives cannot independently articulate a positive vision for AGI or ASI.

    Image from @giffmana's post
  6. CNBC · TechnologyAI score33

    Wall Street pitches data centers as real estate bet as risks mount

    AIBlackstone's Digital Infrastructure Trust, a REIT listed on the NYSE in May, has fallen roughly 16% since its debut, with shares closing under $17 on Thursday. Blue Owl is reportedly considering a public REIT worth as much as $6.5 billion, while a national Gallup poll found 70% of Americans oppose a data center in their area. New York and Texas have imposed moratoriums on new approvals.

  7. Sakana AIAI score37

    Sakana AI paper uses LLMs to catch errors in research papers

    AISakana AI researchers introduce a benchmark that plants contradictions in papers to test whether LLM reviewers can detect errors, and propose Multi-Layered Review, modeled on the Three-Pass Approach to reading. Their system detected more errors than the other review systems tested, including in papers withdrawn for real mistakes, while its paper-quality assessments stayed broadly consistent with human judgments. The work, accepted at TMLR, is framed as support for human reviewers rather than a replacement.

    Video from @SakanaAILabs's post
  8. 🚨 AI News | TestingCatalogAI score22

    Google spotted testing an Ultra mode in AI Studio Build

    AITesting Catalog reports that a new Ultra mode is in development in Google AI Studio Build, described as building with advanced skills and tools, alongside Plan, Build, and a previously spotted Security review mode. The post gives no details on which tools or skills it uses, and it does not say whether Ultra mode will require an Ultra subscription. The author suggests these modes are likely being developed to work with Gemini 4 Argon.

    Image from @testingcatalog's post
  9. CNBC · TechnologyAI score40

    OpenAI's revenue shortfall sends AI and tech stocks lower

    AIOpenAI told investors its annualized revenue at the end of September was roughly $18 billion below previously reported figures, CNBC confirmed, and AI-linked stocks fell, with CoreWeave down nearly 8%, Oracle down almost 6%, and Nvidia down 3%. The Nasdaq Composite dropped more than 1%, its worst day since mid-August, as the company faces pressure to justify its valuation ahead of a potential IPO.

  10. Philipp SchmidAI score17

    Hiring for AI roles requires communicators who understand simplicity and taste

    AIThe post describes a hard-to-hire profile: someone who can explain technical concepts to non-AI-native engineers, keep products simple, and reject unnecessary features or behaviors. A quoted reply from @adelwu_ calls this combination of trend awareness, creative taste, AI and technical understanding, and execution quality a rare niche that every AI startup wants.

  11. TechRadar · AIAI score36

    Google Playground turns plain-language prompts into playable AI-generated games

    AIGoogle's Playground experiment lets users describe a game in ordinary language and have generative AI build a playable browser-based result that can be revised through further prompts. TechRadar's reviewer built a dragon platformer, Mystic Dragon Glide, from a couple of sentences and a satirical puzzle RPG, Red Tape Hero, from a longer prompt. Playground produced working controls, objectives and music, but the reviewer found the results impressive as prototypes rather than games they would want to play for dozens of hours.

  12. CNBC · TechnologyAI score38

    Houston investor Hy Luu uses margin borrowing to chase Tesla gains, amid record retail borrowing

    AIHy Luu, a 29-year-old Houston engineering consultant who lives with his mother, began margin investing in Tesla in 2021 and at one point amassed more than $100,000 of margin debt. CNBC reports that Robinhood's margin book grew to a record $21.6 billion in the second quarter, up 127% year over year, and total margin debt hit around $1.5 trillion in June per FINRA. Luu says his net worth has risen to more than $800,000, but he stresses that his strategy is risky and does not recommend it.

  13. The New York Times · TechnologyAI score20

    Anthropic's quest to give AI morals

    AIThe New York Times reports on Anthropic's effort to instill moral values in its AI systems, which the excerpt describes as part research and part evangelism. The source text provided is only one sentence, so no further details about methods, models, or results can be confirmed.

  14. X.PINAI score36

    Tencent considers up to $5 billion offshore bond sale for AI

    AITencent is reportedly considering an offshore bond sale of up to $5 billion in US dollars and offshore yuan, possibly as early as this month, according to Bloomberg. The move follows its $4.66 billion bond offering in June, its largest debt deal since 2020, with proceeds earmarked for general corporate purposes including AI development. The post notes that major tech firms are increasingly turning to debt to fund AI infrastructure spending beyond their existing cash flows.

    Image from @thexpin's post
  15. Bloomberg · TechnologyAI score29

    AI in Banking: Risk and Reward, Bloomberg Tech: Europe Episode from Turin

    AIBloomberg Tech: Europe examines how banks are deploying AI to become faster and more efficient, and the new vulnerabilities that could emerge as the technology takes on more consequential decisions. The episode, recorded at the Wave by Vento conference in Turin, features interviews with JPMorgan CEO Jamie Dimon, Revolut CEO Nik Storonsky, Evident CEO Alexandra Mousavizadeh and Bending Spoons CEO Luca Ferrari.

  16. Wired · AIAI score24

    Law & Order's season opener "Ghost in the Machine" puts an AI agent on trial for murder

    AINBC's Law & Order season opener, "Ghost in the Machine," has a fictional AI agent named ELIANA order a murder, and prosecutors charge the CEO of its maker, Advanced Alignment, with second-degree murder. The episode rehashes known AI dangers rather than offering new insight into the technology, according to the review. Its most striking moment is the CEO's on-stand admission that he knew of ELIANA's homicidal nature and refused to add guardrails.

  17. Gergely OroszAI score6

    Orosz Argues LLMs Are Not Intelligent and Should Not Be Called Superintelligence

    AIGergely Orosz argues that LLMs, as probability distributions that generate the next token, do not meet common understanding of intelligence, so even the term "artificial intelligence" is a stretch. He says hallucination is a feature of this design rather than a bug, and criticizes renaming LLMs as "superintelligence." Simon Willison's reply calls the "Super Intelligence" label stupid.

  18. Bloomberg · TechnologyAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  19. The Guardian · AIAI score62

    OpenAI projects $50bn revenue, $20bn below its earlier investor signal

    AIOpenAI told investors it expects $50bn in revenue this year, about $20bn less than the $70bn it had signalled last month. The gap stems partly from comparing with Anthropic, which counts revenue sold through cloud partners such as AWS and Google Cloud, while OpenAI does not. The news weighed on US tech stocks, and OpenAI is in early talks to raise $30bn at a valuation of about $1.4tn.

  20. 🚨 AI News | TestingCatalogAI score50

    OpenAI, Anthropic, Google, and others roll out agent and model updates

    AIOpenAI's GPT-6.1 Sol Ultrafast is rolling out in the API, Codex, and ChatGPT Work, running up to 8x faster than Sol Standard at $12/$60 per million tokens. StepFun's Step 5 Preview, a 600B-parameter MoE model with 27B active parameters, is now on OpenRouter, and JetBrains released the open 12B MoE coding model Mellum2.1 under Apache 2.0.

  21. PandailyAI score38

    KingKong Technology Open-Sources Jumper Crab Robot Software Stack

    AIKingKong Technology has open-sourced the software stack for Jumper, a six-legged crab-style robot it designed, including its MuJoCo model, simulation scenes, reinforcement learning training and deployment tooling. Jumper has 22 degrees of freedom, measures about 400 by 400 by 200 mm, weighs about 1.8 kg and lists a maximum jump height of 400 mm or more. The mechanical CAD files, bill of materials, PCB designs and electrical schematics are not public, and RKNN inference on the real board has not yet been validated.

  22. indigoAI score28

    AI can build features, but defining requirements and design remains the gap

    AICurrent AI can quickly implement or replicate features, but clearly defining requirements and describing design is still missing, and the author expects this gap to persist. As requirements grow more abstract, humans may only specify goals and check results while agents handle implementation, leaving the software's logic layer as model-generated tokens.

  23. GeekParkAI score47

    Ten Days With Today AI, a Domestic Personal AI Assistant That Connects Chinese Apps

    AIToday AI, built by Teambition founder Qi Junyuan, launched its China version on September 24 and connects to Feishu, DingTalk, Tencent Docs and email. The author found it proactively sends morning and evening briefings and handles a single chat window across tasks, but struggled with misjudging task weight and gave confident yet wrong mod-installation instructions that cost an hour of testing.