Cursor lets users control computer agents from iOS app
AICursor now lets users control agents running on their computer from its iOS app. Users can check in on agents, reply to them, or start new tasks from their phone.
Updated
Updated
Items with an AI score under 20 are hidden. Show low-relevance items
AICursor now lets users control agents running on their computer from its iOS app. Users can check in on agents, reply to them, or start new tasks from their phone.
AIGitHub says some IDEs that moved Copilot agent sessions to the Copilot SDK left that activity unattributed in usage metrics, and a fix is rolling out by IDE. Visual Studio Code 1.139.0 and later has the fix now, while Visual Studio 18.12, JetBrains, Eclipse, and Xcode are expected between October and November 2026. Billing is unaffected, and missing data from affected versions cannot be backfilled.
AIOpenAI has published a repository called Openai/math, which the author reads as a sign that math problems, or any verifiable problems, are being solved. The author says OpenAI's tools exhausted their Pro token allowance on subagent tests unrelated to their main task, concluding that the work was aimed at verification for its own sake.
AIThomas Wolf's post is a short, playful reply: "how do you like them convolutions," apparently referencing Ben Affleck's comments on convolutional neural networks. The quoted context reports Affleck describing his Python scripting, understanding of CNNs and tensors, GPU work, and private looks at Google and OpenAI's video models.
AIPuffle is launched as a company agent that businesses can consider for internal agents. The post says it is easy to set up, supports multiplayer use, and is highly capable, and can be used like a set of Hermes agents for a company.
AIOpenAI is releasing a broad range of new mathematical results produced by an internal frontier model, with release guidance from the Institute for Advanced Study's Advisory Group on Mathematics and Artificial Intelligence. The results are available on GitHub at openai/math. The post itself is brief and emphasizes the results rather than hype.
AISimon Willison released llm-openai-decisions 0.1a0, a plugin that adds OpenAI's new Decisions API to the LLM command-line tool. The plugin supports yes/no, choices, and score question types, and works with the gpt-6-luna decision model, which accepts both text and image input. OpenAI charges 10 cents per million input tokens for gpt-6-luna, while Jev's rate is 4.2 cents per million, and output is not charged.
AIPollen Robotics' Microduck now runs on its first fully custom single-board computer, built by Seeed around the same Rockchip CPU. The new board adds better memory, WiFi/BT, dual NFC antennas, status LEDs, an RGB flashlight, and improved cooling and boot speed, replacing the earlier Radxa-based prototype stack.
AIVercel's AI Gateway now reruns an agent's decision on a fallback model when the primary model's confidence falls below a user-set threshold, now in beta. Rauch calls the simple feature highly impactful for at-scale AI decision-making.
AIPyTorch's blog describes a Triton-based implementation of Table Batched Embedding (TBE) forward and backward kernels for recommendation-system embedding lookups, which the post says outperforms legacy CUDA kernels on these workloads. On B200, an updated CUDA bounds-check step reaches up to 1.24x speedup on that component, and an optional forward-side preprocessing path cuts combined latency from 79.537 ms to 66.183 ms (−16.8%) on a large configuration.
AIA post by Will DePue titled "Fable 5.1's list" presents 100 mathematical results and says 59% were released today, 87% AI and 13% human. The list includes items attributed to OpenAI, Anthropic, Google DeepMind and human mathematicians, each marked by a colored indicator, and it describes many entries as formalized in Lean or as openai/math family numbers. The post supplies no independent verification of these claims.
AIVercel says its AI Gateway can rerun an agent's decision on a fallback model when the primary model's confidence falls below a user-set threshold. The feature is available in beta.
AIOpenAI says its models solved several important math problems and that the team documented the work in a blog post. The post calls this a remarkable new era of scientific progress, but it does not name the specific problems here.
AIReflection CEO Misha Laskin argues that closed AI models are like renting an apartment, while open models let users own intelligence as AI adoption grows. He says the only way to own intelligence is if it is open. Reflection is preparing to release Beam, its first open-weight model, in a podcast discussion with its co-founders.
AIindigo (@indigox) says Claude now works directly inside Google Docs, Sheets, and Slides, a feature they had been waiting for. They contrast this with Gemini's integration in Google's own apps, which they call a waste.
AIOpenAI is releasing a broad range of new mathematical results produced by an internal frontier model. The company says it consulted the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study and drew on its advice and public recommendations for how the results are released. The results are available at
AITeknium announced Hermes Index, which combines scores from the new HermesBench and three other agent benchmarks. The index aims to help Hermes Agent users find the best model at a given time and at a given price point. It was introduced by Nous Research as a way to inform model choice and show labs their performance in Hermes.
AIThe source provides only a title, a source label, and a comments link, with no body text describing the AI mathematics progress. Specific models, results, and benchmarks cannot be verified from the material provided.
AIClaude Opus 5.5 leads the benchmark with a score of 63.31 at $4.99 per task, ahead of GPT 6 Astra at 56.25 ($11.61) and Sonnet 5.5 at 53.14 ($2.82). At the low end, DeepSeek V4.1 Flash scores 36.91 at $0.26, and Ling 3.0 Flash scores 21.56 at $0.054.
AINous Research's Hermes Bench runs every model through the same harness, with reasoning set high where offered. The index reports each model's mean score and mean cost per task across four suites.
AINous Research says its evaluation uses an index of four benchmarks, including its newly introduced Hermes Bench. The other three are Terminal-Bench 4.0, Terminal-Bench-Science, and SkillsBench.
AINous Research has introduced Hermes Index, an average of benchmarks that measures model performance within Hermes Agent. The index aims to help users choose between models and show AI labs how their models perform in Hermes.
AIPenguin Mail 1.0.5 is a free, GPL-3.0-or-later Rust email and calendar client for x86_64 Linux that handles Gmail, Microsoft, IMAP and POP3 accounts in one inbox. It uses GnuPG for OpenPGP and S/MIME, and its assistant is off until a user picks a model, which can run locally through LM Studio or Ollama.
AIMistral Large 4 is now available in OpenCode. The post says there is a 50% discount for the next two weeks.
AIOpenAI has released its Decisions API in public beta, using GPT-6 Luna to classify text and images, route requests, and score inputs. It is reported to run about 10× faster than the Responses API, starting at $0.10 per million input tokens with no output or cache charges.
AIGitHub reports that Git events on the platform rose from 218.2 billion to 473.3 billion per month between September 2025 and August 2026. It says agent workloads push write throughput and merge contention beyond what its current replica-based architecture handles well, so it is separating durable storage from compute while GitHub keeps running. The article states internal benchmarks reached up to 35 times higher write throughput.
Why it matters: The post links rising Git event volume to specific architectural bottlenecks, showing why agent workloads strain write paths and how GitHub plans to separate storage from compute.
AIand those files can also open inside Claude. In Google Workspace, Claude appears in a sidebar next to the open file, reads its contents, and edits it in place, with the option to approve each edit before it is applied.
AIGoogle's Antigravity agent can take an Android app from prompt to a real device, using the Stitch MCP and Android CLI plugin. The agent pulls designs, builds native Jetpack Compose components, verifies them in the emulator, and runs the final build on a physical phone.
AIAt IROS 2026 in Pittsburgh, at least 17 dexterous hand companies exhibited, 11 of them Chinese, with WUJI reportedly shipping 800 to 900 units a month. Boston Dynamics skipped a booth but released a video on the final day of a new four-finger, 13-degree-of-freedom Atlas hand, down from 7 DOF on its previous gripper. Hand makers are also selling capture gloves and data services, since labs need far more demonstrations than the hardware alone provides.
AIOllama announced that Google DeepMind's EmbeddingGemma 2 is now available on Ollama. The author describes it as made for consumer devices and multimodal, and gives the command ollama pull embeddinggemma-2 to download it. The quoted DeepMind post says the model is a natively multimodal open model for on-device embeddings that unifies code, images, audio, and video in a shared space.
AIGoogle's newest image model, Nano Banana 2.1, is now available on fal. The post says it brings faster generation, better subject consistency, improved visual design, and mask-based editing.
AIFireworks' Head of AI Developer Education fine-tuned a 9B Jev-style classifier on Fireworks using plain SFT, public datasets, and two simple tricks. The full recipe is open source so developers can train their own decision classifiers without a frontier lab budget.
AIReplit can now use context across your projects to flag existing projects that match a new idea before you spin up a duplicate. Users add a custom instructions skill to their workspace to enable the check, which targets duplicate work and project sprawl.
AIAnthropic's Thariq says Claude will increasingly use cloud-based "brains" while operating on users' computers through "local hands." He points to a Latent Space podcast discussion on how this is being implemented in Claude Code.
AIGamma announces Gamma 5, a rebuilt version of its presentation platform that it says overhauls how the product thinks, designs, and edits. The company says the release addresses concerns that AI-generated output looked too similar across tools, and it revamps the agent, design tools, editing, import, export, and connectors. Gamma says teams can build presentations, docs, social assets, and graphics that follow their brand or a new aesthetic, using every frontier and image model under the hood.
AISGLang now supports Kandinsky 6.0 Video, which generates video and synchronized audio together from text or an image. The model comes in Lite (3B) and Pro (29B) sizes, with built-in super-resolution up to 1920×1080. A sample sglang serve command for the Pro distilled model is included.
AIDuring the ongoing Ebola outbreak in the Democratic Republic of Congo, WHO AFRO partnered with Google Earth AI to find transmission blind spots faster. Using Google's Geospatial Reasoning agent prototype, the team identified 48 exposed settlements and more than 45,500 at-risk people in minutes, a process that normally would take weeks.
AIGoogle Earth AI combines environmental signals and other data sources with AlphaEarth Foundations, a Population Dynamics Foundation Model (PDFM), and a prototype Geospatial Reasoning agent. Researchers ask questions such as where a disease is likely to spread next, and the system automatically gathers relevant models and datasets to build a prediction model. By combining satellite views with population patterns, the tool aims to reveal hidden risk factors and identify issues earlier.
AIGoogle Earth AI, according to new research, can help communities respond to public health crises more quickly and proactively. The post says it combines behavioral trends, geospatial AI models, and other insights beyond simple statistics to help public health teams understand complex issues and bridge reporting gaps. The aim is to shift emergency response from reactive management toward proactive prevention.