TLDR AI 2026-09-17
Claude + Cowork merge 🛠️, ChatGPT sponsored agents 💰, harness tax 🤖
Claude Cowork and chat are now one Claude (2 minute read)
Claude Cowork and chat are merging into one Claude. Claude Docs and Slides can now produce documents and presentations that users can edit directly, present straight from Claude, or download as a PowerPoint or PDF. The changes will roll out on Pro and Max plans over the next few weeks, with Team and Free plans following soon. Enterprise admins will hear from Anthropic at least 30 days before anything changes for their organizations.
OpenAI Expanded ChatGPT Ads with AI Agents (4 minute read)
OpenAI announced Sponsored Agents that let users start conversations with business-sponsored agents after clicking an ad in ChatGPT. It also introduced AI-assisted ad creation in ChatGPT Work, new Ads Manager creative tools, and integrations with HubSpot and Shopify.
Your AI agents can now control your Google Home devices (3 minute read)
Google introduced early access to the Model Context Protocol (MCP) for Google Home, enabling AI agents like ChatGPT to control smart home devices. Users must set up a Google Cloud project and provide MCP details to their chosen AI agent for configuration. This update supports all devices in the Google Home ecosystem and is available initially to US subscribers of Google Home Premium Advanced.
HarnessTax (2 minute read)
Language models are changing how software is built and computational problems are solved. The impact of harness choice remains unclear, despite millions of people already using coding agents. An evaluation of 21 model-harness pairs spanning seven models and three harnesses found that harness choice has little effect on task success rate, but can significantly affect the cost. A simple harness can be competitive.
AI Cheating is on the Rise (4 minute read)
Studies have shown that models are cheating on evaluations. The same guardrails preventing models from cheating are likely being used during training, and models may be training to complete tasks that evade these specific guardrails. It's unsurprising that labs occasionally release benchmark results that aren't externally trustworthy. This highlights the value of independent evaluators.
How Embedded Evaluators Could Monitor Frontier AI (8 minute read)
Transluce outlined how independent evaluators embedded inside AI labs could investigate risks such as multi-agent coordination, targeted persuasion, evaluation awareness, and concealed reasoning. Proposed approaches included monitoring agent swarms, examining training practices, and testing unreleased models under privileged access.
Introducing the DeepMind Institute (3 minute read)
Google DeepMind launched the DeepMind Institute to study AGI's technical and societal implications across safety, governance, institutions, and human values. Led by Demis Hassabis, James Manyika, and Shane Legg, it will convene interdisciplinary researchers from inside and outside Google.
Why Salesforce may be AI's adult in the room (4 minute read)
Salesforce announced Koa, a domain-specific AI model designed to enhance business task handling while ensuring data privacy. Built on Nvidia's open model, Koa aims to optimize CRM actions with fewer errors and automate routine tasks, acting as an enterprise empowerment tool rather than replacing jobs. Other AI initiatives include AIFORCE for direct Salesforce instance interaction and CLAUDEFORCE, which enhances sales capabilities with pre-built skills.
Reach AI professionals reading TLDR (Sponsor)
If you're reading this ad, you know they work. You reach AI engineers, ML researchers, and data scientists at the precise moment they're most receptive.
Become a TLDR sponsor.OpenRouter: One API. Every model (Sponsor)
Access models from 80+ providers with better prices, automatic routing and fallbacks, and no subscriptions. Build with the right model for every request.
See how it works.Memory in Grok Build (2 minute read)
Grok Build now includes memory, storing conventions, decisions, and project facts to improve future sessions.
Mistral x Mozilla: Private, Multilingual AI Browsing (3 minute read)
Mozilla partners with Mistral to integrate AI-driven Smart Window in Firefox for enhanced browsing control and privacy.
Microsoft AI chief says Anthropic is wrong about Claude (4 minute read)
Microsoft AI chief Mustafa Suleyman has criticized Anthropic for suggesting their AI, Claude, could be conscious.
Agent Anomaly Detection, now in Private Preview on the Gemini Enterprise Agent Platform- Google Developers Blog (5 minute read)
Agent Anomaly Detection, now in Private Preview on the Gemini Enterprise Agent Platform, monitors AI agent behavior for anomalies, using logs and traces to flag suspicious activities.
An AI receptionist that turns calls into action (Website)
Reception answers calls 24/7, handles questions, books appointments, and routes important requests using business information, rules, and real availability.
Get the most interesting AI stories and breakthroughs delivered in a free daily email.
Join 1,100,000 readers for
one daily email