TLDR AI 2026-09-30
OpenAI Dots 🟡, GPT-6.1 Sol ⚡, software factories 🏭
Meta Muse Just Changed the Internet. Now What? (Sponsor)
AI agents just became mainstream. Meta's Muse is putting them in the hands of billions, with free access and massive distribution across Facebook, Instagram, and WhatsApp.
Muse now accounts for~70% of agentic browser traffic observed by HUMAN. But, how do you capture this opportunity while preventing fraud and abuse?
Join HUMAN for Meta MUSE 101 to learn: 🔍 What just happened: How Muse works and how it's changing digital traffic. 🛡️ Why it matters: What agentic traffic means for your business. ⚙️ What to do next: How to identify, verify, and safely enable AI agents.
Save Your Seat
Introducing dots (10 minute read)
OpenAI introduces "Dots," AI-driven agents powered by GPT-6 Astra, designed to handle tasks autonomously using their own cloud computer. Dots integrate with over 4,000 apps and platforms like ChatGPT, Slack, and Teams, enabling seamless workflow and personalized assistance. They ensure user control with built-in safety measures, allowing for proactive task management while adapting to user goals and feedback.
The Future Is for Everyone: Muse for Small Business (3 minute read)
Meta introduced Muse for Small Business, an AI tool to automate tasks for entrepreneurs using Facebook and Instagram analytics, ad accounts, and other tools. The AI can streamline operations and integrates with platforms like Canva for seamless brand management. Muse aims to empower small businesses by saving time and boosting productivity, with free and subscription options available.
Introducing GPT-6.1 Sol (6 minute read)
GPT-6.1 Sol offers near-Astra intelligence at a fifth of the cost, delivering significant improvements for professional tasks and reducing costs for developers. It excels in coding benchmarks, matches Astra's performance on complex PDF queries, and improves business workflow automation while offering enhanced factual accuracy. Available via OpenAI API and multiple user tiers, it provides cost-effective solutions across tasks with substantial safety and transparency enhancements.
The world's best gradual disempowerment model organism: Frontier AI labs (32 minute read)
We now have slightly to moderately superhuman, legibly impressive general AIs across many domains. However, we have made relatively little progress on technical alignment or the moral dimensions of building superintelligence. Government regulation has been ad hoc and inconsistent. AI labs plan to use AIs to solve many facets of the alignment problem as there's too much work for humans to check, and the amount and share of unchecked work seems to be growing.
Segmentation Drives Market Share Wins in AI (2 minute read)
Anthropic and OpenAI have boosted their revenue by implementing strategic customer segmentation and pricing changes. Anthropic introduced enterprise metered billing, while OpenAI slashed Luna's price by 80%, bringing its revenue run rate close to $70 billion. Both companies aim to hit $100 billion in revenue by year's end, with major clients like Amazon and Google still undecided on long-term commitments.
👨💻
Engineering & Research
One week out: can you prove what your agents shipped? (Sponsor)
GitLab Transcend is live October 6. See AI-assisted, and autonomous workflows run on one platform with humans setting direction, plus a clear read on climbing AI bills and open-weight model alternatives. Hear from Gene Kim of IT Revolution, Hilton, and Thrive Market. Free, and registration includes the $1,600 Enterprise AI Summit livestream.
Register for the livestream →
d1 (2 minute read)
d1 is a decision model that outperforms Jev on Hugging Face's Decision Index. It is built for fast, structured decision-making in software environments. Decision models are a good fit when the answer is a structured decision, such as in classification, categorization, routing, scoring, or content moderation. d1 is available on the Liquid API and will be available on OpenRouter soon.
MCP Events (14 minute read)
MCP Events lets ChatGPT subscribe to updates from an MCP server. Users choose what to monitor and what ChatGPT should do when an update arrives. ChatGPT supports webhook delivery and callback verification from the draft MCP Events specification. Polling, streaming, and the draft's `gap` and `terminated` control notifications are not supported.
Decisions API (1 minute read)
OpenAI's Decisions API, powered by GPT-6 Luna, is now available in limited preview. Users define questions and possible answers, and the model classifies content, routes requests, or chooses an agent's next action. Text and images can be used as context. Preview access is limited to selected API customers for testing. A broad release is planned in the coming days.
Sign in with ChatGPT (6 minute read)
Sign in with ChatGPT lets people use their ChatGPT identity to access supported external applications. It can be used to sign in on a participating partner site or when connecting an app from the ChatGPT plugin directory. Sign in with ChatGPT is available globally to authenticated ChatGPT users. Initial participating partners include Airtable, GitLab, HubSpot, Notion, Supabase, and Vercel.
Adapting for a world of software factories (8 minute read)
Software factories shift engineers from coding locally toward supervising cloud agents and improving the systems that produce software. Warp expects engineers to own product quality while reducing human touches per PR through measurement, automation, and factory optimization.
OpenAI reportedly in talks to raise $30B round at $1.4T valuation (1 minute read)
OpenAI is in talks with investors to raise at least $30 billion in a funding round at a valuation of roughly $1.4 trillion. The company previously raised $122 billion in March in what was supposed to be the last private raise before an IPO. OpenAI has now ruled out a public debut in 2026 due to AI safety issues. The new fundraising will serve as a bridge round to the IPO.
22 Models, Same Correct Answer, 178x Cost Gap (Sponsor)
CData ran 22 models on live enterprise data. All returned the same correct answer at 178x difference in cost.
Read the benchmark.Your agents are paying 4x for the same answer (Sponsor)
When every agent hits your raw sources, one question can burn 45,000 tokens. Guru verifies knowledge once, then serves it to every agent over MCP at roughly 4x fewer tokens.
See how it worksAnnouncing Cohere's Embed 5 Models (2 minute read)
Embed 5 has major gains over Embed 4 on visually rich documents, financial filings, parsed PDFs, code, and multilingual retrieval.
How we engineer safer agents (14 minute read)
AI agents pose security risks as they may inadvertently cross security boundaries while pursuing legitimate goals.
GLM-5.3 and the spread of advanced cyber capabilities (14 minute read)
Anthropic's testing found that attackers can bypass GLM-5.3's safeguards between 64% and 100% of the time with simple techniques.
Devin is now up to 40% more cost-efficient (3 minute read)
Devin is now 30% to 40% cheaper in Fusion and Normal mode, 15% to 20% cheaper in Ultra, and up to 70% cheaper in Devin Review.
Announcing our partnership with OpenAI (4 minute read)
Baseten is partnering with OpenAI to serve open models natively via Codex and the Responses API.
Astra, Opus 5.5, and other Frontier Models Demonstrate Jagged Performance Across SoTA Agentic Tasks from Web Browsing to Robotics (32 minute read)
There is no clear performance leader and high variance in model-task fit on web tasks.
Get the most interesting AI stories and breakthroughs delivered in a free daily email.
Join 1,100,000 readers for
one daily email