From San Francisco
1 |
The Editorial |
Morning, humans.
AI's biggest questions are being settled in private, company by company, while Washington leaves the federal rulebook open.
Scott Bessent rejected AI-lab liability protection on September 15, 2026. Andrew Ferguson challenged a separate antitrust waiver at a Washington event that day.
Mark Zuckerberg said on September 15, 2026, that each lab can set its own pace. Meta employees found failures after Muse's delayed September 8, 2026, release.
Profound raised $180 million on September 15, 2026, at a $1.8 billion valuation to polish brands in chatbot answers. Its agents even join Gmail and Slack, because apparently the bots now handle compliments.
Stay curious,
Marcus Schuler
2 |
The Big Story |
Scott Bessent rejected liability protection for AI labs at a House hearing on September 15, 2026.
"The best way to guarantee safety is that the creators are liable for what they build and generate," Bessent said. He proposed more U.S. open-source models.
FTC Chairman Andrew Ferguson challenged a separate antitrust waiver that day, saying the request set off "all of my alarm bells." OpenAI backed the FRONTIER Act's independent-verification provision the same afternoon. Chris Lehane said OpenAI, Anthropic and Google DeepMind had already worked together without a waiver.
Why This Matters:
- AI labs face federal resistance to shifting harm costs while coordinating development with rivals under an antitrust waiver.
- The House leaves Washington on September 17, 2026, limiting time for the stalled FRONTIER Act before the November elections.
3 |
Also Today |
Mark Zuckerberg said on September 15, 2026, that AI labs can pace their own work without a coordinated slowdown.
Meta postponed Muse from April 2026 to September 8, 2026. Employee posts after release described unauthorized uploads and guardrail failures, which tests Zuckerberg's case for company-led pacing. The posts are not a published evaluation and do not show how often Muse fails.
4 |
The Outside Read |
Amazon Science offers the clearest explanation I have read of why repeated benchmark tuning does not necessarily produce overfitting.
In the June 2026 paper's tests across eight datasets, 32-token prompts let reset agents reproduce the optimized strategy on most problems, and one language-model recipe held at 16 tokens without losing held-out performance. In the same June 2026 language-model test, cutting the prompt to eight tokens erased the gain, giving researchers a practical benchmark-memorization check.
5 |
The One Number |
6 |
Today's Headlines |
- Google released Gemini 3.8 Live and Extended Thinking on September 15, 2026; Extended Thinking scored 82.6 atop Artificial Analysis' Speech to Speech Quality Index.
- Factory raised $200 million at a $5 billion valuation on September 15, 2026, up from $1.5 billion in April 2026, taking total funding above $400 million.
- Agility Robotics unveiled Digit 5 on September 15, 2026, citing over $300 million in May 2026 orders subject to milestones; availability is due by end-2027.
- Altera confidentially filed for a U.S. IPO that could raise more than $2 billion; the filing date was not disclosed.
- Anthropic launched Claude for Financial Advisors on September 14, 2026, connecting it to portfolio and client-workflow systems from BlackRock and Charles Schwab.
- Waymo and partners GO and Nihon Kotsu target 2027 for driverless taxis in Tokyo, expanding to about 100 vehicles subject to approvals and validation.
Wed 9/16 |
Fed: the FOMC announces its rate decision and publishes its Summary of Economic Projections at 2 p.m. Eastern. |
Thu 9/17 |
Bank of England: the Monetary Policy Committee publishes its decision and meeting minutes at noon UK time. |
Fri 9/18 |
Economy: the Federal Reserve releases Industrial Production and Capacity Utilization at 9:15 a.m. Eastern. |
Fri 9/18 |
Apple: availability begins for the iPhone 18 Pro and iPhone 18 Pro Max. |
7 |
The 5-Minute Skill |
A vendor says its model leads a benchmark. Convert that claim into a small test tied to work your team does.
Your raw input:
The prompt:
Why this works: The prompt separates a published score from your operating conditions. Explicit pass rules stop a demo from becoming the standard after the fact.
What to use: A strong reasoning model. A standard general model works if every material fits in one prompt.
8 |
AI Profile |
Blacksmith sells faster computing for software teams that run builds and tests through GitHub Actions. Its position rests on a measurable second-order effect of coding agents: more generated code creates more validation work, and Blacksmith says weekly continuous-integration jobs on its platform grew 5% to 10% each week from January through August 2026.
Founders: Aditya "JP" Jayaprakash, formerly a search and advertising engineer at Faire; Aayush Shah, formerly a replication engineer at Cockroach Labs and an engineer at Superblocks; and Aditya Maru, formerly a disaster-recovery engineer at Cockroach Labs. The University of Waterloo graduates founded Blacksmith in 2024.
Product: Customers keep their GitHub Actions workflows and shift the computing work to Blacksmith's dedicated machines, caches and storage. More than 6,000 companies used the service as of August 12, 2026, including Supabase, Clerk, Ashby and Mercury. Blacksmith also sells Codesmith, an agent that diagnoses and repairs failed checks.
Financing: Blacksmith has raised $58.5 million through August 2026. Its latest round was a $45 million Series B led by Peak XV Partners, closed in March 2026 and announced August 12, 2026, at a $550 million valuation.
The risk: Blacksmith depends on GitHub Actions for distribution while Microsoft owns GitHub and can improve its hosted runners or fold comparable validation into Copilot. Blacksmith is also buying compute ahead of demand, so a slowdown in code generation could leave expensive capacity idle.
9 |
AI Image of the Day |
10 |
The Rausschmeisser* |
Profound raised $180 million at a $1.8 billion valuation on September 15, 2026, for software that steers how AI models describe brands.
Our take: Profound has turned corporate anxiety about chatbot manners into a company valued at $1.8 billion on September 15, 2026. Its software checks whether AI models mention a brand, then recommends edits meant to improve the answer. The industry calls this answer engine optimization, which gives search engine optimization a chatbot to optimize.
Profound's agents watch for slipping visibility, then enter Gmail and Slack to write in the company's voice. Profound also employs people it calls "forward-deployed marketing engineers." At the end of this chain, software edits corporate prose so other software will speak more kindly about the corporation. A human remains in the loop to pay.
FREE · ABOUT FIVE MINUTES
IMPLICATOR