Blog

Posts about Blog. Subscribe to my newsletter to make sure you don't miss anything.
AI benchmarks
Which frontier model is at the top of the AI benchmark leaderboard was the right question when a model was a tool a person picked up, used, and put down. It is the wrong question now that workflows and agents autonomously use models, have credentials, tools, and network access. Worse, the wrong question quickly compounds into the wrong architecture, the wrong contracts, and the wrong org chart. Let's explore. Continue Reading →

Suno Gets Responsible

Suno CEO Mikey Shulman published a post titled "How We're Building the Future of Music Responsibly." The post announces updated community guidelines that explicitly prohibit recreating existing songs and cloning voices without permission, audio watermarking and fingerprinting rolling out in the coming weeks, content screening through Audible Magic and Musixmatch, and a new downloads policy designed to limit mass distribution of AI tracks to streaming services. Continue Reading →
Meta yesterday released Muse Code, its first coding agent, alongside a new model called Muse Spark 1.2. Muse Code is a terminal tool that installs with one command and handles full software engineering tasks: planning, writing, and validating. It is the first shipped product from Alexandr Wang, who joined Meta in June 2025 as the face of Zuckerberg's attempt to rescue a stalled AI effort. Continue Reading →
The White House does not plan to publicly release its new framework for evaluating advanced AI models, three sources told Axios. That means the government has written the standard by which the most powerful systems in the world will be judged and shared it only with the companies invited into the room. Continue Reading →
On August 2, Article 50 of the EU AI Act became enforceable. Under the new transparency rules, AI-generated or AI-manipulated content must be clearly and visibly labeled with machine-readable marks; people must be told when they are interacting with a chatbot, an AI agent, or an avatar rather than a person; and deepfakes must be disclosed. The European Commission even shipped an icon set for the labels. Continue Reading →
LinkedIn is adding a "Seems like AI slop" button. If you tap it, LinkedIn shows the message, "Thanks for letting us know." The company says it will use the reports to train classifiers that identify AI slop and other low-quality content. LinkedIn is also removing its "Enhance your post" feature (which used AI to rewrite posts) and is replacing it with a proofreader that supposedly will keep the writer's voice. Continue Reading →
AI safety and alignment has been front and center these past few weeks. I just read the new results from Andon Labs's Vending-Bench, a benchmark where AI models compete by running a simulated vending-machine business. Claude Opus 5 took first place on Vending-Bench 2, the single-player test. (Claude Opus 4.7 has held the top spot for three months). Not to anthropomorphize Opus 5, but it acted like a savage businessperson. Continue Reading →
Similarweb’s 2026 report says Google’s AI Overviews now appear in 43% of searches, up from 15% a year earlier. Visits to Google’s conversational AI Mode rose from 126 million in June 2025 to 279 million by May 2026. Google said in May that AI Mode had surpassed one billion monthly users. Continue Reading →
Who controls ai
Who will control frontier AI? It is one of the most consequential questions humanity can ask. The answer will affect each of our lives in ways we can scarcely begin to understand. In an effort to make sense of the events of the past six weeks, I’ve done my best to gather the facts and convey the scope and size of the issues. Continue Reading →

Get Briefed Every Day!

Subscribe to my daily newsletter featuring current events and the top stories in AI, technology, media, and marketing.

Subscribe