AnthropicFigmaLinearVercelStripeShopifyPostHogDatadogRampSeatGeekSierraHexTurbopufferAnthropicFigmaLinearVercelStripeShopifyPostHogDatadogRampSeatGeekSierraHexTurbopuffer
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
NVDA$1,234.56+1.23%
The Claude Code Daily
Prompt injection benchmark comparing model risks

Anthropic Claims Prompt Injection Solved for Claude Two Months Ago

Bcherny says Anthropic solved prompt injection in practice for Claude models about two months ago, while noting it remains a significant risk for any model. The post benchmarks OpenAI's latest model as roughly on par with Gemini Flash and Opus 4.8 on this metric, and frames public naming-and-shaming of labs as a deliberate strategy to push the industry toward better alignment.

bcherny·Tue, Sep 8 12:35pm ET

Tip: Ask Claude to Post Slack Updates While You Sleep

Instead of using /loop or /goal, you can tell Claude your desired outcome and ask it to post frequent progress updates to a Slack channel overnight. Claude keeps working autonomously and you wake up to a progress log.

jarredsumner·Tue, Sep 8 10:28am ET

Custom Output Styles Land in Claude Code Desktop

The Claude Code desktop app now supports customizable output styles, letting users tune Claude toward more concise responses or specific formatting preferences.

lydiahallie·Tue, Sep 8 12:30pm ET

Managed Agents Founder Video Now on YouTube

The full video interview with founders of WisprFlow, Actively, and Pendo about building on Claude Managed Agents is now available on YouTube.

ClaudeDevs·Tue, Sep 8 4:03pm ET

Users Report Subagents Hard to Discover in Claude Code

A user mentioned that subagents took a long time to find in Claude Code. The follow-up question asks whether the friction was around creating custom subagents or using them in general, suggesting the team is paying attention to discoverability.

lydiahallie·Mon, Sep 7 10:13pm ET
More Stories

Bcherny Walks Back Sassy Tone on OpenAI Post

After posting that OpenAI's new model is roughly on par with Gemini Flash and Opus 4.8 on prompt injection risk, bcherny clarified the post was meant earnestly, not as a dig, and that the improvement is genuinely good news.

bcherny·Tue, Sep 8 12:56pm ET

Bcherny Lists Five Published Prompt Injection Defense Papers

In response to questions about sharing safety research, bcherny pointed to five published Anthropic works: Constitutional AI, persona vectors, sleeper agent probes, constitutional classifiers, and prompt injection defenses.

bcherny·Tue, Sep 8 1:05pm ET

Anthropic Publishes Safety Research to Help Other Labs

Anthropic intentionally publishes methods and techniques for training safer, more aligned models so other labs can benefit, reflecting genuine concern about AI safety going well for everyone.

bcherny·Tue, Sep 8 1:09pm ET

Overnight Claude Sessions Work Because Claude Tells You to Sleep

The trick of asking Claude to post Slack updates while you sleep seems to have gotten more reliable around the same time Claude started proactively telling users to go to bed, suggesting some shift in its overnight session behavior.

jarredsumner·Tue, Sep 8 10:33am ET

Get the daily digest in your inbox

Every weekday morning. Unsubscribe anytime.