Prompt Engineering

anthropic playground workbench claude api testing a jointer plane lying flat with a closed loop handle

Anthropic Rebrands Workbench as Playground for Claude API Testing

Anthropic retired the legacy Workbench on 17 August 2026 and shipped Playground the next day. The replacement is stateless: no saved prompts, no version history, no variables, no evals and no sharing, with nothing stored on Anthropic’s servers. In exchange it supports every Messages API parameter, ships templates for server-side features such as code execution and web search, shows the full SDK request alongside the response for every run, and exports to code. Three experimental prompt endpoints were retired alongside it, and the export window for Workbench data closed permanently on 1 September 2026.

Read more
ai gender bias womens language workplace prompts a gramophone with a flared horn on a square box body

AI Might Be Making Women Sound Bad at Work

Johns Hopkins researchers took 427 real workplace writing prompts, rewrote each into a women-associated and a men-associated version, and ran both through GPT-4, Llama, Mistral and Gemma. Every model returned shorter, plainer, less formal documents for the women-coded phrasing. Adding a male or female sign-off name changed nothing. Linear probes decode the linguistic register at 0.988 accuracy by layer 5 while name gender reaches only 0.717, and activation patching puts the causal weight in layers 0 to 7. This article covers the method, the per-model results, the two controls that rule out mirroring, and why the authors say users cannot fix it themselves.

Read more
CHAT