The Alignment Problem: How To Tell If An LLM Is Trustworthy
New research attempts to put together a complete taxonomy for trustworthiness in LLMs. Before that on the Brief: The FEC is considering new election rules around deepfakes. Also on the Brief: self-driving cars approved in San Francisco; an author finds fake books under her name on Amazon; and Anthropic releases a new model. Today's Sponsor: Supermanage - AI for 1-on-1's - https://supermanage.ai/breakdown ABOUT THE AI BREAKDOWN The AI Breakdown helps you understand the most important news and dis