Anthropic
entities · 74 notes linked
Related: Openai · Agentic Coding · Large Language Models · AI Safety · AI Agents · Claude Code · Zvi Mowshowitz · Claude
Notes
- 10 things I learned from burning myself out with AI coding agents — Practical limits of AI coding agents after 50 real projects
- 255 — Paul Christiano's defense of RLHF research positive net impact
- 27 AI models were ranked by the public and ChatGPT came 8th — these are the models that beat it — Prolific's Humaine leaderboard ranks Gemini first in public user experience study
- A Complete Breakdown of the Claude Mythos 1 Leak and Features — Leaked Claude Mythos 1 capabilities including 69% Exploit Bench score
- A Deep Dive Into MCP and the Future of AI Tooling — MCP as standard interface for AI tool use
- AI #171: False Flag — Weekly roundup; OpenAI-linked PAC false-flag scandal
- AI #172: The First Fable — Weekly roundup excluding the Fable model itself
- AI #173: AI Pauses — Weekly AI roundup during the Fable takedown
- AI #174: You're It — Weekly AI roundup, Claude Tag, medical scanners, agent security
- AI #175: The Fable Continues — Weekly AI roundup covering Fable's return aftermath
- AI 2027 — HN debate on AI 2027 scenario: AGI timelines, LLM limits, alignment risks
- AI 50 2025: AI Agents Move Beyond Chat — Forbes AI 50 marks shift from chatbots to full workflow automation
- American Government Takes Down Claude Fable — Commerce export controls abruptly shut down Fable
- Boris Cherny (@bcherny) on X — Claude Code creator Boris Cherny shares his minimal vanilla setup
- Building a C compiler with a team of parallel Claudes — 16 parallel Claude agents autonomously build Rust C compiler
- Claude Code Ralph Plugin Breaks LLM Performance (And a Simple Bash Loop Wins) — Ralph methodology critique — stateless bash loops outperform official Claude Code Ralph plugin
- Claude Code Tasks Are Here (New Update Turns Claude Code ToDos to Tasks) — Claude Code Tasks feature upgrades ephemeral Todos to persistent cross-session tasks
- Claude Fable 5 and Mythos 5: Capabilities — Fable 5 capability review, benchmarks, classifiers, reception
- Claude Fable 5 and Mythos 5: The System Card — Reading the Fable/Mythos 319-page system card
- Claude Skills are awesome, maybe a bigger deal than MCP — Claude Skills as lightweight Markdown-based agent capability extension beating MCP
- Claude Sonnet 5 Is Not Frontier But Has Its Uses — Sonnet 5 system card review, cheaper faster non-frontier model
- Claude for Life Sciences — Anthropic launches Claude specialization for life sciences research workflows
- Claude's Chrome plugin is now available to all paid users - Engadget — Anthropic's Claude Chrome plugin expands beyond Max plan subscribers
- Clawdbot vs Claude Code: One for Coding, One for Everything Else (Don't Get Confused) — Comparing Claude Code terminal coding tool with Clawdbot general assistant
- Cost calculations for LLM providers — Comparative USD-per-million-token pricing table with LMSYS ELO scores
- Does Prompt Caching Make RAG Obsolete? — Prompt caching economics vs RAG for LLM context management
- Does Vibe Coding Really Work? We Built a Game With Claude—Here's How It Turned Out - Decrypt — Hands-on vibe coding experiment building a game with Claude 3.7 Sonnet
- Employment for computer programmers in the U.S. has plummeted to its lowest level since 1980—years before the internet existed | Fortune — AI correlates with historic drop in US programming jobs
- Fable #6: The Return of the King — Anthropic Fable 5 restored after government export-control blip
- Fable and Mythos: Model Welfare — Model welfare assessment of Fable 5 / Mythos 5
- From Google To Nvidia, Tech Giants Have Hired Hackers To Break AI Models — AI red teams at major tech companies probing model vulnerabilities
- GPT-5.6: The System Card — Review of OpenAI GPT-5.6 Sol/Terra/Luna system card
- Google DeepMind boss hits back at Meta AI chief over 'fearmongering' claim — Hassabis vs LeCun debate on AI safety, regulation, and open-source control
- Here are the top 10 generative-AI startups founded by ex-Googlers that are taking on ChatGPT — 2023 roundup of ex-Google generative-AI startups
- How Claude Helps Me Manage My Calendar (but ChatGPT Stumbles!) — Claude processes ICS calendar files reliably where ChatGPT fails context limits
- How I Use Every Claude Code Feature — Practitioner's guide to Claude Code features and workflows
- How We Use Claude Code Skills to Run 1,000+ ML Experiments a Day — Claude Code skills registry for team ML knowledge
- How to Create Powerful Loops in Claude Code | Towards Data Science — Using /goal command to create self-verifying autonomous coding agent loops
- I Tested (New) Claude Code /Insights (It Roasted My Coding Habits) — Claude Code /insights command analyzes 30-day coding history into personalized report
- I Tried New Claude Code Ollama Workflow ( It's Wild & Free) — Claude Code integrates with Ollama for free local model workflows
- I created over a dozen personal apps using AI in 60 days, here's what I learned — Non-programmer builds 15+ apps using Claude Sonnet and Bolt.new
- I used Claude to vibe-code my wildly overcomplicated smart home — Non-coder uses Claude Code to configure Home Assistant
- I've been using Claude Code for a couple of days — Hacker News debate on Claude Code coding experience
- IDEcline: How the world's most powerful coding tools became second-class citizens overnight — Three-wave shift from IDE-centric to agent-control-plane software development
- Introducing Claude for education — Anthropic launches Claude specialized version for higher education
- MCP + Google Sheets: A Beginner's Guide to MCP Servers — Beginner tutorial connecting MCP servers to Google Sheets
- MCP doesn't move data. It moves trust — MCP as AI governance and intent-control layer over APIs
- MCP: An (Accidentally) Universal Plugin System — MCP as accidentally universal interoperability layer beyond AI context
- MCP: The Missing Link Between AI Agents and APIs — MCP standardizes how AI agents access external APIs
- Mapping the Mind of a Large Language Model — Anthropic extracts millions of interpretable features from Claude 3 Sonnet
- Orchestrate teams of Claude Code sessions - Claude Code Docs — Claude Code agent teams for parallel coordinated work
- Perspectives on the Social Impacts of Reinforcement Learning with Human Feedback — Social and ethical impacts of RLHF across seven societal dimensions
- Prompt Engineering Urges 'Hermeneutic Prompting' As A Powerful Technique Unlocking The True Value Of Generative AI — Hermeneutic circle prompting technique for richer LLM responses
- Replit CEO on AI breakthroughs: 'We don't care about professional coders anymore' — Replit Agent targets non-coders after Claude 3.5 Sonnet SWE-bench breakthrough
- Senior Developer Skills in the AI Age: Leveraging Experience for Better Results — Senior-developer practices for guiding AI coding agents
- Silicon Valley bets big on 'environments' to train AI agents | TechCrunch — RL environments emerge as critical training infrastructure for capable AI agents
- Super Nested Claude Code Offers Next Level Vibe Coding — Nested Claude Code instances coordinated via tmux for parallel workflow automation
- The Batch | DeepLearning.AI | AI News & Insights — DeepLearning.AI weekly AI news and insights newsletter homepage
- The Capacity for Moral Self-Correction in Large Language Models — RLHF-trained LLMs can self-correct harmful outputs when instructed
- The Once And Future Fable #2 — Reconstructing the 24 hours of the Fable takedown
- The Once And Future Fable #3: Fix This Code — Debunking the Fable "jailbreak," export-control fallout
- The Once And Future Fable #4 — Fable takedown aftermath, cyber-defense, restoration odds
- The Once And Future Fable #5 — AI-policy roundup on ad hoc licensing regime and open weights
- The Pentagon says AI is speeding up its 'kill chain' | TechCrunch — AI accelerating Pentagon kill chain planning amid usage policy tensions
- The author of SB 1047 introduces a new AI bill in California | TechCrunch — California SB 53 creates AI whistleblower protections and public compute cluster
- The challenges of reinforcement learning from human feedback (RLHF) - TechTalks — RLHF limitations across feedback, reward modeling, and policy
- The second wave of AI coding is here — Next-gen AI coding agents targeting functional correctness via process data
- Three Labs With a Plan and A Memorandum — Trump AI memorandum plus OpenAI's benefit-everyone plan
- WSJ Article Claiming China Has Matched Anthropic Is Obvious Nonsense — Debunking WSJ headline that China matched Mythos on cyber
- We Got Claude to Fine-Tune an Open Source LLM — Hugging Face skill enables Claude Code to submit and manage LLM training jobs
- What AI Models for War Actually Look Like — Military-specialized AI startup building models for mission planning and decision dominance
- White House Will Ad Hoc Decide Who Can Individually Access GPT-5.6 — Critique of ad hoc White House frontier-model access regime
- Xiaomi's MiMo Code claims it beats Claude Code past 200 steps — Long-horizon coding agent endurance gap and competing harness approaches
- claude code's DX is too good. and that's a problem. — Claude Code's expanding feature surface creates developer experience tension