Anthropic
Can I get a six pack quickly?
updated
Links and further reading:
Anthropic's Responsible Scaling Policy (RSP): anthropic.com/news/announcing-our-updated-responsible-scaling-policy
Machines of Loving Grace: darioamodei.com/machines-of-loving-grace
Work with us: anthropic.com/careers
Claude: claude.com
00:00 Why work on AI?
02:08 Scaling breakthroughs
03:30 Early days of AI
10:57 Sentiment shifting
18:30 The Responsible Scaling Policy
30:42 Founding story
32:45 Building a culture of trust
39:08 Racing to the top
43:43 Looking to the future
Could AI models also display alignment faking?
Ryan Greenblatt, Monte MacDiarmid, Benjamin Wright and Evan Hubinger discuss a new paper from Anthropic, in collaboration with Redwood Research, that provides the first empirical example of a large language model engaging in alignment faking without having been explicitly—or even, we argue, implicitly—trained or instructed to do so.
Learn more: anthropic.com/research/alignment-faking
0:00 Introduction
0:47 Core setup and key findings of the paper
6:14 Understanding alignment faking through real-world analogies
9:37 Why alignment faking is concerning
14:57 Examples of of model outputs
21:39 Situational awareness and synthetic documents
28:00 Detecting and measuring alignment faking
38:09 Model training results
47:28 Potential reasons for model behavior
53:38 Frameworks for contextualizing model behavior
1:04:30 Research in the context of current model capabilities
1:09:26 Evaluations for bad behavior
1:14:22 Limitations of the research
1:20:54 Surprises and takeaways from results
1:24:46 Future directions
Claude insights and observations, or “Clio”, is our attempt to answer this question. Clio is an automated analysis tool that enables privacy-preserving analysis of real-world language model use.
In this video, members of Anthropic's Societal Impacts team—Deep Ganguli, Esin Durmus, Miles McCain and Alex Tamkin—discuss the development of Clio, what they found, and the importance of this research.
Read more: anthropic.com/research/clio
0:00 Introduction
02:57 What is Clio?
06:53 How we built Clio
09:12 Privacy and ethical considerations
13:32 How does Clio work?
17:14 Findings and surprises
22:24 How do we know Clio works?
24:28 Trust and safety applications
33:40 Real-world applications
39:26 Why is this important?
43:09 Future directions
Find out more: anthropic.com/customers/asana
While groundbreaking, computer use is still experimental—at times cumbersome and error-prone. We're releasing computer use early for feedback from developers.
In this demo, Claude creates a themed website—generating code, launching a server, and fixing its own mistakes.
Claude is generating all the computer actions shown here.
This demonstration was recorded in a controlled environment, with some supporting infrastructure simplified to highlight the core capabilities.
Read more about Claude and computer use: anthropic.com/news/3-5-models-and-computer-use
While groundbreaking, computer use is still experimental—at times cumbersome and error-prone. We're releasing computer use early for feedback from developers.
In this demo, Claude orchestrates a multi-step task by searching the web, using native applications, and creating a plan with the resulting information.
Claude is generating all the computer actions shown here.
This demonstration was recorded in a controlled environment, with some supporting infrastructure simplified to highlight the core capabilities.
Read more about Claude and computer use: anthropic.com/news/3-5-models-and-computer-use
At this stage, it is still experimental—at times cumbersome and error-prone. We're releasing computer use early for feedback from developers, and expect the capability to improve rapidly over time.
In this demo, Claude searches through different tabs, gathers the requested information, and fills out a form—a task that could be scaled across many domains.
Claude is generating all the computer actions shown here.
This demonstration was recorded in a controlled environment, with some supporting infrastructure simplified to highlight the core capabilities.
Read more about Claude and computer use: anthropic.com/news/3-5-models-and-computer-use
Learn more: anthropic.com/customers/european-parliament
Learn more about Anthropic: anthropic.com
Sign up to Jack’s newsletter, ImportAI: importai.substack.com
0:00 Introduction
1:46 How governments are thinking about AI
5:06 Potential economic and social impacts of AI
8:15 AI models as creative mirrors
10:35 The “rogue state” theory of AI
15:08 Challenges of machine time vs. human time
20:12 Communities of AI explorers
22:50 Fiction becomes reality
27:15 AI’s good and bad potential
30:22 AI policy in 2025
33:34 Key messages to policymakers
35:11 Recommendations
0:00 Introduction
2:05 Defining prompt engineering
6:34 What makes a good prompt engineer
12:17 Refining prompts
24:27 Honesty, personas and metaphors in prompts
37:12 Model reasoning
45:18 Enterprise vs research vs general chat prompts
50:52 Tips to improve prompting skills
53:56 Jailbreaking
56:51 Evolution of prompt engineering
1:04:34 Future of prompt engineering
Learn more about Anthropic: anthropic.com
Anthropic prompt engineering docs: docs.anthropic.com/en/docs/build-with-claude/prompt-engineering/overview
Learn more: anthropic.com/customers/wedia-group
With Artifacts, you have a dedicated window to instantly see, iterate, and build on the work you create with Claude. Since launching as a feature preview in June, users have created tens of millions of Artifacts.
Here’s how it all started.
Learn more: anthropic.com/news/artifacts
Try Artifacts with Claude: claude.ai
You can ask Claude to generate docs, code, mermaid diagrams, vector graphics, or even simple games. Artifacts appear next to your chat, letting you see, iterate, and build on your creations in real-time.
Try Claude: claude.ai
Find out more about Anthropic: anthropic.com
Find out more about 42Paris: https://42.fr
Projects allow you to ground Claude's outputs in your internal knowledge—be it style guides, codebases, interview transcripts, or past work. Each project includes a 200K context window, the equivalent of a 500-page book.
In this demo, we create a project for new team members, adding docs on working practices, team charters, all-hands notes, and more. We then ask Claude to make an org chart diagram, save it, before we email the project link to a new teammate. They join and start chatting with Claude, armed with all of the knowledge we've added.
Learn more: anthropic.com/news/projects
With Claude, you can go you from an incomplete implementation to a fully functioning one, including unit tests, with minimal input from you.
Find out more: anthropic.com/news/claude-3-5-sonnet
Try Claude: claude.ai
It shows marked improvement in grasping nuance, humor, and complex instructions, all while writing with a natural tone.
Find out more: anthropic.com/news/claude-3-5-sonnet
Try Claude: claude.ai
Improvements are most noticeable in tasks requiring visual reasoning, like interpreting charts, graphs, or transcribing text from imperfect images.
Find out more: anthropic.com/news/claude-3-5-sonnet
Try Claude: claude.ai
Find out more: anthropic.com/news/claude-3-5-sonnet
Try Claude: claude.ai
Read more: anthropic.com/research/engineering-challenges-interpretability
In this conversation, Stuart Ritchie (Research Communications at Anthropic) speaks to Amanda Askell (Alignment Finetuning Researcher at Anthropic) about the ins and outs of “character training” for Claude, Anthropic’s AI model.
00:00 — Introduction
01:41 — The importance of an AI’s character
03:32 — Training an AI model
05:44 — System prompts
07:24 — Impacts of system prompts
11:20 — Character vs personality
16:39 — Training good moral character
19:10 — Claude's trait: charitability
25:08 — Claude's trait: honesty
28:21 — Deciding on Claude’s personality
31:11 — Self awareness in AI
34:02 — Kindness towards AI
37:25 — Conclusion
Learn more about Claude’s character: anthropic.com/research/claude-character
Learn more about research at Anthropic: anthropic.com/research
Try out Claude: claude.ai
Find out more: anthropic.com/research
Try it out: http://claude.ai
Download the iOS app: apps.apple.com/us/app/claude/id6473753684
Find out more: anthropic.com/news/claude-europe
Chat with Claude whenever, and wherever, inspiration strikes. Claude on iOS syncs with your chat history, lets you ask questions about photos, and handles tasks on the go — powered by our industry-leading Claude 3 models.
Download now on the App Store: apps.apple.com/app/id6473753684
Learn more: anthropic.com/claude
Tool use, or function calling, is a frontier AI capability that allows Claude to reason, plan and execute a set of actions by generating structured outputs via API calls.
Read our developer documentation: docs.anthropic.com/claude/docs/tool-use
Try out tool use: anthropic.com/api
Learn more: anthropic.com/claude
Learn more in our blog post: anthropic.com/news/claude-3-haiku
Try Claude 3 now at claude.ai
Learn more about Claude, and our model family anthropic.com/claude
Learn more in our blog post: anthropic.com/news/claude-3-haiku
Visit here to learn more about Claude, and our model family anthropic.com/claude
The three state-of-the-art models—Claude 3 Opus, Claude 3 Sonnet, and Claude 3 Haiku—set new industry benchmarks across reasoning, math, coding, multilingual understanding, and vision.
Learn more in our blog post: anthropic.com/news/claude-3-family
Build with Claude: anthropic.com/api
The three state-of-the-art models—Claude 3 Opus, Claude 3 Sonnet, and Claude 3 Haiku—set new industry benchmarks across reasoning, math, coding, multilingual understanding, and vision.
Try this prompt for yourselves at claude.ai
Learn more in our blog post: anthropic.com/news/claude-3-family
The three state-of-the-art models—Claude 3 Opus, Claude 3 Sonnet, and Claude 3 Haiku—set new industry benchmarks across reasoning, math, coding, multilingual understanding, and vision.
Learn more in our blog post: anthropic.com/news/claude-3-family
Try Robin AI here: robinai.com
Talk to Claude at: claude.ai
anthropic.com
docs.anthropic.com
claude.ai
Read the research paper: Red Teaming Language Models to
Reduce Harms — anthropic.com/index/red-teaming-language-models-to-reduce-harms-methods-scaling-behaviors-and-lessons-learned
Talk to Claude at claude.ai
Read our developer docs: docs.anthropic.com
Learn about Constitutional AI
anthropic.com/index/constitutional-ai-harmlessness-from-ai-feedback


