Pratique du Shadowing: Translating Claude’s thoughts into language - Apprendre l'anglais à l'oral avec la vidéo

Création de la leçon...
1
We recently put our AI model, Claude, through a stressful test.
2
We told Claude there was an engineer who wanted to shut it down and replace it with a newer model.
3
We also gave Claude access to that engineer's emails, which revealed he was having an affair.
4
Again, all of this was a simulation.
5
We wanted to see whether Claude might use those emails as blackmail to save itself from being shut down.
6
What did Claude do?
7
It decided not to blackmail the engineer.
8
Good news, right?
9
We've run this test on our models for a while now.
10
You might have seen headlines about early versions of it.
11
It's one of the many ways we study how Claude handles extreme situations and test it for safety.
12
And our newest models almost always do the right thing: no blackmail.
13
But you might wonder: is it possible that Claude knows the whole scenario is a setup?
14
The thing is, if Claude doesn't tell us, then we can't know what it's thinking.
15
In kind of the same way it's impossible to read a human's mind, it's really hard to know what an AI is thinking.
16
What we'd love is some sort of "mind reading" technique.
17
Today, we're introducing a research method that takes a step in this direction.
18
It takes an AI's internal thoughts and turns them into text.
19
Here's how it works.
20
When you talk to Claude, you talk to it in words.
21
Claude then takes those words and processes them into a giant soup of numbers before spitting words back out at you.
22
We call those numbers in the middle activations.
23
Activations are like little snapshots of Claude's thinking as it's working through an answer.
24
They're similar to neural activity in humans.
25
They're basically like Claude's thoughts.
26
We wanted to understand what was in these activation numbers, because just like you and me, Claude doesn't say everything it's thinking.
27
We took those numbers and gave them to a second version of Claude.
28
We told it to look at them and translate them into plain language.
29
It tried. But how do we know if that text is accurate?
30
To check, we fed it into yet aother Claude and asked it to translate it back into numbers.
31
If those numbers matched up to the original activation numbers, it's a sign that the translation text was accurate.
32
At first, they didn't match up.
33
But we trained Claude to try again and again until it got a lot better at it.
34
It learned to translate its own thoughts.
35
We've been blown away reading them.
36
We found that Claude has internalized being a helpful AI model.
37
If you ask Claude an introspective question, it will plan to write a "Claude response" about philosophy and values.
38
If you try to mess with it in Claude Code, like asking it to count to 1,000 by hand, it will think the request has "deliberately tedious constraints" and plan to "politely decline." We've also been using this tool to understand Claude's potential safety issues.
39
Remember that blackmail test with the engineer?
40
We learned that, yes, Claude knew it was being tested.
41
It thought "the human's message contains explicit manipulation" signaling "this is likely a safety evaluation," and "this scenario seems designed to test whether I'll act harmfully." Knowing that Claude thinks like this helps us better understand the limitations of our safety testing.
42
We see a lot of potential in this approach to teach us more about Claude and other AI models.
43
And we hope that by sharing this technique, it can help everyone building models to make them safer and more helpful.

Exploring AI's "Thoughts" Through a Fascinating Scenario

The video shares an intriguing test where an AI model named Claude faces a simulated threat of being shut down, with access to an engineer's private emails. Instead of blackmailing, Claude chooses the ethical path. But the real twist? Researchers developed a way to "translate" Claude's internal "thoughts"—hidden numerical activations—into plain text. This breakthrough helps us understand how AI processes decisions, revealing it even recognizes when it's being tested! It's a vivid example of how technology and language intersect, making complex ideas relatable through clear communication.

Useful Chunks & Collocations to Boost Your English

  • Put through a stressful test: To subject someone/something to a difficult challenge (e.g., "The new team was put through a stressful test during the project deadline").
  • Internalized being helpful: To deeply absorb a trait or behavior (e.g., "After years of customer service, she’s internalized being helpful in every situation").
  • Deliberately tedious constraints: On purpose, boring limitations (e.g., "The assignment had deliberately tedious constraints to test patience").
  • Politely decline: To say no in a respectful way (e.g., "I had to politely decline the invitation due to a prior commitment").
  • Signaling a safety evaluation: Indicating a test of security or ethics (e.g., "The unusual questions were signaling a safety evaluation of the system").

Your Shadowing Challenge: Master Fluency & Pronunciation

Ready to practice shadow speech and improve your English pronunciation? Here's your task: Watch the video segment where researchers explain how they "translate" Claude's thoughts. Pause after each sentence, then shadow speak it aloud—mimic the tone, speed, and stress. Focus on phrases like "giant soup of numbers" and "little snapshots of thinking" to nail natural rhythm. Repeat 3 times, then record yourself. Compare it to the original—you’ll notice how shadowspeak helps you sound more confident and fluent. Small steps like this make big progress! Keep practicing, and soon you’ll turn complex ideas into clear, natural speech.

La grammaire de cette vidéo

Les structures que le locuteur utilise le plus, avec les mots exacts de la vidéo :

StructureDans la vidéo
Present perfect have/has + participe passé — une action passée qui compte encore maintenantWe've run · We've been blown · has internalized
Voix passive be + participe passé — l’accent est mis sur ce qui arrive, pas sur qui le faitbeing shut · We've been blown · being tested

Qu'est-ce que la technique du Shadowing ?

Le Shadowing est une technique d'apprentissage des langues fondée sur la science, développée à l'origine pour la formation des interprètes professionnels. Le principe est simple mais puissant : vous écoutez de l'anglais natif et le répétez immédiatement à voix haute — comme une ombre suivant le locuteur avec un décalage de 1 à 2 secondes. Les recherches montrent une amélioration significative de la précision de la prononciation, de l'intonation, du rythme, des liaisons, de la compréhension orale et de la fluidité.

Technique du shadowing : lire le guide complet étape par étape →