Pratique du Shadowing: System Design: How to Build an API Rate Limiter - Apprendre l'anglais à l'oral avec la vidéo

Création de la leçon...
1
Let's design an API rate limiter for an application that takes thousands of requests per second.
2
User journey is simple.
3
User can make API calls to your server to do something.
4
Each user is allowed to make 100 API calls per minute.
5
If a user makes more than 100 calls per minute, then your server will not process those API calls.
6
Most engineers approach this problem with a counter.
7
For each user, you store user ID, count of API calls, and current minute.
8
Every time a request is made by user, you increase this count, and you will reject the requests once it passes 100 requests.
9
This implementation is good enough for internal API calls, but let's take two specific examples why this solution is not good.
10
Let's say user Alan is a burst user, and makes 100 API calls in first two seconds.
11
Then he will get rejected for the next 58 seconds.
12
Alan will be able to make calls when next minute starts.
13
And our second user, John, sends 100 API requests, but these requests are sent at the last second of the current minute.
14
The next second clock will reset to new minute, and John can send another 100 requests.
15
Basically, John sent 200 requests in 2 seconds, because according to Counter, these 100 requests belong to different minutes.
16
Basically, users will experience different behavior of APIs based on when they make the API call.
17
There is a fundamental issue with the approach we just discussed.
18
The Counter is watching the clock instead of watching the user.
19
Let's see a better approach.
20
So, we will give every user a bucket that holds 100 tokens.
21
Every API request from user takes one token out, and we will keep filling the bucket at a rate of 100 tokens per minute, which is approximately 1.67 tokens per second.
22
For simplification, let's assume we will fill two tokens per second.
23
If user doesn't make any API call, the bucket will remain at 100, and we don't need to fill.
24
If the bucket is empty, we will reject the request.
25
This token bucket approach is what Stripe, GitHub, and most API gateways actually use for tracking API calls.
26
Now let's take a look at Alan and John's requests again, and see how this bucket system will work.
27
Alan drains all 100 tokens in 2 seconds.
28
Then Alan has to slow down.
29
Bucket fills about 2 tokens per second.
30
That's why he can now make 2 calls per second.
31
John, who is making 100 API calls at the end of the minute, will also need to wait for Bucket to fill, because once he makes 100 requests in a second, Bucket will get empty.
32
So he can also make about 2 per second, as Bucket fills with time.
33
Next question is, how do we store these buckets?
34
We just need to store the token count, and the time of the last refill, so that you can add tokens per second.
35
But what if your app works at large scale, and you have 20 different servers accepting API calls.
36
Two requests for the same user can land on two servers in the same millisecond.
37
Therefore, you keep the bucket outside the servers, in one common shared storage.
38
This storage needs to be fast, and that's why we usually use Redis.
39
When any of these 20 servers processes any API call, we change the count at this common Redis.
40
To make this crystal clear in your mind, you need to remember that request limit is not just one number.
41
It's two numbers.
42
First, how much a user can burst in one go, and second, how fast they can go after the burst.

Vocabulaire et conseils d’expression pour cette leçon

Cette leçon d’expression orale de niveau C1 s’appuie sur la vidéo « System Design: How to Build an API Rate Limiter ». Les mots qui reviennent le plus souvent : Api, user, request, bucket, token. Cette vidéo contient 42 phrases et 594 mots à répéter en shadowing. La partie parlée dure 2:47. Le locuteur parle vite, environ 213 mots par minute : attendez-vous à des liaisons et à des sons réduits. 82 % des mots font partie des 3 000 mots les plus courants en anglais ; le reste mérite d’être vu avant de commencer.

Vocabulaire clé de cette vidéo

Les 13 mots les plus avancés de la vidéo, avec leur prononciation et leur sens :

MotPrononciationSens
bucket nom/ˈbʌkɪt/seau, chaudière
token nom/ˈtoʊkən/symbole
server nom/ˈsɝvɚ/serveur
burst verbe/bɜːst/éclater
clock nom/klɑk/indicateur de vitesse, compteur de vitesse
crystal nom/ˈkɹɪstəl/cristal
belong verbe/bɪˈlɒŋ/avoir sa place
fundamental adjectif/ˌfʌn.dəˈmɛn.təl/fondamental
implementation nom/ˌɪmplɪmənˈteɪʃən/mise en œuvre
drain nom/dɹiːn/drain, bonde
reset verbe/ɹiːˈsɛt/remettre à zéro, réinitialiser
millisecond nom/ˈmɪlɪˌsɛkənd/milliseconde
simplification nom/ˌsɪm.plɪ.fɪˈkeɪ.ʃən/simplification

Les verbes à particule que vous entendrez

MotSens
slow down verberalentir, décélérer

La grammaire de cette vidéo

Les structures que le locuteur utilise le plus, avec les mots exacts de la vidéo :

StructureDans la vidéo
Phrases conditionnelles if + proposition, will/would + verbe — une condition et son résultatIf the bucket is empty, we will reject
Voix passive be + participe passé — l’accent est mis sur ce qui arrive, pas sur qui le faitis allowed · is made · are sent

Prononciation à surveiller

  • Les sons « sh » et « zh »: implementation /ˌɪmplɪmənˈteɪʃən/, simplification /ˌsɪm.plɪ.fɪˈkeɪ.ʃən/
  • Mots longs — placez bien l’accent: fundamental /ˌfʌn.dəˈmɛn.təl/, implementation /ˌɪmplɪmənˈteɪʃən/, millisecond /ˈmɪlɪˌsɛkənd/, simplification /ˌsɪm.plɪ.fɪˈkeɪ.ʃən/

Les sons difficiles pour les francophones :

  • /r/ anglais — langue recourbée, sans frotter la gorge: crystal /ˈkɹɪstəl/, drain /dɹiːn/, reset /ɹiːˈsɛt/, refill /ˈɹiː.fɪl/

Comment s’entraîner avec cette vidéo

  1. Écoutez la vidéo en entier une fois sans parler et notez les mots que vous ne connaissez pas.
  2. Commencez à la vitesse 0,75×, répétez phrase par phrase, puis revenez à la vitesse normale quand cela devient facile.
  3. Enregistrez-vous et comparez avec l’original, en faisant attention à des mots comme bucket, token, server.

Qu'est-ce que la technique du Shadowing ?

Le Shadowing est une technique d'apprentissage des langues fondée sur la science, développée à l'origine pour la formation des interprètes professionnels. Le principe est simple mais puissant : vous écoutez de l'anglais natif et le répétez immédiatement à voix haute — comme une ombre suivant le locuteur avec un décalage de 1 à 2 secondes. Les recherches montrent une amélioration significative de la précision de la prononciation, de l'intonation, du rythme, des liaisons, de la compréhension orale et de la fluidité.

Technique du shadowing : lire le guide complet étape par étape →