Механика алгоритма токенов в ведре
Изучите алгоритм токенов в ведре, его гибкость при допущении всплесков и распространённые варианты применения в современных системах.
«Механика алгоритма токенов в ведре» — бесплатный урок API Rate Limiting & Scalability Patterns на CoddyKit. Это урок 3 из 4. Ты можешь прочитать весь урок бесплатно ниже — а потом практиковать его прямо в браузере с встроенным редактором кода и ИИ-репетитором 24/7. Это часть пути обучения API Rate Limiting & Scalability Patterns, и твой прогресс синхронизируется между веб-версией и приложением CoddyKit. Курс API Rate Limiting & Scalability Patterns содержит 4 уроков всего.
Части этого урока еще не переведены и отображаются на английском.
Meet the Token Bucket
Welcome to the Token Bucket algorithm! After exploring fixed windows and leaky buckets, we'll now dive into a flexible approach that’s widely used in modern systems.
The Token Bucket is a rate-limiting algorithm that allows for bursts of traffic while still enforcing an average rate limit.
Tokens in a Virtual Bucket
Imagine a virtual 'bucket' that holds a certain number of 'tokens'. Each token represents permission for one API request.
- When a request arrives, it tries to take a token.
- If a token is available, the request proceeds, and the token is removed.
- If no tokens are available, the request is typically denied or queued.
Filling Up the Bucket
Tokens are continuously added to the bucket at a constant, predefined rate. This rate determines the average number of requests allowed over time.
For example, if tokens are added at 5 tokens per second, your API can sustain an average of 5 requests per second.
The Bucket's Maximum Size
Just like a real bucket, our virtual token bucket has a maximum capacity. This means it can only hold a certain number of tokens at any given time.
- If tokens are generated but the bucket is full, the new tokens are discarded.
- This capacity limits the maximum size of a 'burst' of requests that can be handled.
Requesting a Token
When an API client makes a request, the rate limiter checks the token bucket:
- If tokens are available: One token is consumed, and the request is allowed to proceed.
- If no tokens are available: The request is blocked, rejected (e.g., with HTTP 429 Too Many Requests), or deferred.
The Power of Bursts
The key advantage of the Token Bucket algorithm is its ability to allow bursts. If the bucket has accumulated many tokens (up to its capacity), a sudden rush of requests can be served immediately.
Once the accumulated tokens are used up, the rate limit reverts to the sustained token generation rate.
Token Bucket in Action
Try running this simplified Java example to see how tokens are refilled and consumed. Notice how initial requests can burst, but subsequent requests depend on refills.
public class Main {
// Simple TokenBucket class for demonstration
static class TokenBucket {
private int capacity;
private int tokens;
private int refillRate; // tokens per "tick"
public TokenBucket(int capacity, int refillRate) {
this.capacity = capacity;
this.tokens = capacity; // Start full
this.refillRate = refillRate;
}
public void refill() {
tokens = Math.min(capacity, tokens + refillRate);
System.out.println("Refill. Tokens: " + tokens);
}
public boolean tryConsume(int numTokens) {
if (tokens >= numTokens) {
tokens -= numTokens;
System.out.println("Consume " + numTokens + ". Left: " + tokens);
return true;
}
System.out.println("Fail to consume " + numTokens + ". Left: " + tokens);
return false;
}
public int getTokens() {
return tokens;
}
}
public static void main(String[] args) {
// Bucket: capacity 5, refills 1 token per tick
TokenBucket bucket = new TokenBucket(5, 1);
System.out.println("Start. Tokens: " + bucket.getTokens());
// 1. Initial burst
System.out.println("\n--- Request 1 (cost 3) ---");
bucket.tryConsume(3); // OK: 5 -> 2
// 2. Simulate time passing (refill)
System.out.println("\n--- Tick 1 ---");
bucket.refill(); // 2 -> 3
// 3. Another request
System.out.println("\n--- Request 2 (cost 2) ---");
bucket.tryConsume(2); // OK: 3 -> 1
// 4. Simulate time passing (refill)
System.out.println("\n--- Tick 2 ---");
bucket.refill(); // 1 -> 2
// 5. Try to consume more than available
System.out.println("\n--- Request 3 (cost 3) ---");
bucket.tryConsume(3); // FAIL: 2 tokens available
// 6. Simulate time passing (refill)
System.out.println("\n--- Tick 3 ---");
bucket.refill(); // 2 -> 3
// 7. Try again with enough tokens
System.out.println("\n--- Request 4 (cost 3) ---");
bucket.tryConsume(3); // OK: 3 -> 0
}
}Why Choose Token Bucket?
The Token Bucket algorithm offers several compelling advantages, especially when compared to simpler methods:
- Allows Bursts: It's perfect for APIs that expect occasional spikes in traffic.
- Simple to Implement: The core logic is straightforward to code.
- Smooth Average Rate: While allowing bursts, it still enforces a consistent average request rate over the long term.
- Flexible: You can tune both the refill rate and bucket capacity to suit different use cases.
Common Use Cases
Token Bucket is widely used in various scenarios where controlled burstiness is desirable:
- API Gateways: To protect backend services from sudden traffic surges.
- Network Traffic Shaping: To smooth out data transmission and prevent network congestion.
- Resource Management: Limiting access to shared resources in distributed systems.
- Client-Side Rate Limiting: Implementing rate limits within client SDKs to prevent excessive requests.
Token Bucket Check
Let's test your understanding of the Token Bucket algorithm.
Token Bucket Summary
Great job! In this lesson, you've learned about the Token Bucket algorithm, a powerful rate-limiting method.
- It uses a virtual bucket that accumulates tokens at a fixed rate.
- Requests consume tokens, and if no tokens are available, requests are denied.
- Its main strength is allowing controlled bursts of traffic, up to the bucket's capacity.
This flexibility makes it a popular choice for many real-world API and network applications.
Часто задаваемые вопросы
Урок «Механика алгоритма токенов в ведре» бесплатный?
Да — полный текст урока «Механика алгоритма токенов в ведре» бесплатно доступен здесь в веб-версии. Чтобы практиковать его интерактивно (встроенный редактор кода и ИИ-репетитор 24/7) и разблокировать остальной курс API Rate Limiting & Scalability Patterns, подпишись на CoddyKit PRO. Курс API Rate Limiting & Scalability Patterns содержит 4 уроков всего.
Чему я научусь в уроке «Механика алгоритма токенов в ведре»?
Изучите алгоритм токенов в ведре, его гибкость при допущении всплесков и распространённые варианты применения в современных системах. Ты практикуешь API Rate Limiting & Scalability Patterns с помощью реального кода, который запускаешь прямо в браузере, и ИИ-репетитор 24/7 отвечает на твои вопросы во время урока.
Нужен ли мне опыт, чтобы начать API Rate Limiting & Scalability Patterns?
Предыдущий опыт не требуется. API Rate Limiting & Scalability Patterns на CoddyKit структурирован для всех уровней — от новичков до продвинутых, поэтому ты можешь начать отсюда или с самого начала и учиться в своем темпе. Это урок 3 из 4.
Сколько времени занимает урок «Механика алгоритма токенов в ведре»?
Большинство уроков CoddyKit занимают около 5–10 минут. Каждый из них компактный и интерактивный, поэтому ты постоянно делаешь прогресс и продолжаешь с того же места в веб-версии и приложении.
Можно ли писать и запускать код в этом уроке API Rate Limiting & Scalability Patterns?
Да. Каждый урок API Rate Limiting & Scalability Patterns включает встроенный редактор кода, поэтому ты пишешь и запускаешь реальный код прямо в браузере и получаешь моментальную обратную связь от AI — локальная установка не требуется.
Все уроки этого курса
- Счётчик фиксированного окна
- Подробно об алгоритме дырявого ведра
- Механика алгоритма токенов в ведре
- Выбор подходящего алгоритма