Tickd.ai
API errors

How to Fix Claude API 429 Rate Limit Error

Updated 8/21/2026

An Anthropic API 429 error indicates that you have exceeded your rate limits. This occurs either because you sent too many requests per minute (RPM) or consumed too many tokens per minute (TPM). Anthropic enforces these limits dynamically based on your account's workspace billing tier.

To restore API functionality and prevent your application from dropping requests, follow these troubleshooting and optimization steps.

What Causes the Claude API 429 Error?

Anthropic categorizes rate limits into three distinct metrics: 1. Requests per minute (RPM): The total number of separate API calls made in a 60-second window. 2. Tokens per minute (TPM): The combined count of input and output tokens processed in a 60-second window. 3. Tokens per day (TPD): The total volume of tokens processed in a 24-hour window.

When your system crosses any of these thresholds, the API rejects subsequent calls with a 429 status code and a rate_limit_error payload.

---

Step 1: Read the Rate Limit Response Headers

Anthropic includes diagnostic headers in every single API response. Inspecting these headers in your application logs is the fastest way to determine whether you are hitting RPM or TPM limits, and exactly when the block will lift.

Examine your response objects for the following headers: * anthropic-ratelimit-requests-limit: Your maximum allowed requests per minute. * anthropic-ratelimit-requests-remaining: How many requests you have left in the current window. * anthropic-ratelimit-requests-reset: The time remaining (formatted as ISO 8601 duration, or an absolute UTC timestamp) until your RPM limit resets. * anthropic-ratelimit-tokens-limit: Your maximum allowed tokens per minute. * anthropic-ratelimit-tokens-remaining: The number of tokens you can still process before being blocked. * anthropic-ratelimit-tokens-reset: The time remaining until your TPM limit resets.

If anthropic-ratelimit-tokens-remaining is close to zero, you need to reduce payload sizes or throttle request speed.

---

Step 2: Implement Exponential Backoff with Jitter

Do not let your application spam the API immediately after receiving a 429 error. Instead, write a retry mechanism that uses exponential backoff and jitter (randomized delay). This spaces out retry requests, preventing "thundering herd" problems.

Here is a standard Python implementation using the official anthropic SDK and a backoff strategy:

`python import time import random from anthropic import Anthropic, RateLimitError

client = Anthropic(api_key="your_actual_api_key")

def call_claude_with_retry(prompt, model="claude-3-5-sonnet-20241022"): base_delay = 1.0 # Start with a 1-second delay max_delay = 60.0 max_retries = 5

for attempt in range(max_retries): try: message = client.messages.create( model=model, max_tokens=1024, messages=[{"role": "user", "content": prompt}] ) return message except RateLimitError as e: if attempt == max_retries - 1: raise e # Calculate exponential delay with jitter delay = min(base_delay * (2 ** attempt) + random.uniform(0, 1), max_delay) print(f"Rate limited. Retrying in {delay:.2f} seconds...") time.sleep(delay) `

---

Step 3: Optimize and Reduce Input Payload Sizes

If you frequently hit TPM limits, your prompts or system contexts may be excessively large. Use these practices to reduce token usage: 1. Use Prompt Caching: For large context sets (such as documents or system instructions), enable Anthropic's prompt caching feature. This dramatically reduces the real-time processing overhead and lowers your token consumption footprints. 2. Prune System Messages: Trim repetitive boilerplate text from your system prompts. Keep context instructions lean. 3. Set strict max_tokens limits: Explicitly define the maximum output size in your API calls to control return token volume.

---

Step 4: Verify Your Billing Tier and Request an Upgrade

If your code is optimized and you still hit limits, your application has outgrown your current billing tier. Anthropic scales your rate limits based on your lifetime deposit amount.

  1. Open your browser and log into the Anthropic Console.
  2. Navigate to Settings > Plans & Billing.
  3. Check your current Tier level (Tier 1 through Tier 5).
  4. If you are on Tier 1 (Evaluation), deposit at least $40 to automatically upgrade to Tier 2, which immediately raises your Sonnet limits from 20,000 TPM to 80,000 TPM.
  5. To request custom limits, navigate to Limits in the console and click Request limit increase next to the specific model family.

---

When to Escalate

If you have deposited enough funds to reach a higher tier but your rate limits on the Console's "Limits" tab have not updated within 2 hours, or if you are receiving 429 errors despite remaining well below the values indicated in your response headers, contact Anthropic Support. Log into the Anthropic Console, click the Help button in the bottom corner, and submit a ticket under "API Billing & Rate Limits" with copies of your API header logs.

Quick fixes

  • Claude is down or not loading
  • Claude Pro billing or payment problem
  • Can't sign in to Claude

While you're here

Tickd is more than troubleshooting — these three are free and take seconds.

Agent BuilderDesign your own AI agent and export it to ChatGPT, Claude, Gemini or Grok.Build one free