한국어English日本語简体中文繁體中文DeutschไทยTiếng ViệtРусскийPortuguês (Brasil)EspañolBahasa Indonesia

Game Lag White Paper › L13 Server architecture and operations

External service dependency External dependencies (auth, billing, platform)

Cause ID in-external · Primary owner External (External) · Also Game team (Server development)

Open the interactive card with figures and simulations →

When an external service such as platform login, payments, or identity verification is slow or down, players get stuck at that step.

Why An external authentication or payment service is down or slow → Effect That step waits for a response → On screen Can’t log in, payments fail. Players already in the game are fine

Symptoms
Can’t connect / infinite loading, Dropped action / rollback
Factors
Stall
Who’s affected
Whole server, One feature only
When
Right after login or maintenance, During specific actions
Owner
Primary owner External (External) · Also Game team (Server development)
Game team action items
Put timeouts and a friendly message on external calls, cache authentication results, set up a retry and compensation process for payments.
External action items
Ask the authentication, payment, or platform provider to confirm the outage and restore service, tell players the problem is an external service outage.
On the graph
Step change · External call response time and error rate, successful logins
Where to look
Response time, error rate, and timeout count per external call (platform login, payments, identity verification), plus the provider’s status page
Confirmed if
From the time login and payment failures pile up, errors and timeouts for one specific external call step up and stay there, and the provider’s status page shows an outage at the same time
Ruled out if
External calls are healthy but logins are blocked: points to the login server itself (thread pool exhaustion, DB) or the OS connection queue (backlog)
Check with
Game server or client logs and metrics
Real incidents
Fastly 2021: Worldwide Fastly CDN errors
AWS 2021: AWS us-east-1 internal network congestion
AWS 2025: AWS us-east-1 DynamoDB DNS outage and long recovery

Sources

  1. Timeouts, retries, and backoff with jitter AWS
    Amazon Builders’ Library. Waiting for a response holds resources such as threads and connections, so set timeouts, and retry APIs with side effects only when they’re idempotent
  2. Circuit Breaker Pattern Microsoft Azure
    Calls likely to fail are rejected immediately without waiting for the timeout, which protects response time
  3. REL05-BP01 Implement graceful degradation to transform applicable hard dependencies into soft dependencies AWS
    AWS Well-Architected. Keep core functions running when a dependency fails, even on slightly stale data (the basis for caching authentication results)

See also

Same layer: L13 Server architecture and operations

Same symptom (Can’t connect / infinite loading), other layers

View the interactive card with figures and simulations