The high-performance architecture behind Gatling's load generation engine, and why it holds up when everything else starts dropping requests.
Everything on this page, as a PDF, for sharing with your team or your architecture review.
What engineering teams consistently run into when trying to run reliable, high-performance load tests.
Standing up a reliable load testing environment takes heavy engineering effort.
Most engines saturate compute, memory, or sockets long before realistic traffic.
Inefficient engines need too many machines for meaningful load.
Systems now span SaaS, Kubernetes, private VPCs, and on-prem, testing must adapt.
Gatling Enterprise Edition is powered by one of the most scalable, resource-efficient, and battle-tested load generation architectures in the industry.
While most legacy tools rely on heavy, thread-per-user architectures, Gatling's engine is fully asynchronous, event-driven, and optimized for massive concurrency, simulating millions of virtual users and sustaining extreme throughput with far fewer machines than traditional solutions.
Why Gatling's engine delivers real concurrency, predictable performance, and extreme throughput with fewer machines.
A single load generator can realistically sustain up to 60,000 concurrent virtual users or 300,000 requests per second, depending on protocol complexity and service architecture.
Monoliths, APIs, microservices, Kubernetes clusters, and globally distributed streaming platforms, all handled by the same engine.
What customers reach with Gatling Enterprise Edition
Gatling Enterprise Edition adapts to any network, security, or infrastructure context.
Hosted and operated by Gatling on AWS, pre-configured for optimal performance. Default instance: c6i.xlarge (4 vCPU).
Deploy and manage your own injectors in AWS, Azure, GCP, Kubernetes, OpenShift, or on-premises. Full control over sizing, scaling, and security.
Whichever model you choose, test orchestration, reporting, data aggregation, and scenario execution stay fully centralized within Gatling Enterprise Edition.
Gatling Enterprise Edition includes enterprise-grade support designed for mission-critical load testing.
Recommendations tailored to your systems.
Design guidance from testing experts.
Help finding system bottlenecks.
Guidance for demanding test setups.
Ongoing monitoring & troubleshooting.
JVM warmup, TLS optimization, scaling patterns.
Whether your services are public, firewalled, or fully isolated on-premises, there's a deployment model built for it.
| Endpoint infrastructure | Testing requirement | Recommended solution |
|---|---|---|
| Public services Internet-exposed |
No strict firewall rules or rate limits | Gatling-managed load generatorsFully managed on AWS |
| Public services with firewall / DDoS protection | Client must allow Gatling traffic through security layers | Gatling-managed with dedicated IPFixed IPs for whitelisting |
| Cloud-native or microservices environments | Need load close to the application to minimize latency | Private locationsCloud-managed |
| Secure or regulated environments | Sensitive data, secret keys, or regulated credential handling | Private locationsCloud-managed |
| Private services Behind firewalls, no direct internet access |
Application or API can't be reached from public networks | Private locationsCloud-managed |
| No access to public cloud providers On-prem, managed datacenters |
Fully isolated infrastructure | Private locationsDedicated machines |
All private-location options deploy across AWS, Azure, GCP, Kubernetes, or OpenShift, with full control over the entire environment.
Whether you're scaling APIs, migrating to the cloud, or handling flash-traffic spikes, Gatling helps you deliver fast, reliable performance.