Bare-metal managed
- nodes
- 1 dedicated server
- scaling
- Vertical scaling
- database
- Local MariaDB
- cache
- Varnish, Redis, OPcache
- failover
- Restore from backup
- rpo
- 24h (daily backup)
- rto
- ~90 min
Tuned LAMP stack, multi-DC architecture, automatic scaling, HA on demand. Managed end-to-end by senior engineers on hardened Debian, in European datacenters.
01, ARCHITECTURE TIERS
Each tier maps a real e-commerce load profile. The HA module grafts on T3 and T4 to push RPO to zero on isolated node failure.
Indicative pricing on /en/prices. Every infrastructure is sized after an audit, no off-the-shelf product.
02, HIGH-AVAILABILITY MODULE
3 nodes minimum, spread across 2 datacenters minimum, with rack anti-affinity. Designed to keep the service running through isolated node loss without human intervention.
Synchronous MariaDB replication. Every write is committed across all nodes before validation.
Shared sessions, automatic failover. Cache and session continuity through node loss.
Automatic traffic rerouting on node failure. No human intervention for single-node loss.
3-node cluster minimum, spread across 2+ datacenters. Rack-aware placement to survive infrastructure-level events.
Real-time replication between nodes via lsyncd. Application files stay coherent across the cluster.
Synchronous DB replication, in-memory session replication. Designed for zero data loss on isolated node failure.
03, STACK
Every layer is provisioned via versioned Ansible playbooks, with project-level Unix isolation and a dedicated PHP-FPM pool per site.
| Layer | Components |
|---|---|
| Edge | Cloudflare WAF, DDoS mitigation, bot management, rate limiting |
| Load balancer | HAProxy or Nginx, health checks, sticky sessions when needed |
| Web | Nginx or Apache, PHP-FPM tuned per project, isolated pools, dedicated Unix user |
| Cache | Varnish full-page, Redis Sentinel for sessions and object cache, OPcache |
| Search | OpenSearch or Elasticsearch, dedicated nodes on T3 and T4 |
| Database | MariaDB, master/slave on T3, Galera multi-master on HA module |
| OS | Hardened Debian 12+, kernel-level isolation via LXC, project-level Unix isolation |
| Provisioning | Ansible playbooks versioned, internal inventory, lsyncd file replication |
| Monitoring | 100+ custom probes, Monit, Munin, log analysis, 24/7 alerting |
04, OBSERVABILITY AND SCALING
Monitoring catches anomalies before they hit the application. LXC isolation lets us add CPU and RAM without restart.
Console views, live probes
05, DISASTER RECOVERY
Targets validated during the commissioning phase, measured from incident detection. Detailed DR plan available to prospects on request.
| Scenario | Standard, RPO / RTO | Resilience, RPO / RTO | HA module, RPO / RTO |
|---|---|---|---|
| Application service crash (PHP-FPM, Redis, MariaDB, OpenSearch, Varnish) | 0 RPO / minutes | 0 RPO / seconds (LB) | 0 RPO / automatic |
| VM unreachable, hardware failure, filesystem corruption | 24h RPO / ~90 min | Seconds RPO / automatic (LB), 30-45 min (master/slave) | 0 RPO / automatic |
| Datacenter loss, prolonged provider outage | 24h RPO / 1-2h | Variable, automatic if multi-DC | 0 RPO / automatic (minority DC), 15-30 min (majority DC) |
| Database corruption | 24h RPO / 1-4h | Seconds RPO / 30-60 min (slave restore) | 0 RPO / automatic (Galera resync) |
| Configuration error (Cloudflare, system) | 0 RPO / under 1h | 0 RPO / under 1h | 0 RPO / under 1h |
9+ restore points by default: 7 daily, plus 2 bi-monthly. Extended retention (30, 60, 90 days) on demand.
Stored on 2+ separate servers, in 2 different European countries, with providers distinct from production.
Annual joint restore drill included on T3 and HA tiers. Run with you, on your infrastructure.
CONSOLE.FAST-MAGE.COM
Every backup is tracked: timestamp, scope (files, database, full), size, storage target. Restore opens as a ticket from this view, no tier 1 filter.
History retained for the full contract duration. Client-side audit available at any time, no prior request needed.
06, CONTRACTUAL SLA
Full SLA defined in the Cloud Hosting Particular Conditions, article 5. Penalties on miss, article 5.2.2.
| Engagement | SLA |
|---|---|
| Monthly uptime | 99.9 % |
| Network restore | 2h |
| Hardware restore | 2h |
| System restore | 3h |
| Functional anomaly restore | 12h |
| High-priority intervention, business hours | 30 min |
| High-priority intervention, after hours | 6h |
On real incidents, in-hours pickup is typically under 10 min, well below the 30 min contractual SLA. After hours, automated SMS escalation reaches an on-call engineer with end-to-end pickup under 20 to 30 min, then resolution under the 6h contractual window.
Every high-priority incident gets a written post-mortem on request: timeline, root cause, corrective actions.
07, SECURITY AND SOVEREIGNTY
Production stays in the EU, on infrastructure providers we audit for performance and isolation. PCI DSS available on dedicated offers.
Production runs on a mix of:
Provider selection per project: latency, location, capacity, regulatory constraints. GCP and AWS available when client requirements demand it.
08, SUPPORT MODEL
Direct contact with the engineers who run your infrastructure. No qualification script, no bot, no disposable first-level filter.
Stack maintenance, kernel and OS patching, MariaDB tuning, Varnish VCL, Redis, OpenSearch indexing, PHP-FPM pool sizing. Provisioning via Ansible.
Performance bottleneck triage, plugin conflict diagnosis, CMS hack containment, indexer freeze, queue stalls. Direct collaboration with your dev team or agency.
Provider escalation, hardware replacement coordination, DNS cutover, capacity planning, backup restore drills. PowerSupport option covers client-owned infrastructure.
60 % of new clients come via referrals from agencies and IT directors. The DR plan (PRA) is shared with prospects under NDA, contact us for the full version.
06, FREQUENTLY ASKED
Five key points.
Answers reviewed by our engineers. If your question is missing, just write to us, quick reply, no sales script.
Yes. 99.9 % monthly uptime is the contractual SLA written in our Cloud Hosting Particular Conditions, article 5. Penalties apply if the threshold is missed. Higher targets are achievable on the HA module by design (synchronous Galera, automatic failover), validated during the one-month commissioning phase.
Minimum 9 restore points available at any time: 7 daily rolling backups, plus 2 bi-monthly snapshots (1st and 16th of each month, kept until the next cycle). Backups live on at least 2 separate servers, in 2 different European countries, with infrastructure providers distinct from the production server. Extended retention (30, 60, 90 days) and Tier-3 backup to your own storage are available.
Designed RPO of zero on isolated single-node failure: synchronous Galera writes are committed across the cluster, Redis Sentinel handles session failover, the load balancer reroutes traffic automatically. On loss of a majority DC (2 nodes out of 3), the remaining node loses quorum and refuses writes by default (split-brain protection); a manual bootstrap reinstates service in degraded mode within 15 to 30 min in business hours. Targets are validated during the one-month commissioning acceptance test before production.
European datacenters only: OVH, Scaleway, Hetzner, UpCloud. No US providers in the production path. Each project runs under a dedicated Unix user with isolated PHP-FPM pool. Configuration and credentials are centralized in our internal inventory; access is restricted to authenticated Fast-Mage engineers and logged. GDPR-compliant. PCI DSS available on dedicated offers.
Yes. The managed-services-only model is supported. RTO targets then apply from the moment you provision an operational server on your account. The PowerSupport option lets Fast-Mage interact directly with your provider's console and ticket system, otherwise hardware escalation stays on your side. Audit, takeover playbook, and a one-month observation period are included.
Automated monitoring (100+ custom probes) detects incidents around the clock. In business hours, the technical team on duty acts immediately. After hours, the alerting system escalates by SMS to the first-level on-call engineer, with automatic escalation if no acknowledgement. End-to-end pickup time after hours can reach 20 to 30 minutes. Effective intervention then starts under the contractual 6h after-hours SLA, in practice much faster on real incidents.
Question not answered?
Contact us →CONTACT, FREE AUDIT
Free audit, reply within 24 business hours.
Tell us about your stack and situation, a Fast-Mage engineer will call you back. No salesperson, no script. The audit stays yours, even if you don't migrate.