During the webinar, ART-Fintech experts will talk about modern approaches to building a highly accessible container infrastructure for financial organizations that ensures the continuous operation of critical services even in the event of hardware, software, or entire sites failures.
In the webinar program:
1. Fault-tolerant container cluster
- fault tolerance levels of container infrastructure;
- clustering of management and work nodes;
- service replication and automatic application recovery;
- eliminating single points of failure.
2. Load distribution and balancing
- balancing external traffic;
- Ingress controllers and service balancing;
- horizontal application scaling;
- peak load handling;
- automatic service availability monitoring.
3. Data and storage fault tolerance
- ensuring high availability of PostgreSQL, MongoDB and other databases;
- OpenSearch/ELK fault-tolerant search clusters;
- distributed file and block storage (Piraeus/LINSTOR, Ceph);
- data protection, replication, and disaster recovery;
- differences between replication, backup, and archiving.
4. Distributed Platforms and Disaster Recovery
- building infrastructure in multiple data centers;
- Active-Passive, Active-Active, and "warm" reserve schemes;
- one distributed cluster or several independent clusters;
- synchronization of applications, databases, and storage;
- RTO and RPO indicators;
- scenarios for switching to a backup site and returning after an accident.
A practical example
At the end of the webinar, an example of building a fault-tolerant container platform, including application services, databases, a search cluster and distributed storage, will be considered.