Overview
TRaaS is a technical risk control platform. It is built on the proven methodologies and internal tools of the Ant Group Site Reliability Engineering (SRE) team. The platform resolves operations and maintenance (O&M) problems that arise during cloud migration and distributed transformation. These problems include observability, incident response, disaster recovery, chaos engineering, fund security, and stress testing.
High Availability Management Platform
The High Availability Management Platform, also known as High Availability Service (HAS), is a high availability (HA) management platform that focuses on disaster recovery. It provides end-to-end capabilities for managing disaster recovery plans across the entire stack, from customer services to middleware, Platform as a Service (PaaS), and Infrastructure as a Service (IaaS). Key features include disaster recovery switchovers, recovery, planning, and drills. The platform also offers monitoring of data centers and disaster recovery status, a disaster recovery dashboard, environment inspections, and risk response.
HAS provides disaster recovery service views, plan orchestration, and switchover and recovery capabilities. It supports one-click disaster recovery switchovers and recovery at the data center level for multi-data-center deployment architectures.
Fund Security Monitoring
The Fund Security Monitoring platform is a real-time verification platform that ensures fund security. It uses a bypass method to analyze fund flows within business processes and provides real-time alerts to prevent fund loss as money moves through your business systems.
End-to-end Stress Testing
End-to-end Stress Testing (Loadcenter) is a one-stop stress testing service for businesses that includes performance testing, report generation, and risk management. Built on Ant Group's extensive experience with online end-to-end stress testing, Loadcenter provides a high-fidelity, low-cost online testing experience with effective risk identification.