Site Reliability Engineer
Established in 2018, Bybit is one of the world’s leading cryptocurrency exchanges and digital financial platforms, serving over 80 million users across more than 200 countries and regions.
Established in 2018, Bybit is one of the world’s leading cryptocurrency exchanges and digital financial platforms, serving over 80 million users across more than 200 countries and regions. Powered by world-class technology and a user-first mindset, Bybit delivers a seamless ecosystem across trading, payments, wealth management, custody, institutional services, and Web3 — connecting users to the future of digital finance.
Our core values define how we build. We listen, care and improve to create products and experiences that put users first. Backed by a global team of ambitious builders, problem-solvers, and innovators, we foster a high-performance and fast-moving environment where talent is empowered to drive real impact at the global scale. Supported by 24/7 multilingual customer service and a strong commitment to innovation, we are shaping the future of finance through technology, collaboration, and bold execution.
Today, Bybit is recognized as one of the most trusted and transparent platforms in the digital asset industry, continuing to expand its global presence while building the infrastructure for the next generation of financial services.
Responsibilities:
Control all releases to EU sites, access all relevant systems and data on AWS Singapore and EU cloud services.
Implement system change control, monitor performance, and prepare capacity planning documents to support scalability.
Regularly evaluate the provided security requirements and fix vulnerabilities.
Ensure that all systems are securely backed up before any changes, and maintain a strong backup strategy to protect data integrity.
Responsible for managing recovery efforts, including rollback procedures when necessary, to minimize downtime and maintain operational stability.
Ensure reliable backup and recovery processes to support business continuity and system resilience.
Qualifications:
Strong command of English and Chinese
Proficient in Linux operating systems, familiar with Shell, Python or other scripting languages.
Have rich experience in system operation and troubleshooting, and be able to quickly identify and resolve problems.
Familiar with common operation and maintenance tools, such as Ansible, Puppet, Chef, SaltStack, etc.
Familiar with container technology and related ecosystems, such as Docker, Kubernetes, etc.
Familiar with cloud computing platforms, such as AWS, Tencent Cloud.
Familiar with network protocols and network architecture, and have certain network troubleshooting capabilities.
Posted July 27, 2026