Join The Everforth GlideFast Team
Work with the best consultants and developers in the industry.
GlideFast Consulting - Cloud Engineer (Remote)
Reference
JBRQ0001010
Job Details
What is the role?
- Administer, monitor, and maintain MySQL and MariaDB database environments supporting mission-critical enterprise applications
- Manage database availability, performance, security, capacity, replication, and recoverability across production and non-production environments
- Perform database patching, version upgrades, configuration tuning, maintenance windows, and post-change validation
- Support backup, restore, and disaster recovery procedures, including periodic testing and documentation of recovery steps
- Maintain database user access, privileges, service accounts, and administrative controls in alignment with security and change management practices
- Monitor and troubleshoot MariaDB replication topology, replication lag, failover readiness, and data consistency issues
- Diagnose and resolve database performance issues using query analysis, indexing, execution plans, connection/session analysis, slow query logs, and resource utilization metrics
- Proactively identify database bottlenecks, poorly performing jobs, storage constraints, and configuration risks before they impact availability
- Partner with Linux, application, middleware, and monitoring teams to resolve cross-platform performance issues
- Contribute to database modernization planning, including potential future migrations or platform changes
- Administer, monitor, and maintain Red Hat Enterprise Linux environments supporting mission-critical enterprise applications
- Perform system patching, OS upgrades, package management, configuration changes, maintenance windows, and post-change validation
- Manage user access, sudo privileges, SSH configuration, service accounts, file permissions, and administrative controls in alignment with security and change management practices
- Maintain core Linux services and system components, including systemd services, logs, cron, networking, firewalls, storage mounts, LVM, and certificate-related support
- Support backup, restore, and disaster recovery procedures by validating host readiness, service recovery steps, and operational runbooks
- Participate in a scheduled 24x7 on-call rotation for P1/P2 incident response and escalated production support
- Act as an escalation resource for tier 2/3 and advanced database incidents, including root cause analysis and corrective action planning
- Create and maintain runbooks, operational procedures, health check reports, change documentation, and knowledge articles
- Support data clone, refresh, and environment synchronization activities as part of regular maintenance cycles