Job Summary
We are seeking a dynamic and highly skilled Cloud Operations Specialist to lead the management, deployment, and optimization of cloud infrastructure across multiple platforms. In this role, you will be responsible for ensuring the stability, security, and scalability of cloud services, leveraging your expertise in cloud architecture, automation, and distributed systems. Your proactive approach will drive continuous improvement in cloud operations, supporting our organization’s digital transformation and innovation initiatives. This position offers an exciting opportunity to work with cutting-edge technologies and contribute to a robust cloud ecosystem that empowers our business growth.
Responsibilities
- Manage and monitor cloud infrastructure across public cloud providers such as AWS, Google Cloud Platform, Azure, and Rackspace, ensuring high availability and performance.
- Automate deployment processes using Infrastructure as Code (IaC) tools like Terraform, Ansible, and Puppet to streamline provisioning and configuration management.
- Implement and maintain scalable cloud architectures utilizing microservices, containers (Docker, Kubernetes), and virtualization technologies like VMware and OpenStack.
- Develop and optimize cloud services including SaaS (Software as a Service), PaaS (Platform as a Service), IaaS (Infrastructure as a Service), Virtual Private Clouds (VPCs), and Virtualization environments.
- Ensure security through effective identity and access management (IAM) solutions, system hardening practices, VPN configurations, and compliance with security standards.
- Collaborate with development teams to facilitate DevOps automation, CI/CD pipelines using Jenkins or similar tools, and scripting in Python, Bash, PowerShell or Ruby for system management tasks.
- Monitor distributed systems’ health using RESTful APIs, web services, and logging tools; troubleshoot issues promptly to minimize downtime.
- Maintain comprehensive documentation of cloud architecture, deployment procedures, system configurations, and incident reports to support ongoing operations.
Skills
- Extensive experience with cloud computing platforms such as AWS, Google Cloud Platform, Azure, OpenStack or Rackspace.
- Strong knowledge of service-oriented architecture (SOA), RESTful APIs, web services development and integration.
- Proficiency in virtualization technologies including VMware and OpenStack; experience with Docker containers and Kubernetes orchestration.
- Expertise in managing IT infrastructure with Infrastructure as Code (IaC) tools like Terraform, Ansible or Puppet for automated deployment.
- Familiarity with cloud management tools for monitoring performance, security compliance, and resource utilization.