We are seeking a dynamic and highly skilled Site Reliability Engineer (SRE) specializing in Data Platforms to join our innovative team. In this role you will be at the forefront of designing implementing and maintaining robust scalable and highly available cloud-based data infrastructure. Your expertise will ensure seamless data operations across multiple cloud environments empowering our organization to deliver reliable and efficient data services. This position offers an exciting opportunity to work with cutting-edge cloud technologies automate deployment processes and optimize distributed systems to support our data-driven initiatives.
Develop and maintain scalable reliable and secure cloud infrastructure for data platforms utilizing Service-oriented architecture (SOA) principles. Automate deployment processes using Infrastructure as Code (IaC) tools such as Terraform Ansible or Puppet to streamline operations and reduce manual intervention. Manage and optimize cloud environments across multiple platforms including AWS Google Cloud Platform Azure VMware OpenStack Rackspace Citrix and other public cloud services. Design and implement resilient distributed systems capable of handling high volumes of data with scalability and performance in mind. Collaborate with development teams to ensure best practices in microservices architecture RESTful API integrations web services and container orchestration using Kubernetes or Docker. Monitor system health and performance metrics troubleshoot issues proactively to minimize downtime and ensure high availability. Implement security best practices including identity and access management (IAM) system hardening VPN configurations and compliance standards to safeguard sensitive data.
Proven experience managing IT infrastructure in a cloud environment with a strong understanding of cloud architecture principles. Extensive knowledge of cloud computing platforms such as AWS Google Cloud Platform Azure VMware OpenStack Rackspace or Citrix. Hands-on expertise in automating deployment processes using tools like Terraform Ansible Chef or Puppet. Strong background in managing distributed systems with a focus on scalability and reliability experience with Kubernetes and Docker is highly desirable. Proficiency in scripting languages such as Python Bash (Unix shell) PowerShell or Ruby for automation tasks. Familiarity with database systems including MySQL PostgreSQL Oracle or NoSQL solutions like Cassandra or DynamoDB. Deep understanding of service-oriented architecture (SOA) RESTful APIs web services development and management of cloud services including SaaS (Software as a Service) PaaS (Platform as a Service) IaaS (Infrastructure as a Service). Knowledge of virtualization technologies such as VMware or VirtualBox experience with system hardening and security protocols is essential. Experience working within Agile development methodologies strong problem-solving skills combined with excellent communication abilities are vital for success in this role. Join us to be part of a forward-thinking team that thrives on innovation! Your expertise will directly impact the reliability of our data platform infrastructure while helping us push the boundaries of what’s possible in cloud computing and data management. Pay $60.00 - $65.00 per hour Experience Airflow Configuration 4 years (Required) Snowflake 6 years (Required) Python 5 years (Required)
Remote
Remote Poland
Remote Spain
Remote Mexico
Remote Argentina
Remote United States
Remote Brazil
Unlock: Sign Up for free / Sign In and use the searches from your home page or the links in the footer.