Talent.com
Grupo Myth
Site Reliability EngineerGrupo Myth • São Paulo, Brasil
Procure outras vagas
Site Reliability Engineer

Site Reliability Engineer

Grupo Myth • São Paulo, Brasil
Há 3 dias
Descrição da vaga

As a Site Reliability Engineer, you will play a key role in supporting existing customers with their managed or private cloud deployments, as well as in launching new deployments on major cloud platforms such as Azure, AWS, and GCP. Your mission will include
ensuring the smooth operation, scalability, and security of cloud services, as well as automating processes to increase both efficiency and reliability.

Responsibilities:

1. Deployment Setup and Management:

Lead the design and implementation of new cloud deployments, tailoring solutions to meet stakeholder requirements on platforms like Azure, AWS, GCP, and Kubernetes.

Optimize cloud architectures for scalability and cost-effectiveness, adhering to best practices for networking, security, and access controls.

Gain and maintain deep knowledge of cloud infrastructure providers to create robust solutions.

2. Automation and CI/CD::

Craft and manage automation scripts and infrastructure as code (IaC) with Terraform, Ansible, or CloudFormation.

Deploy CI/CD pipelines to streamline software delivery, testing, and deployment processes, ensuring efficient version control and configuration management.

3. Managed Cloud Support:

Ensure the availability of the services by configuring system monitors and alerts and attending to critical alerts in a timely manner.

Offer continuous support and maintenance for existing deployments, monitoring system performance and swiftly resolving issues to maintain high availability and reliability.

Implement strategies for performance optimization and failure prevention, conducting thorough root cause analyses to avoid future issues.

4. Monitoring and Security:

Establish comprehensive monitoring and alerting systems to oversee customer deployments, setting thresholds for incident response.

Conduct regular security assessments and stay abreast of the latest threats and trends to fortify cloud environments against risks.

5. Collaboration and Knowledge Sharing:

Foster a collaborative environment with product developers, operations, and QA teams to enhance workflows and product quality.

Share knowledge and best practices, contributing to the team’s collective expertise through documentation, training, and mentorship.

  • Location São Paulo
  • Working hours: from 12PM to 9PM BRT
  • On duty: availability needed for On duty work over the weekend (frequency is about 1 weekend every 6 weeks).

Criar um alerta de emprego para esta pesquisa

Site Reliability Engineer • São Paulo, Brasil