Skill diminta
AWSAirflowAttention to DetailElasticsearchGCPGoGrafanaKubernetesMongoDBMySQLPrometheusPythonRedisTerraform
Deskripsi
What Will You Do
- Build automation to improve operational efficiency and reduce manual work.
- Develop monitoring, alerting, logging, and dashboards to ensure system reliability.
- Manage and improve infrastructure, scalability, availability, security, and cost efficiency.
- Support infrastructure and service connectivity across cloud and containerized environments.
- Troubleshoot incidents and drive root cause analysis and continuous improvement.
- Maintain technical documentation, including runbooks, SOPs, postmortems, and system diagrams.
What Are The Requirements
- Bachelor’s degree in Computer Science, Information Technology, or related fields.
- Minimum 5 years of experience in SRE, DevOps, Cloud Engineering, or related roles.
- Strong experience with Kubernetes, AWS/GCP, Terraform, and Terragrunt.
- Familiar with ArgoCD/FluxCD and observability tools such as Grafana, Prometheus, and Opsgenie.
- Experience with Kafka, Airflow, Databricks, and databases such as MySQL, MongoDB, Elasticsearch, or Redis.
- Proficiency in Golang and/or Python for automation.
- Strong ownership, collaboration, problem-solving, and attention to detail.
- Willing to work WFO in Malang or Jakarta.
- This position will be probation to permanent.