Skill diminta
CI/CDDockerElasticsearchGrafanaKubernetesLinuxPrometheusTroubleshooting
Gaji pasar untuk posisi ini
Median Rp9,7 jt/bulan · rentang umum Rp7,5 jt – Rp11,5 jt
Berdasarkan 212 lowongan DevOps & Infrastructure se-Indonesia (semua level).
Gaji lowongan ini 17% di bawah median pasar.
Lihat data gaji selengkapnya →Deskripsi
- Manage and maintain Linux-based production servers supporting telco/VAS platforms
- Ensure system availability, stability, and performance in line with SLA (≥99.9%)
- Handle high transaction volumes and traffic spikes in production environments
- Monitor system performance, logs, and infrastructure health
- Perform troubleshooting, incident response, and root cause analysis
- Manage and optimize Docker-based services
- Operate container orchestration platforms such as Kubernetes or Docker Swarm
- Collaborate with developers on system deployment and CI/CD processes
- Implement system security measures and anomaly monitoring
- Prepare technical documentation and incident reports
- Participate in on-call support for critical system incidents
- Administered and maintained the Elasticsearch Stack (ELK) including Elasticsearch, Kibana, Logstash, Fluent-bit, Filebeat, Metricbeat, Heartbeat, Elastic APM, and Kibana Machine Learning (Mandatory).
- Performed Elasticsearch cluster maintenance, including index lifecycle management (ILM), replica tuning, shard relocation, and storage optimization.
- Job Responsibilities
- Manage and maintain Linux-based production servers supporting telco/VAS platforms
- Ensure system availability, stability, and performance in line with SLA (≥99.9%)
- Handle high transaction volumes and traffic spikes in production environments
- Monitor system performance, logs, and infrastructure health
- Perform troubleshooting, incident response, and root cause analysis
- Manage and optimize Docker-based services
- Operate container orchestration platforms such as Kubernetes or Docker Swarm
- Collaborate with developers on system deployment and CI/CD processes
- Implement system security measures and anomaly monitoring
- Prepare technical documentation and incident reports
- Participate in on-call support for critical system incidents
- Administered and maintained the Elasticsearch Stack (ELK) including Elasticsearch, Kibana, Logstash, Fluent-bit, Filebeat, Metricbeat, Heartbeat, Elastic APM, and Kibana Machine Learning (Mandatory).
- Performed Elasticsearch cluster maintenance, including index lifecycle management (ILM), replica tuning, shard relocation, and storage optimization.
- Job Requirements
- Bachelor’s degree in Information Technology, Computer Science, or related field
- Minimum 5–8 years of experience as a System Administrator
- Strong experience managing Linux production servers (Ubuntu / CentOS / Debian)
- Experience working with high-traffic or production-critical systems
- Hands-on experience with Docker containerization (mandatory)
- Experience with Kubernetes or Docker Swarm
- Good understanding of networking concepts, system architecture, and API integration
- Experience using monitoring and logging tools such as Grafana, Prometheus, Elasticsearch Stack (ELK) including Elasticsearch, Kibana, Logstash, Fluent-bit, Filebeat, Metricbeat, Heartbeat, Elastic APM, and Kibana Machine Learning (a must).
- Experience handling incident management and performance optimization and elasticsearch cluster maintenance, including index lifecycle management (ILM), replica tuning, shard relocation, and storage optimization.
- Willing to participate in 24/7 on-call rotation for critical incidents
- Preferred Background
- Candidates with experience in the following industries are highly preferred:
- Telecommunication companies
- VAS (Value Added Services)
- Content Provider / SMS Gateway companies
- High traffic digital platforms
Experience supporting platforms such as
- SMS Gateway / Messaging Platforms
- OTP or Authentication Systems
- Telco API services