Skill diminta
ApacheApache KafkaDockerKubernetesSQL
Deskripsi
About Us
PT Datacaraka Solusindo
is a leading Information and Communication Technology (ICT) company based in Jakarta, Indonesia. We specialize in delivering innovative IT solutions and support services across various industries. Our mission is to provide high-quality technology solutions that help our clients achieve their business objectives efficiently and securely.
With a strong customer-centric approach, technical expertise, and commitment to excellence, we continuously embrace technological advancements to deliver exceptional value to our clients and partners.
Qualifications
- Bachelor's Degree in Information Technology, Information Systems, Computer Science, or a related field.
- Minimum 3 years of experience as a Data Engineer, Integration Engineer, or in a similar role.
- Hands-on experience with
- Apache NiFi
- , including processors, controller services, and provenance.
- Strong knowledge of
- Apache Kafka
- , including cluster setup, topics, producers/consumers, and stream processing.
- Proficient in writing and optimizing SQL queries using
- Impala
- and/or
- Hive
- for big data analytics.
- Familiar with
- Cloudera Platform (CDP/CDH)
- and the Hadoop ecosystem, including HDFS, Hive, Spark, and Kudu.
- Experience with
- Docker
- and
- Kubernetes
- for application deployment.
- Knowledge of
- Java, Python, or Scala
- is an advantage.
- Excellent analytical thinking, problem-solving, and debugging skills.
- Strong communication and interpersonal skills with the ability to collaborate effectively across cross-functional teams.
- Able to produce clear, well-structured technical documentation.
- Capable of working independently as well as collaboratively in a team-oriented environment.
- Key Responsibilities
- Design, develop, implement, and maintain scalable data pipelines using
- Apache NiFi
- for data ingestion, transformation, and routing.
- Manage and maintain
- Apache Kafka
- as an event streaming platform, including topic configuration, producers, consumers, and cluster administration.
- Develop, optimize, and maintain SQL queries in
- Impala
- to support large-scale data analytics and reporting.
Integrate data pipelines from multiple sources, including databases, APIs, IoT devices, and cloud storage into enterprise analytics platforms.
- Monitor, troubleshoot, and optimize the performance, reliability, and security of Cloudera, NiFi, Kafka, and Impala environments.
- Collaborate with the DevOps team to deploy, manage, and optimize applications using Docker and Kubernetes.
- Perform root cause analysis and resolve issues related to data integration, processing, and system performance.
- Ensure data quality, reliability, and availability across the end-to-end data engineering lifecycle.