Hadoop Developer at Persistent System Ltd · Bengaluru (Bangalore), Pune, Hyderabad · 4 - 6 years · ₹6L - ₹22L / yr · Posted 15 Jul 2022
Location: Bangalore/Pune/Hyderabad/Nagpur
4-5 years of overall experience in software development.
- Experience on Hadoop (Apache/Cloudera/Hortonworks) and/or other Map Reduce Platforms
- Experience on Hive, Pig, Sqoop, Flume and/or Mahout
- Experience on NO-SQL – HBase, Cassandra, MongoDB
- Hands on experience with Spark development, Knowledge of Storm, Kafka, Scala
- Good knowledge of Java
- Good background of Configuration Management/Ticketing systems like Maven/Ant/JIRA etc.
- Knowledge around any Data Integration and/or EDW tools is plus
- Good to have knowledge of using Python/Perl/Shell
Please note - Hbase hive and spark are must.

Similar jobs (2)
We are looking for a Big Data Engineer to build large-scale data processing systems.
Responsibilities
- Build batch and streaming pipelines with Spark and Kafka
- Manage data in the Hadoop ecosystem (HDFS, Hive)
- Write Spark jobs in Scala or PySpark
- Tune jobs for performance and cost
Requirements
- 2+ years of big data engineering
- Strong hands-on Spark experience
- Experience with Hadoop, Hive and Kafka
Job Summary:
- We are seeking an experienced Hadoop Engineer with strong hands-on expertise in MAPR and Hortonworks platform engineering and administration. The successful candidate will join our Hadoop Platform Engineering & Automation team to support our critical enterprise data lake platform. This role emphasizes platform support, patching, vulnerability remediation, L3 issue resolution, and automation initiatives utilizing Linux, Shell Scripting, and Ansible.
Responsibilities:
- Platform Engineering & Administration: Administer, maintain, and optimize MAPR and Hortonworks Hadoop clusters.
- Provide L3 support for complex technical issues related to cluster functioning, performance, and reliability.
- Monitor cluster health, capacity, performance, and security configurations.
- Patching & Vulnerability Remediation: Execute OS-level and platform-level patching across large-scale Hadoop clusters.
- Implement remediation for platform vulnerabilities in accordance with organizational InfoSec policies.
- Collaborate with other support teams to ensure compliance with standards and mitigation timelines.
- Automation Development: Develop, enhance, and maintain automation workflows using Linux, Shell Scripting, and Ansible.
- Automate recurring operational tasks including cluster patch deployment, configuration management, monitoring and ing integrations, and system health checks.
- Operational Support: Troubleshoot node failures, service crashes, cluster imbalance, and distributed computing issues.
- Perform root cause analysis for high-severity incidents.
- Ensure high availability and optimal performance of Hadoop platform services.
- Work closely with engineering teams to support consistency in deployment and configuration processes.
Mandatory Skills:
- Hands-on experience in Hadoop engineering and administration.
- Strong proficiency in MAPR and Hortonworks Administration.
- Deep understanding of Hadoop ecosystem components including HDFS, YARN, MapReduce, Hive, HBase, Spark, and Zookeeper.
- Experience with Linux system administration.
- Strong expertise in Shell Scripting and Ansible Automation.
- Experience with patching and security remediation for large-scale distributed systems.
- Understanding of configuration management, service orchestration, and cluster operations.
Preferred Skills:
- Exposure to DevOps tools such as Git.
- Experience with monitoring tools like Grafana, Prometheus, and Ambari.
- Understanding of ITIL processes including Incident, Change, and Problem management.
Qualifications:
- Strong analytical and troubleshooting skills for L3 support.
- Ability to work independently and in cross-functional teams.
- Excellent communication and documentation skills.
- Strong ownership mindset toward reliability and stability of platforms.







