Big Data Developer at IntraEdge · Remote only · 4 - 16 years · ₹11L - ₹27L / yr · Remote only · Posted 15 Jun 2022
Company Name: Intraedge Technologies Ltd (https://intraedge.com/" target="_blank">https://intraedge.com/)
Type: Permanent, Full time
Location: Any
A Bachelor’s degree in computer science, computer engineering, other technical discipline, or equivalent work experience
- 4+ years of software development experience
- 4+ years exp in programming languages- Python, spark, Scala, Hadoop, hive
- Demonstrated experience with Agile or other rapid application development methods
- Demonstrated experience with object-oriented design and coding.
Please mail you rresume to poornimakattherateintraedgedotcomalong with NP, how soon can you join, ECTC, Availability for interview, Location

About IntraEdge
About
WE ARE A LARGE PRODUCTS
AND SERVICES ORGANIZATION
We operate with the agility and flexibility of a much smaller firm, which allows us to network talent, manage projects and conduct more business opportunities at a much faster and larger scale.From helping you build the perfect teams, to building products and platforms, we are here to provide the strategic vision and execution of your digital transformation initiatives. Our products include: Truyo, Byndr, learn .
Visit us on on website below:
Connect with the team
Company social profiles
Similar jobs (7)
We are looking for a skilled Python & PySpark Developer with strong expertise in Big Data technologies, Spark, SQL/PL-SQL, and REST API development using Flask or Django. The ideal candidate should have experience building scalable data pipelines, processing large datasets, developing APIs, and working with distributed computing frameworks.
Key Responsibilities
- Develop, optimize, and maintain scalable data pipelines using PySpark and Apache Spark.
- Design, develop, and optimize complex SQL and PL/SQL queries, stored procedures, functions, and database objects.
- Build and maintain RESTful APIs using Flask or Django.
- Develop robust Python applications for data engineering and backend services.
- Process and analyze large-scale datasets using Big Data technologies.
- Optimize Spark jobs for performance, scalability, and reliability.
- Integrate APIs with internal and external systems.
- Collaborate with cross-functional teams including Data Engineers, Data Scientists, and Application Developers.
- Troubleshoot production issues and implement performance improvements.
- Follow coding standards, version control, and CI/CD best practices.
Mandatory Skills
- Strong proficiency in Python programming.
- Hands-on experience with PySpark and Apache Spark.
- Strong SQL coding skills.
- Experience with PL/SQL development.
- Experience in Big Data ecosystem.
- REST API development using Flask or Django.
- Experience in developing and consuming Python APIs.
- Knowledge of data processing, ETL, and distributed computing.
- Experience with Git/version control.
Preferred Skills
- Experience with Hadoop ecosystem (Hive, HDFS, YARN).
- Exposure to cloud platforms such as AWS, Azure, or GCP.
- Knowledge of Airflow or other workflow orchestration tools.
- Experience with Docker and Kubernetes.
- Familiarity with Kafka or other streaming technologies.
- Understanding of CI/CD pipelines.
Qualifications
- Bachelor's or Master's degree in Computer Science, Information Technology, or a related field.
- 4–8+ years of experience in Python and Big Data development (can be adjusted based on the role).
Required Experience
- Strong hands-on experience in Python, PySpark, and Apache Spark.
- Extensive experience writing optimized SQL and PL/SQL code.
- Experience developing REST APIs using Flask or Django.
- Experience working with large-scale data processing and ETL pipelines.
- Strong analytical, debugging, and problem-solving skills.
Mandatory Skills: Python, PySpark, SQL Coding, Apache Spark, Big Data, Flask/Django (REST API), PL/SQL, Python APIs.
Company: Wissen Technology
Position: Databricks Engineer
Experience: 6-10 Years
Location: Bengaluru/Mumbai
Employment Type: Full-time
About Wissen Technology
Wissen Technology is a global technology services company focused on delivering innovative software engineering, data, and technology solutions to leading enterprises. The company works with clients across industries to build scalable, high-performance technology platforms and solve complex business and technology challenges.
With a strong focus on engineering excellence, innovation, and collaboration, Wissen Technology brings together skilled technology professionals across areas such as software engineering, data engineering, cloud, analytics, and digital transformation.
At Wissen Technology, employees have the opportunity to work on challenging technology projects, collaborate with experienced engineering teams, and contribute to solutions that create measurable business impact.
Key Responsibilities
- Design, develop, and maintain scalable data processing applications using Python, PySpark, and Spark.
- Develop and optimize data pipelines and workflows on Databricks.
- Collaborate with data engineers, data scientists, business stakeholders, and other technical teams to understand requirements and deliver high-quality solutions.
- Ensure data integrity, quality, performance, and reliability across data processing pipelines.
- Write clean, maintainable, scalable, and efficient code following established coding standards and best practices.
- Perform data analysis and implement appropriate data validation and quality checks.
- Monitor, troubleshoot, and optimize performance issues across data workflows and pipelines.
- Work with relational databases and develop efficient SQL queries for data extraction and transformation.
- Participate in code reviews, testing, deployment, and continuous improvement of data engineering solutions.
- Use Git/version control and follow established software development and deployment practices.
Required Skills & Qualifications
- Bachelor’s or Master’s degree in Computer Science, Engineering, Information Technology, or a related field.
- Proven experience as a Databricks Developer, Data Engineer, or similar role.
- Strong hands-on expertise in Apache Spark and PySpark.
- Strong programming skills in Python; experience with Scala is an advantage.
- Strong proficiency in SQL and hands-on experience with relational databases.
- Practical experience developing and optimizing data pipelines and data processing applications.
- Familiarity with Git and version control systems.
- Strong understanding of data engineering concepts, data transformation, and data validation.
- Excellent analytical and problem-solving skills.
- Strong communication and collaboration skills, with the ability to work effectively with cross-functional teams.
The recruiter has not been active on this job recently. You may apply but please expect a delayed response.
Must have experience in Java
Must have experience in Spark
Must have experience in ETL coding
Strong expertise in coding
1st virtual , 2nd round F2F
Python pyspark, SQL, data engineer
5+yrs
9+yrs
Bangalore/Hyderabad
immediate to 15days.
Dear Candidate,
Thanks for showing interest in the opportunity!!!
As discussed, we have an opening for ETL Developer – Hadoop & PySpark with Mphasis for the Bangalore location – Permanent.
Job Description:
We are looking for an experienced ETL Developer with strong expertise in ETL development, Hadoop, and PySpark. The candidate should have hands-on experience in data processing, building and maintaining ETL pipelines, data transformation, and handling large datasets using Big Data technologies.
Primary & Mandatory Skills:
- ETL Development
- Hadoop / HDFS
- PySpark / Apache Spark
- Data Extraction, Transformation & Loading
- SQL and Data Processing
Good to Have:
- Hive / Spark SQL
- Data Pipeline Development
- Big Data Processing
- Python Programming
Company: Mphasis
Role: ETL Developer – Hadoop & PySpark
Experience: As per requirement
Location: Bangalore
Employment Type: Permanent
Description
We are looking for Senior Data Engineers to join our Data Platform team and build scalable, high-performance data platforms that power data processing, analytics, and downstream applications.
The ideal candidate will have strong experience in distributed data processing, ETL pipelines, and Big Data technologies, with hands-on expertise in Apache Spark and Python Scala.
You will be responsible for designing, developing, and optimizing large-scale data pipelines while collaborating closely with cross-functional engineering teams to build reliable, production-grade data solutions.
Key Responsibilities
- Design, develop, and maintain scalable ETL and data processing pipelines for large-scale datasets.
- Build and optimize distributed data applications using Apache Spark and Python Scala.
- Develop reliable, high-performance data pipelines for batch and streaming workloads.
- Design and manage data workflows using Apache Airflow.
- Build and operate data workloads on AWS, with strong usage of Amazon S3 for large-scale data storage.
- Work with large datasets to ensure data quality, consistency, reliability, and performance.
- Collaborate with engineering, product, analytics, and other platform teams to deliver robust data solutions.
- Optimize data workflows for scalability, reliability, performance, and cost efficiency.
- Troubleshoot production issues, identify bottlenecks, and continuously improve platform performance.
Requirements
Candidates who demonstrate:
- 5+ years of experience in Data Engineering, Big Data Engineering, or a similar role.
- Strong hands-on experience with Apache Spark and Scala.
- Experience designing, building, and maintaining large-scale ETL pipelines.
- Strong hands-on experience with AWS, particularly Amazon S3.
- Hands-on experience with Apache Airflow for workflow orchestration and scheduling.
- Strong SQL skills and a solid understanding of distributed data processing concepts.
- Experience working with batch and/or streaming data pipelines.
- Excellent debugging, problem-solving, and performance optimization skills.
- Strong communication and collaboration skills.
Good to Have
- Experience with Databricks and the broader Databricks data platform.
- Familiarity with streaming technologies such as Apache Kafka.
- Experience working on large-scale data platforms handling high-volume data workloads.
- Exposure to additional AWS data services and cloud-native data architectures.
Company Name – Wissen Technology
Group of companies in India – Wissen Technology & Wissen Infotech
Work Location – Whitefield, Bangalore
Website and Company profile:
www.wissen.com
LinkedIn Page:
https://www.linkedin.com/company/wissen-technology/
While you may already know about Wissen and the company history, here is a quick rundown for you.
About Wissen Technology:
· The Wissen Group was founded in the year 2000. Wissen Technology, a part of Wissen Group, was established in the year 2015.
· Wissen Technology is a specialized technology company that delivers high-end consulting for organizations in the Banking & Finance, Telecom, and Healthcare domains. We help clients build world class products.
· Our workforce has highly skilled professionals, with leadership and senior management executives who have graduated from Ivy League Universities like Wharton, MIT, IITs, IIMs, and NITs and with rich work experience in some of the biggest companies in the world.
· Wissen Technology has grown its revenues by 400% in these five years without any external funding or investments.
· Globally present with offices US, India, UK, Australia, Mexico, and Canada.
· We offer an array of services including Application Development, Artificial Intelligence & Machine Learning, Big Data & Analytics, Visualization & Business Intelligence, Robotic Process Automation, Cloud, Mobility, Agile & DevOps, Quality Assurance & Test Automation.
· Wissen Technology has been certified as a Great Place to Work®.
· Wissen Technology has been voted as the Top 20 AI/ML vendor by CIO Insider in 2020.
· Over the years, Wissen Group has successfully delivered $650 million worth of projects for more than 20 of the Fortune 500 companies.
· We have served client across sectors like Banking, Telecom, Healthcare, Manufacturing, and Energy. They include likes of Morgan Stanley, Goldman Sachs, MSCI, StateStreet, Flipkart, Swiggy, Trafigura, GE to name a few.
About Role :
Key Responsibilities
- Build and maintain data transformation pipelines using java Spark
- Develop and optimize large-scale/CPU intensive data processing using Apache Spark
- Orchestrate workflows using Airflow
- Implement data quality checks, testing, and monitoring for pipeline. Good to have exposer into managing metadata, cataloguing, and lineage
- Support schema evolution, backfills, and incremental processing
- Ensure pipelines meet SLAs for freshness, reliability, and performance
- Expertise/working knowledge in Spark and HBase(semantic layer, virtual datasets, Reflections)
Required Skills & Qualifications
- Strong hands-on experience with
- HBase
- Apache Spark
- Experience with HBase or similar lakehouse query engines
- Airflow
- Understanding of data catalogs and lineage (e.g., OpenLineage, DataHub, Apache Polaris , openlineage)
- Proficiency in Java
- Experience with Git-based development and CI/CD
Nice-to-Have Skills
- OpenTable format/Iceberg ,Apache Arrow
- CDC-based analytics pipelines
- Cloud platforms (AWS)
- Kubernetes-based data platforms








