Infosys
Spark
Bangalore, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Join a data-driven team where your work with distributed processing helps turn complex datasets into meaningful insights. In this role, you’ll collaborate closely with engineers, analysts, and stakeholders to build reliable, scalable data solutions using Spark, contributing to faster decision-making and better customer outcomes. You’ll be encouraged to take ownership of deliverables, improve performance, and bring clarity to ambiguous problem statements through structured analysis and thoughtful implementation. If you enjoy solving large-scale data challenges, optimizing pipelines, and working in a collaborative environment that values learning and continuous improvement, this opportunity will help you grow your technical depth while making a visible impact across projects and teams.
Responsibilities
Key Responsibilities:
Design, develop, and maintain scalable data processing jobs using Spark for batch and/or near-real-time workloads.
Analyze large datasets to identify trends, anomalies, and data quality issues; implement validation and reconciliation checks.
Optimize Spark applications for performance by tuning partitions, caching strategies, memory usage, and execution plans.
Collaborate with cross-functional teams to translate business requirements into technical solutions and well-defined deliverables.
Implement robust error handling, logging, and monitoring to ensure reliability and easier troubleshooting.
Participate in code reviews, follow engineering best practices, and contribute to reusable components and standards.
Support deployments and production issues by performing root-cause analysis and implementing preventive fixes.
Technical requirements
Primary skills:Technology->Big Data - Data Processing->Spark
Additional responsibilities
Minimum Qualifications:
Bachelor’s degree (or equivalent) in Engineering/Technology/Computer Science or related field (BTech/BE/MSc or equivalent).
3–5 years of experience working on data engineering or big data processing initiatives.
Hands-on experience building and maintaining Spark-based data processing solutions.
Strong understanding of distributed computing concepts and data processing fundamentals.
Ability to work independently on assigned modules and collaborate effectively within a team.
Preferred Qualifications:
Master’s degree (MTech/MCA or equivalent) in a relevant discipline.
Proven experience delivering end-to-end Spark pipelines, including development, testing, and production support.
Experience improving job performance and stability through Spark tuning and structured troubleshooting practices.
Familiarity with building reusable frameworks/components to standardize Spark development across projects.
Strong communication skills to explain technical trade-offs and align solutions with stakeholder expectations.
Good to have skills:
Hadoop, Hive, Kafka, Airflow, Delta Lake
Education
MCA,MSc,MTech,Bachelor of Engineering,BTech
Similar jobs
- Spark or Pyspark & Scala & Snowflake/ Databrick DeveloperIqvia · Bangalore, IndiaFirst seen 2d ago
- ADF (Azure Data Factory)DatabricksPysparkInfosys Limited · Bangalore, Karnataka, IndiaFirst seen 3d ago
- Software Engineer III - Big Data, AWS, Spark - Data EngineeringJPMorganChase · Bangalore, Karnataka, IndiaFirst seen yesterday
- Java Spark developerCitigroup · Chennai, Tamil Nadu, IndiaFirst seen yesterday
- Azure Databricks, PysparkCognizant · Chennai, Tamil Nadu, IndiaFirst seen yesterday
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job