Convegenius
Junior Data Engineer
Shimla, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.7M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Data & ML
- Seniority
- Entry level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 30 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Role Overview
We are looking for a Junior Data Engineer to join our Data Platform team. You will design and maintain scalable data pipelines and architectures using AWS services, enabling reliable data movement, transformation, and analytics at scale. You will collaborate with analytics, product, and engineering teams to support reporting, dashboards, and insights for millions of students and schools.
Key Responsibilities
- Design, build, and maintain ETL/ELT pipelines for large-scale data ingestion, transformation, and loading.
- Develop and optimize Spark and PySpark jobs for batch and real-time data processing.
- Work with AWS services (S3, Glue, Lambda, Redshift, Athena, EMR) to manage the data ecosystem.
- Support the design and implementation of Data Lake and Data Warehouse architectures.
- Implement data validation, partitioning, and schema management for efficient query performance.
- Collaborate with data analysts and BI teams to ensure data availability and consistency.
- Maintain data lineage, metadata, data quality, and governance.
- Implement monitoring and alerting for data pipelines.
- Use Git and CI/CD tools to manage code and automate deployment of data workflows.
Qualifications
- Education: Bachelor's degree in Computer Science, Information Technology, Data Engineering, or a related field.
- Experience: 3–5 years of hands-on experience in data engineering, pipeline development, or cloud-based data systems.
- Core Skills: Strong knowledge of SQL and Python/PySpark.
- Cloud & Data Stack: Practical experience with AWS (S3, Glue, Lambda, Redshift, Athena, EMR, Step Functions).
- Architecture: Solid understanding of Data Lake architecture, ETL/ELT frameworks, and data warehousing concepts.
- Frameworks: Familiarity with Delta Lake, Spark SQL, or other big data frameworks, along with data modeling and performance tuning.
Nice to Have
- Exposure to GCP (BigQuery, Dataflow) or Azure (Data Factory, Synapse, Databricks).
- Experience with Databricks, PostgreSQL, MySQL, NoSQL, or Airflow.
- Understanding of DevOps practices, CI/CD pipelines, and infrastructure automation.
- Prior experience in an EdTech or public data ecosystem.
Listed on hirly, a job board. hirly is not the employer: Convegenius is hiring for this role.
Similar jobs
- Sr Data Engineer IThe Walt Disney Company · IndiaFirst seen today
- Data Engineer, Specialist (PITech- Core - Middle Office Team 1)The Vanguard Group · Hyderabad, Telangana, IndiaFirst seen today
- AI Data Engineer - SeniorCummins · Pune, Maharashtra, IndiaFirst seen today
- Azure Data EngineerCGI · IndiaFirst seen today
- Data EngineerSolventum · IndiaFirst seen today
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job