hirly

Convegenius

Junior Data Engineer

Shimla, India

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Convegenius first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.7M live jobs from 200,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Data & ML
Seniority
Entry level
Country
IN
Work mode
On-site / unstated
First seen by hirly
30 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Role Overview

We are looking for a Junior Data Engineer to join our Data Platform team. You will design and maintain scalable data pipelines and architectures using AWS services, enabling reliable data movement, transformation, and analytics at scale. You will collaborate with analytics, product, and engineering teams to support reporting, dashboards, and insights for millions of students and schools.

Key Responsibilities

  • Design, build, and maintain ETL/ELT pipelines for large-scale data ingestion, transformation, and loading.
  • Develop and optimize Spark and PySpark jobs for batch and real-time data processing.
  • Work with AWS services (S3, Glue, Lambda, Redshift, Athena, EMR) to manage the data ecosystem.
  • Support the design and implementation of Data Lake and Data Warehouse architectures.
  • Implement data validation, partitioning, and schema management for efficient query performance.
  • Collaborate with data analysts and BI teams to ensure data availability and consistency.
  • Maintain data lineage, metadata, data quality, and governance.
  • Implement monitoring and alerting for data pipelines.
  • Use Git and CI/CD tools to manage code and automate deployment of data workflows.

Qualifications

  • Education: Bachelor's degree in Computer Science, Information Technology, Data Engineering, or a related field.
  • Experience: 3–5 years of hands-on experience in data engineering, pipeline development, or cloud-based data systems.
  • Core Skills: Strong knowledge of SQL and Python/PySpark.
  • Cloud & Data Stack: Practical experience with AWS (S3, Glue, Lambda, Redshift, Athena, EMR, Step Functions).
  • Architecture: Solid understanding of Data Lake architecture, ETL/ELT frameworks, and data warehousing concepts.
  • Frameworks: Familiarity with Delta Lake, Spark SQL, or other big data frameworks, along with data modeling and performance tuning.

Nice to Have

  • Exposure to GCP (BigQuery, Dataflow) or Azure (Data Factory, Synapse, Databricks).
  • Experience with Databricks, PostgreSQL, MySQL, NoSQL, or Airflow.
  • Understanding of DevOps practices, CI/CD pipelines, and infrastructure automation.
  • Prior experience in an EdTech or public data ecosystem.
Original posting on Convegenius's site ↗

Listed on hirly, a job board. hirly is not the employer: Convegenius is hiring for this role.

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job