Infosys
Iceberg, Doris, Trino
Bangalore, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.6M live jobs from 190,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Step into a high-impact role where you’ll lead the design and evolution of modern analytics platforms powered by Iceberg, Doris, and Trino. You’ll work at the intersection of data engineering and query performance, helping teams unlock fast, reliable insights from large-scale datasets. This role is ideal for someone who enjoys solving complex data architecture challenges, optimizing distributed systems, and guiding engineers toward clean, scalable implementations. You’ll collaborate closely with data engineers, platform teams, and stakeholders to build dependable data products, improve developer experience, and ensure data is accessible, governed, and performant. If you’re excited by open table formats, high-concurrency analytical workloads, and building systems that make data truly usable across an organization, this is a great place to grow and lead with purpose.
Responsibilities
Key Responsibilities:
Lead the architecture and implementation of lakehouse and analytics solutions using Iceberg, Doris, and Trino for scalable querying and reporting.
Design and maintain Iceberg table layouts, partitioning strategies, schema evolution patterns, and data lifecycle management (compaction, snapshots, retention).
Build and optimize distributed query workflows in Trino, including connector configuration, query tuning, resource governance, and workload management.
Develop and optimize analytical data models and ingestion patterns leveraging Doris for high-performance OLAP workloads.
Implement robust batch/stream processing pipelines using Spark, ensuring correctness, scalability, and cost efficiency.
Establish performance benchmarks, monitor SLAs, and troubleshoot production issues across compute, storage, and query layers.
Drive best practices for data quality, reliability, and operational excellence through automation, documentation, and runbooks.
Mentor engineers, conduct design reviews, and lead technical decision-making aligned with long-term platform goals.
Minimum Qualifications:
Bachelor’s or Master’s degree in BTECH, MTECH, MCA, MSC or a related field.
6–8 years of experience in data engineering, analytics engineering, or building distributed data platforms.
Strong hands-on expertise with Iceberg, including table design, partitioning, schema evolution, and maintenance operations.
Strong hands-on expertise with Trino for federated/distributed querying, performance tuning, and operational troubleshooting.
Strong hands-on expertise with Doris for OLAP use cases, data modeling, and query performance optimization.
Proven experience building data pipelines using Spark in production environments.
Solid understanding of distributed systems, query execution concepts, and data storage formats for analytics workloads.
Technical requirements
Iceberg, Doris, Trino
Additional responsibilities
Preferred Qualifications:
Experience designing end-to-end lakehouse architectures integrating Iceberg with multiple compute engines and downstream consumers.
Advanced expertise in query optimization techniques (statistics, partition pruning, file sizing, caching strategies) across Trino and OLAP systems.
Experience with Spark optimization (shuffle tuning, join strategies, adaptive execution) and building reusable pipeline frameworks.
Strong operational ownership: monitoring, alerting, incident management, and capacity planning for analytics platforms.
Ability to lead cross-team technical initiatives, influence standards, and improve platform adoption through enablement and documentation.
Education
MCA,MTech,Bachelor of Engineering,BTech
Listed on hirly, a job board. hirly is not the employer: Infosys is hiring for this role.
Similar jobs
- Data Engineer – Apache Spark | Kafka | Flink | Trino | Iceberg | Big Data | Streaming | Data Platform 4–8 YearsCisco · Bangalore, IndiaFirst seen today
- Data Engineer – Apache Spark, Iceberg & KafkaApplicantz · IndiaFirst seen 7d ago
- Iceberg, Doris, TrinoInfosys · Bangalore, IndiaFirst seen 6d ago
- Data Engineer (PySpark, Redshift, Iceberg, Airflow)EY · São Paulo JK, SP, BRFirst seen 7d ago
- Senior Data Engineer (GCP • Python • Iceberg • Delta Lake • Kafka • Snowflake • Databricks)Railroad19 · U.S. RemoteFirst seen 22d agoremote
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job