Infosys
Iceberg, Doris, Trino
Bangalore, India
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.5M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Seniority
- Mid level
- Country
- IN
- Work mode
- On-site / unstated
- First seen by hirly
- 28 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Step into a high-impact data platform leadership role where you’ll shape how teams store, query, and serve analytics at scale. You’ll lead the design and evolution of a modern lakehouse ecosystem powered by Iceberg, Doris, and Trino, enabling fast, reliable insights across diverse workloads. Working closely with data engineering, analytics, and infrastructure teams, you’ll drive technical decisions, establish best practices, and mentor engineers to deliver production-grade solutions. This role is ideal for someone who enjoys solving complex performance and reliability challenges, building scalable architectures, and turning data into a trusted product for the business. If you’re excited about modern table formats, distributed query engines, and collaborative engineering culture, this is your chance to lead from the front and make a measurable difference.
Responsibilities
Key Responsibilities:
Lead architecture and implementation of lakehouse solutions using Iceberg for table management, governance, and scalable storage patterns.
Design and optimize distributed query workloads using Trino, including catalog configuration, connector strategy, and query performance tuning.
Build and operate high-performance analytics serving layers using Doris, focusing on ingestion patterns, schema design, and workload isolation.
Drive end-to-end data pipeline execution leveraging Spark for batch processing, transformations, and data quality enforcement.
Establish standards for data modeling, partitioning, compaction, file sizing, and lifecycle management to improve cost and performance.
Own production readiness: monitoring, alerting, incident response, root-cause analysis, and continuous performance improvements across the stack.
Collaborate with stakeholders to translate analytical needs into scalable technical designs, delivery plans, and measurable outcomes.
Mentor engineers, conduct design/code reviews, and guide best practices for reliability, maintainability, and secure data access.
Minimum Qualifications:
BTECH, MTECH, MCA, or MSC in Computer Science, Engineering, or a related field.
8–12 years of experience in data engineering, data platform, or analytics infrastructure roles with leadership/ownership responsibilities.
Strong hands-on expertise with Iceberg, Doris, and Trino in production environments, including performance tuning and operational support.
Strong experience with Spark for scalable data processing and pipeline development.
Solid understanding of distributed systems, data storage formats, query optimization, and production troubleshooting practices.
Technical requirements
Iceberg, Doris, Trino
Additional responsibilities
Preferred Qualifications:
Proven experience designing lakehouse architectures, including table layout strategies, compaction approaches, and multi-engine interoperability.
Advanced Trino optimization experience (query plans, statistics, resource groups, connector tuning) for high concurrency and large datasets.
Strong Doris operational expertise including ingestion optimization, indexing strategy, and workload management for BI/analytics use cases.
Experience strengthening platform reliability through observability, capacity planning, and performance benchmarking.
Ability to lead cross-team technical initiatives, influence architecture decisions, and communicate trade-offs clearly to technical and non-technical stakeholders
Education
MCA,MSc,MTech,Bachelor of Engineering,BTech
Similar jobs
- Data Engineer – Apache Spark | Kafka | Flink | Trino | Iceberg | Big Data | Streaming | Data Platform 4–8 YearsCisco · Bangalore, IndiaFirst seen today
- Data Engineer – Apache Spark, Iceberg & KafkaApplicantz · IndiaFirst seen 7d ago
- Iceberg, Doris, TrinoInfosys · Bangalore, IndiaFirst seen 5d ago
- Data Engineer (PySpark, Redshift, Iceberg, Airflow)EY · São Paulo JK, SP, BRFirst seen 6d ago
- Senior Data Engineer (GCP • Python • Iceberg • Delta Lake • Kafka • Snowflake • Databricks)Railroad19 · U.S. RemoteFirst seen 22d agoremote
Browse similar roles
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job