Takealotcom
Data Principal Engineer
Cape Town
Get past the screening software and onto a recruiter's desk
hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.
- Keywords matched to this posting
- Fit score before you apply
- Cover letter included
Matched against 2.4M live jobs from 200,000+ employers in 200+ countries.
Tailor my resume for this job →Apply from your AI assistant
Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.
Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.
hirly's read of this role
- Role family
- Engineering
- Seniority
- Lead / management
- Country
- ZA
- Work mode
- On-site / unstated
- First seen by hirly
- 16 Sept 2026
Derived automatically from the posting. Upload your resume above to see how the role scores against it.
the posting
Takealot.com , South Africa’s leading online retailer, is looking for a highly talented Data Principal Engineer to join our team based in Cape Town.
We are a young, dynamic, hyper-growth company looking for smart, creative, hard-working people with integrity to join us!
Think you’ve been challenged before? Think again!
Scale: 4 million happy shoppers shop online on takealot.com. Show them what you can do.
Learn: We work with the best of the best, and then some. Code alongside industry leaders and up-skill in record time.
Grow: Expand your career in the fast-growing Takealot Group: takealot.com , Mr D and TFS. We like to promote from within: Here’s your chance.
Purpose of the role:
Our data and analytics platform powers the operational backbone of Takealot Group, from logistics and supply chain to distribution centre analytics and data monetisation products. It's grown fast, and the business increasingly needs real-time operational data that our historically batch-oriented platform wasn't built for.
As Data Principal Engineer, you are the highest individual-contributor technical authority in the data division. The mandate: simplify the ecosystem, build the real-time data layer the business now needs, set the technical standards the Group's data governance programme runs on, and design the automation that keeps governance and platform operations manageable as the Group scales. You'll leave behind an architecture the team can understand, build on, and be proud of.
This is a transformation role, with executive sponsorship, direct business impact, and real autonomy over consequential technical decisions. This is a formalised, senior individual contributor milestone on our technical career ladder, a long-term seat for strong engineers who want to keep growing technically without moving into people management.
Key Responsibilities:
Ecosystem Architecture & Platform Simplification:
Produce a definitive, up-to-date master blueprint of the Group's data architecture, sources, flows, models, KPI mappings and use it to drive a structured simplification programme; cutting over-engineering and technical debt, standardising technology choices across teams, and making it faster to deliver new data products.
Real-Time Operational Enablement:
Logistics, Supply Chain, and Distribution Centre operations need faster access to operational data than our current batch-oriented warehouse provides.
Design and build the event-driven, live data layer that decouples these systems from the historical reporting warehouse, enabling faster and more precise operational analytics across the Group.
Data Governance Standards & Documentation:
Establish and own the technical standards that underpin the Group's data governance programme.
Defining data quality standards, lineage documentation requirements, and data management practices, and building them into sprint workflows as normal engineering practice.
Working with central team SMEs (Data Engineering, Analytics Engineering, BI, DataOps) to turn existing engineering practice into formal, Group-level domain standards.
Building and maintaining a centralised architecture repository as the Group's single source of truth for how data flows across the ecosystem. Given the scale involved, hundreds of systems across 15-20+ business units. This is a phased build: the first 90 days should produce the repository's structure and the first few highest-priority domains, with full coverage growing over the following quarters.
Keeping standards and documentation current as the platform evolves.
AI Enablement, Integration & Automation Platform:
Define the technical guardrails that let teams innovate safely: data contracts, security and privacy controls, model governance patterns, and clear standards for embedding automation into data engineering operations.
Own platform alignment for AI consumption, so BigQuery, Dataform, and Looker expose data that AI tools and copilots can use reliably and safely: documented schemas, semantic layers, data contracts, consistent access patterns.
Lead integration planning for AI tooling on the platform, secure, well-governed connection patterns for AI agents, in line with the Group's AI Data Policy.
Design and build automation that keeps governance and platform operations manageable as the Group scales, automated maturity telemetry from platform metadata (classification tag coverage, lineage completeness, Dataform test coverage, Looker documentation completeness), self-service onboarding tooling, and AI-assisted copilots that cut manual facilitation work across the Group.
Technical Onboarding, Training & Knowledge Transfer:
Replace fragmented, course-heavy onboarding with a structured learning path and reference architecture guides tailored to different skill levels (engineering, analytics, BI).
Design specialist technical training modules covering architecture and standards, and contribute content into the Group's tiered governance training.
Success here means the team can operate without depending on any one person's knowledge.
Technical Leadership & Mentorship:
As the highest individual contributor in the data domain, you weigh in on design disagreements, unblock the most complex cross-functional challenges, and set the engineering standard for the division.
Where a technical recommendation conflicts with a team's delivery priorities, the accountable manager makes the final call, your job is to bring the strongest technical case to that decision.
You coach and mentor senior and staff-level engineers, advise the Engineering Director on platform and governance strategy, and contribute to the technical career path matrix for the department.
Minimum Required Qualification:
A Bachelor's degree in Computer Science, Engineering, Information Systems, or a related field is preferred; equivalent demonstrated experience at this scale and seniority will be considered in place of formal qualifications.
Minimum Required Experience:
8+ years experience in data engineering, including at least 3 years at principal or architect level in a complex, high-scale environment.
A track record designing and delivering large-scale data platforms: lakehouse/warehouse, real-time event streaming, data ingestion frameworks.
Experience leading platform simplification or technical debt reduction, not just greenfield builds.
Experience contributing technical standards into a formal data governance programme (DMBOK familiarity a plus).
Hands-on experience with AI or ML data infrastructure: feature stores, model serving pipelines, data contracts, drift monitoring.
Experience integrating AI/LLM tooling (copilots, agents, RAG) with a data platform, access patterns, semantic layers, and data contracts that let AI tools consume data safely.
A track record of building automation or internal tooling, dashboards, scripts, copilots that cuts manual operational work.
Exposure to logistics, e-commerce, or supply chain data is a plus.
Technical Skills:
Deep GCP expertise, particularly BigQuery and Dataform.
Strong experience with stream processing frameworks (Kafka, Pub/Sub, or equivalent).
Solid command of dimensional and event-driven data modelling; working knowledge of Looker/LookML.
Experience with Infrastructure as Code (Terraform or equivalent) and CI/CD for data pipelines.
Strong Python and SQL; familiarity with dbt-style transformation frameworks and orchestration tools.
Understanding of POPIA obligations as they apply to data processing and governance.
Architecture & Governance:
Can produce clear architecture documentation: data flow diagrams, ADRs, technical blueprints, onboarding guides.
Experience designing for auditability, data lineage tracing, and compliance requirements.
Comfortable weighing build vs. buy trade-offs and driving technology standardisation across teams.
Leadership & Commun
Similar jobs
- Lead Data EngineerAbsa · JohannesburgFirst seen yesterday
- Data EngineerWundermanthompson · Cape Town, Western Cape, South AfricaFirst seen 4d ago
- Data EngineerVmlenterprisesolutions · Cape Town, South AfricaFirst seen 4d ago
- Senior Data EngineerMyhcm · Cape TownFirst seen 6d ago
- Data EngineerMyhcm · Cape TownFirst seen 6d ago
Want this one?
Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.
Tailor my resume for this job