hirly

Trintech

Principal AI/ML Engineer Lead

India - Bangalore

See how you match this job — and similar ones. Free.

Upload your resume and hirly scores it against this role at Trintech first, then against similar open jobs, and shows where you fit and why.

PDF or DOCX, up to 12MB. No sign-up to see your matches.

Get past the screening software and onto a recruiter's desk

hirly rewrites your resume for this job — matching the keywords and skills in the posting, moving your most relevant experience to the top, and writing a cover letter to fit. About 30 seconds.

  • Keywords matched to this posting
  • Fit score before you apply
  • Cover letter included

Matched against 2.6M live jobs from 190,000+ employers in 200+ countries.

Tailor my resume for this job →

Apply from your AI assistant

Connect hirly to Claude and ask it to apply to this job. hirly tailors your resume, fills the employer’s form and asks before sending. ChatGPT: manual setup today.

Some employer sites stop an application at a CAPTCHA or sign-in and hand it back with a link. Applying needs a paid plan. Works with any assistant that supports MCP.

hirly's read of this role

Role family
Data & ML
Seniority
Lead / management
Country
IN
Work mode
On-site / unstated
First seen by hirly
27 Sept 2026

Derived automatically from the posting. Upload your resume above to see how the role scores against it.

the posting

Description

WHAT YOU’LL DO

  • LLM Engineering Standards Across Agent Pods
  • Define and own LLM engineering standards across all Agent Stream pods — agent framework conventions, prompt lifecycle standards, eval harness design patterns, guardrail implementation, and confidence threshold calibration methodology that all Senior AI/ML Engineers follow.
  • Own the Langfuse observability framework at platform level — instrumentation standards, trace validation patterns, eval pipeline design, prompt regression testing, and model version regression detection. Pod-level Langfuse usage is consistent because this role defines how it is done.
  • Standardise RAG pipeline architecture across pods — embedding strategy, vector database selection and management, retrieval strategy, reranking, and structured output design for financial document reasoning. Shared RAG infrastructure is your design.
  • Review and challenge agent design proposals from pod AI/ML engineers — raise the bar on prompt design, memory architecture, evaluation rigour, and production reliability across the stream.
  • Contribute to AI/ML hiring — define the technical bar for AI/ML engineers across the stream, participate in interviews, and ensure hiring standards are consistent across pods.
  • Memory Architecture Ownership
  • Own the memory architecture strategy for the AI Platform — designing the personalised, complex memory layer that agents depend on for continuity, context, and adaptive behaviour across sessions and users.
  • Design and implement the tiered memory architecture for agent workflows — working memory (in-context), episodic memory (past interactions), semantic memory (extracted facts and preferences), and procedural memory (agent instruction updates). Select and implement the right memory framework for each tier: LangMem for LangGraph-native flows, Mem0 for managed personalisation, Zep/Graphiti for temporal and knowledge-graph reasoning, or Letta for explicit OS-style memory management.
  • Define memory hygiene standards — extraction policies, deduplication, contradiction resolution, and forgetting policies for agents that write aggressively to memory at scale.
  • Own context window management strategy — how long-running agent workflows handle context pressure, when to compress, when to retrieve from memory, and how to maintain coherence across multi-step financial close workflows.
  • Platform AI/ML Contribution
  • Contribute hands-on to Platform Team AI capabilities — working with the Platform Architect on how RAG, memory, and eval infrastructure is exposed as shared platform services that agent pods consume.
  • Stay current on the LLM and agent engineering landscape — evaluate new frameworks, protocols, and tooling (MCP, A2A, new memory systems, emerging eval approaches) and bring informed recommendations to the Director of Engineering on what to adopt and when.
  • Identify and address systemic AI/ML quality gaps across pods — inconsistent eval practices, weak guardrail implementations, or memory architectures that will not scale.

WHO YOU ARE

  • Extensive experience in AI/ML engineering, LLM engineering, data science, software engineering, or a related technical discipline, including experience delivering production-grade AI or ML solutions.
  • Strong hands-on experience with production-grade LLM agent development, including LangChain, LangGraph, or similar agent frameworks.
  • Experience defining platform-level engineering standards, architecture patterns, reusable frameworks, or technical practices across multiple teams or product areas.
  • Strong understanding of prompt lifecycle management, including prompt versioning, rollback, environment-specific configuration, evaluation harnesses, and regression testing.
  • Experience with LLM evaluation and observability practices, including confidence threshold calibration, guardrail design, model regression detection, trace validation, and tools such as Langfuse or similar platforms.
  • Experience designing RAG pipeline architecture, including embedding strategies, vector database selection, retrieval strategies, hybrid search, reranking, and structured output design.
  • Experience designing or implementing agent memory systems, context management strategies, or long-running agentic workflows.
  • Strong data science foundation, including model evaluation, statistical reasoning, experimental design, and understanding of model behavior in non-deterministic or edge-case scenarios.
  • Production engineering experience with Python and related API frameworks, such as FastAPI or similar tools.
  • Experience with PostgreSQL, pgvector or similar vector database technologies, Docker, Kubernetes, and Azure OpenAI or equivalent LLM providers.
  • Ability to evaluate emerging AI/ML technologies, make informed architecture recommendations, and guide technical decisions across teams.
  • Strong communication, collaboration, technical leadership, and problem-solving skills.
  • Experience with Ragas or equivalent RAG evaluation frameworks preferred.
  • Experience with model fine-tuning or RLHF preferred.
  • Experience with financial close, Record-to-Report, accounting, or enterprise finance software preferred.

At our core, Trintechers stand committed to fostering a culture rooted in our core values – Humble, Empowered, Reliable, and Open. Together, these values guide our actions, define our identity, and inspire us to continuously strive for excellence in everything we do.

All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin or disability.

Original posting on Trintech's site ↗

Listed on hirly, a job board. hirly is not the employer: Trintech is hiring for this role.

Browse similar roles

Want this one?

Upload your resume and hirly rewrites it for this job and writes the cover letter — in about thirty seconds, before you sign up.

Tailor my resume for this job