Logo for LeewayHertz

Technical Architect- Remote, India

Role overview

Qualifications

  • 10+ years in software engineering, data or AI roles, including at least 3 years in an architect or technical lead capacity.
  • Demonstrated experience architecting and delivering production Generative AI systems.
  • Strong hands-on Python, with the ability to prototype designs and review production code.
  • Deep expertise in LLM application architecture: prompt and context engineering.

Responsibilities

  • Own the end-to-end architecture of GenAI solutions across the retrieval, orchestration, model, integration, and deployment layers.
  • Translate ambiguous business problems into AI solution designs with clear scope, feasibility assessment and success metrics.
  • Define reference architectures, design patterns and reusable accelerators for RAG, agentic workflows and LLM integration.
  • Lead model and platform selection, documenting the cost, latency, accuracy and data-residency trade-offs behind each decision.

Key facts

  • Remote from: India
  • Full time
  • Expert & Leadership (>10 years)
  • Technical Architect
  • English

Hard skills

Other skills

  • Consulting
  • Communication
  • Mentorship
  • Problem Solving

About the company

LeewayHertz logo

LeewayHertz

LeewayHertz is one of the first few companies to build and launch a commercial app on Apple's App Store. Our certified designers and developers have designed and developed over 100 digital platforms using AI, IoT, Web3, Metaverse and Blockchain technologies. At LeewayHertz, we have developed digital solutions for Fortune 500 companies and startups to ease their business functions with the latest technologies. Some of our reputed clients include ESPN, NASCAR, Hershey's, McKinsey, P&G, Siemens, 3M, Pearson and more.Being an award-winning software development company, we have also proven our expertise in AI development, offering a comprehensive suite of AI services, including AI/ML strategy consulting, custom model and solutions development, and AI integration and deployment. Our team of certified experts specializes in advanced AI technologies such as ML, computer vision, and natural language processing. LeewayHertz's expertise extends to multiple AI models, including GPT-4, LLaMA, PaLM-2, and more. If you're looking to harness the power of AI for transformative business success, LeewayHertz is your trusted technology partner. Join us at https://www.leewayhertz.com/

Company details

Company size51 - 200

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

This is a remote position.

We are seeking a Technical Architect with 10+ years of overall technology experience, including proven experience architecting and delivering production Generative AI systems. The ideal candidate will combine deep hands-on command of large language models, retrieval-augmented generation, and Agentic architectures with the judgement to design AI solutions that hold up under real enterprise constraints of cost, latency, security, and compliance. This role owns the technical architecture of AI engagements end-to-end: shaping the solution during discovery and pre-sales, defining reference architectures and integration patterns, selecting models and platforms, and guiding delivery teams through implementation. You will work directly with client stakeholders, product managers and engineering leads, and will help set the technical standards of our AI practice. This is a senior, hands-on architecture role rather than a purely advisory one. This role owns multiple concurrent AI engagements from discovery through production. Beyond technical architecture, the Technical Architect is expected to act as a trusted advisor to clients, lead engineering teams, mentor future technical leaders, and drive engineering excellence through hands-on involvement.

Requirements

Responsibilities

  • Own the end-to-end architecture of GenAI solutions across the retrieval, orchestration, model, integration, and deployment layers.
  • Translate ambiguous business problems into AI solution designs with clear scope, feasibility assessment and success metrics.
  • Define reference architectures, design patterns and reusable accelerators for RAG, agentic workflows and LLM integration.
  • Lead model and platform selection, documenting the cost, latency, accuracy and data-residency trade-offs behind each decision.
  • Design the non-functional envelope: scalability, latency budgets, availability, observability and inference cost control.
  • Architect data and retrieval pipelines covering ingestion, chunking, embedding strategy, vector store selection and hybrid search.
  • Define evaluation strategy and guardrails so accuracy, groundedness, safety and hallucination rates can be measured and governed.
  • Embed security, privacy and compliance into the design: PII handling, tenancy isolation, access control and audit.
  • Support pre-sales and discovery through solution workshops, effort estimation, technical proposals and client presentations.
  • Guide delivery teams, run design reviews and mentor engineers, while staying hands-on in prototyping and unblocking hard problems.
  • Maintain architecture documentation and decision records, and assess which advances in generative AI are ready for enterprise adoption.
  • Own multiple client engagements simultaneously while maintaining delivery quality.
  • Lead discovery workshops, challenge assumptions, and refine business requirements into technically sound solutions.
  • Push back on unrealistic timelines, architectures, or requirements using engineering judgement and data.
  • Build strong relationships with Team, product owners, and executive stakeholders.
  • Mentor senior engineers and cultivate future architects and technical leaders.
  • Lead architectural governance, design reviews, and technical decision records.
  • Set engineering standards, coding guidelines, AI development best practices, and review critical code.
  • Remain hands-on by building prototypes, solving difficult technical problems, and contributing production-quality code when needed.
  • Drive cross-project reuse through internal frameworks, accelerators, and reference implementations.
  • Present architecture, trade-offs, risks, and implementation strategy confidently to executive audiences.

Essential Skills

Job

  • 10+ years in software engineering, data or AI roles, including at least 3 years in an architect or technical lead capacity.
  • Demonstrated experience architecting and delivering production Generative AI systems, not only prototypes or POCs.
  • Strong hands-on Python, with the ability to prototype designs and review production code.
  • Deep expertise in LLM application architecture: prompt and context engineering, structured output, tool calling and orchestration.
  • Proven experience designing RAG systems end-to-end, including chunking, embedding selection, hybrid retrieval and re-ranking.
  • Proven ability to manage multiple enterprise AI programs simultaneously.
  • Strong client-facing consulting experience with executive communication.
  • Excellent presentation, whiteboarding, and workshop facilitation skills.
  • Demonstrated experience influencing technical decisions across multiple teams.
  • Experience managing senior engineers and mentoring future technical leaders.
  • Strong engineering judgement balancing quality, cost, delivery timelines, and business value.
  • Comfortable making architectural decisions with incomplete information.
  • Experience with agent and orchestration frameworks such as LangChain, LangGraph, LlamaIndex or CrewAI.
  • Strong knowledge of vector databases (Pinecone, Weaviate, Qdrant, FAISS, pgvector) and their operational trade-offs.
  • Solid machine learning and deep learning fundamentals, including fine-tuning and adaptation approaches (LoRA/QLoRA, PEFT).
  • Strong cloud architecture skills on AWS, Azure or GCP, including their AI/ML and data services.
  • Experience with microservices, API design, event-driven patterns and enterprise system integration.
  • Working knowledge of MLOps and LLMOps: CI/CD, containerization, model versioning, monitoring and rollback.
  • Experience defining LLM evaluation and observability approaches (RAGAS, LangSmith, DeepEval or equivalent).
  • Personal
  • Strong communication and stakeholder management skills, including with non-technical and client-side audiences.
  • Sound engineering judgement, with the confidence to defend a design and the openness to revise it.
  • Strong ownership across the full delivery lifecycle, not only the design phase.
  • Ability to mentor engineers and lead through influence rather than authority.
  • Comfortable operating with ambiguity in a fast-moving technology space.

Personal

  • Strong communication and stakeholder management skills, including with non-technical and client-side audiences.
  • Sound engineering judgement, with the confidence to defend a design and the openness to revise it.
  • Strong ownership across the full delivery lifecycle, not only the design phase.
  • Ability to mentor engineers and lead through influence rather than authority.
  • Comfortable operating with ambiguity in a fast-moving technology space

Preferred Skills

Job

  • Experience architecting multi-agent systems and complex autonomous workflows.
  • Exposure to multimodal AI covering vision, speech or document understanding.
  • Experience with inference optimization and serving at scale (vLLM, TensorRT-LLM, Triton, quantization).
  • Experience deploying open-weight models on-premise or in-VPC for data-sensitive clients.
  • Knowledge of graph-based retrieval (GraphRAG) and knowledge-graph modelling.
  • Familiarity with AI governance and responsible AI frameworks (EU AI Act, NIST AI RMF, ISO/IEC 42001).
  • Experience with data platform architecture and pipelines (Airflow, dbt, Spark, lakehouse patterns).
  • Pre-sales, solutioning or client-facing consulting experience in a services organisation.
  • Domain depth in one or more of BFSI, healthcare, retail, supply chain or manufacturing.
  • Personal
  • Proactive mindset with a genuine interest in tracking a fast-moving field.
  • Consulting orientation, balancing technical ideals against client timelines and budgets.

Personal

  • Proactive mindset with a genuine interest in tracking a fast-moving field.
  • Consulting orientation, balancing technical ideals against client timelines and budgets.

Other Relevant Information

  • Bachelor's or Master's degree in Computer Science, Data Science, Engineering, or a related field.
  • Relevant certifications in AI/ML, cloud architecture (AWS/Azure/GCP), or enterprise architecture are a plus.
  • A portfolio of production Generative AI architectures, open-source contributions or published work is highly desirable.

Benefits

  • This role offers the flexibility of working remotely in India.
LeewayHertz is an equal opportunity employer and does not discriminate based on race, colour, religion, sex, age, disability, national origin, sexual orientation, gender identity, or any other protected status. We encourage a diverse range of applicants.



Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Technical Architect Related jobs

Other jobs at LeewayHertz

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.