Logo for Venon Solutions

#1007 - Principal AI Engineer

Role overview

Qualifications

  • Advanced English level (C1/C2)
  • 10 or more years of experience in software/AI engineering
  • 4/5 years of hands-on experience with LLM orchestration frameworks
  • Strong stakeholder collaboration and problem-solving skills

Responsibilities

  • Design the orchestration and abstraction layers of the central AI system
  • Design, build, and operate MCP servers and set standards for tools
  • Establish when to use specialized sub-agents versus directly exposing tools
  • Define evaluation for orchestration and retrieval quality

Key facts

  • Remote from: Latin America
  • Full time
  • Senior (5-10 years)
  • AI Engineer
  • English

Hard skills

Other skills

  • Communication
  • Collaboration
  • Problem Solving

About the company

Venon Solutions logo

Venon Solutions

IT Services & IT Consulting

We provide nearshoring of fully dedicated IT professionals from Latin America and build remarkable process-driven + user-centered software to help companies to scale their teams quickly with the highest customer satisfaction standards We have 14+ years developing software and nearshoring for global companies such as Volkswagen, FIAT, IVECO, government agencies and other companies from around the globe.

Company details

Company typeSME
IndustryIT Services & IT Consulting
Company size51 - 200

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Job Opportunity only available for professionals located in Latin America.

The Principal Engineer, AI Orchestration & Retrieval defines how the enterprise's central AI system is composed – the orchestration and abstraction layers that connect LLMs to tools, data, and one another, and the retrieval systems that ground them. This role sets the strategy and builds the reality for how we build and expose tools (including MCP servers), how we structure retrieval and chunking, and when to rely on specialized sub-agents versus directly exposing tools to a model.

Requirements:

  • Advanced English level (C1/C2) to ensure fluent communication across teams.
  • 10 or more years of experience in software/AI engineering, with hands-on experience building LLM orchestration, agents, and retrieval systems.
  • +4/5 years of deep hands-on experience with LLM orchestration frameworks (e.g., LangGraph, LlamaIndex, Semantic Kernel, or equivalents) and agentic patterns.
  • Direct experience building MCP servers and tool/function-calling integrations.
  • Evidence-based opinions on the optimal number of tools to expose to an LLM and the optimal number of APIs per MCP server, and on overall tool-surface design.
  • A clear, defensible point of view on specialized sub-agents versus direct tool exposure, and the tradeoffs of each.
  • Deep experience with retrieval/RAG: chunking strategies, embeddings, vector databases, hybrid search, and re-ranking.
  • Experience designing abstraction layers and platform APIs that many teams build on top of.
  • Strong understanding of context-window management, prompt/context assembly, and cost/latency optimization.
  • Experience with evaluation and observability for agentic and retrieval systems.
  • Ability to set strategy and standards while remaining hands-on in code.
  • Strong stakeholder collaboration and problem-solving skills.

Essential Responsibilities:

  • Design the orchestration and abstraction layers of the central AI system that connect LLMs to tools, data, and sub-agents.
  • Design, build, and operate MCP (Model Context Protocol) servers and set standards for how tools are defined, exposed, and versioned.
  • Define tool-surface strategy: the optimal number of tools exposed to an LLM, the optimal number of APIs per MCP server, and how to keep tool surfaces coherent and discoverable.
  • Establish when to use specialized sub-agents versus directly exposing tools to a model, and design the corresponding multi-agent patterns.
  • Design retrieval (RAG) systems: chunking strategies, embedding models, vector stores, hybrid/keyword search, re-ranking, and context assembly.
  • Define abstraction layers that decouple product teams from the underlying models, tools, and providers.
  • Build routing, context-window management, and memory strategies for agentic workflows.
  • Define evaluation for orchestration and retrieval quality (retrieval precision/recall, tool-selection accuracy, task success, latency, and cost).
  • Establish observability and tracing across multi-step agent and tool calls.
  • Address safety, guardrails, authentication, and access control across tools and agents.
  • Partner with product teams to onboard their capabilities as tools and agents into the central AI system.
  • Mentor engineers and raise orchestration and retrieval maturity across teams.

What do we offer?

  • Competitive salary in USD.
  • 100% Remote Work.
  • Type of contract: Independent Contractor with Venon Solutions LLC.
  • Contract duration: long-term
  • Benefits: 2 weeks of PTO (paid time off).
  • Holidays: from the North American calendar.
  • Working hours: Full-time EST, fully committed.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

AI Engineer Related jobs

Other jobs at Venon Solutions

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.