Logo for CodeRoad Inc

Senior Data Engineer

Role overview

Qualifications

  • 4+ years of dedicated data engineering experience with strong command of SQL
  • Solid foundation in dimensional modeling and semantic layer architecture
  • Proven production experience building and maintaining resilient ETL/ELT data pipelines
  • Advanced English fluency (written and spoken)

Responsibilities

  • Design and implement modern semantic models to abstract complex operational databases
  • Build curated, high-performance analytical views and production ELT/ETL pipelines
  • Optimize database query performance and schema definitions for low-latency analytical workloads
  • Lead the implementation of granular data governance frameworks

Key facts

Hard skills

Other skills

  • Problem Solving
  • Collaboration

About the company

CodeRoad Inc logo

CodeRoad Inc

Software Development

CodeRoad provides end-to-end software development services, helping businesses scale with ideal infrastructure solutions. From staff augmentation to dedicated IT teams and general software engineering, our nearshore technology services empower businesses to thrive in an ever-evolving digital landscape.

Company details

IndustrySoftware Development
Company size201 - 500

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Senior Data Engineer

Location/Region

Latin America | 100% Remote

About CodeRoad

CodeRoad provides end-to-end software development services, helping businesses scale with ideal infrastructure solutions. From staff augmentation to dedicated IT teams and general software engineering, our nearshore technology services empower businesses to thrive in an ever-evolving digital landscape.

About the Role

As a Senior Data Engineer at CodeRoad, you will serve as the technical backbone responsible for architecting and optimizing an agent-ready data foundation and semantic layer. You will model complex analytical views, optimize low-latency query performance for downstream LLM retrieval, and anchor robust access control and data privacy policies across modern database environments.

This role is critical to bridging complex operational data systems with cutting-edge AI frameworks, ensuring that downstream AI agents and analytical engines access accurate, highly secure, and optimized data in real time.

Key Responsibilities

  • Design and implement modern semantic models (facts, dimensions, and metrics stores using tools like dbt) to abstract complex operational databases for AI querying and analytical consumption.

  • Build curated, high-performance analytical views and production ELT/ETL pipelines that streamline data retrieval, enhance freshness, and minimize redundant database queries.

  • Optimize database query performance, indexing, and schema definitions for low-latency analytical workloads through systematic query log analysis.

  • Lead the implementation of granular data governance frameworks, including Role-Based Access Control (RBAC), Row-Level Security (RLS), and dynamic data masking for PII protection.

  • Establish comprehensive data quality, pipeline monitoring, and anonymization standards across all analytical layers.

  • Collaborate with cross-functional AI engineering teams to align data structures with the operational requirements of RAG architectures and Agentic AI workflows.

Requirements

  • 4+ years of dedicated data engineering experience with strong, production-grade command of SQL (Azure SQL, PostgreSQL, or modern cloud data warehouses).

  • Solid foundation in dimensional modeling (Kimball methodology) and semantic layer / metrics store architecture (e.g., dbt).

  • Proven production experience building and maintaining resilient ETL/ELT data pipelines.

  • Demonstrated experience implementing database security policies, granular role-based permissions, and dynamic data masking / PII protections.

  • Hands-on expertise tuning complex queries for low-latency analytical workloads.

  • Practical experience working with RAG Applications and Agentic AI Frameworks.

  • Soft Skills: Strong ownership mindset, proactive problem-solving ability, and a highly collaborative approach within nearshore team environments.

  • Language: Advanced English fluency (written and spoken) is strictly required.

Nice to Have

  • Experience with orchestration platforms such as Apache Airflow, Prefect, or Dagster.

  • Exposure to vector databases (e.g., pgvector, Pinecone, Qdrant) used in generative AI stacks.

  • Familiarity with streaming or event-driven technologies (e.g., Apache Kafka, AWS Kinesis).

  • Hands-on experience with cloud infrastructure as code (Terraform) or cloud-native CI/CD automation.

What You’ll Love

  • 100% Remote culture allowing you to work from anywhere in Latin America.

  • Holidays off aligned with national and company calendars.

  • Paid Time Off (PTO) to ensure work-life balance.

  • Health insurance assistance coverage.

  • Competitive USD compensation paid directly to you.

  • Growth opportunities within high-impact, modern tech projects and multi-client ecosystems.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Data Engineer Related jobs

Other jobs at CodeRoad Inc

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.