Logo for SMX Services & Consulting, Inc.

Data Engineer

Role overview

Qualifications

  • Minimum of three years of professional experience in data engineering or a closely related discipline.
  • Demonstrated experience designing, building, and maintaining scalable ETL or ELT pipelines.
  • Strong SQL and Python skills or equivalent data-engineering capabilities.
  • Experience ingesting and transforming flat files, JSON, XML, Excel, APIs, relational data, and graph data.

Responsibilities

  • Design, build, test, deploy, and maintain scalable ETL and ELT pipelines.
  • Ingest and transform data from flat files, JSON, XML, Excel, APIs, relational databases, graph databases, streaming sources, and other formats.
  • Develop reusable SQL and Python processes for data ingestion, validation, standardization, transformation, enrichment, and loading.
  • Implement data-quality checks, completeness checks, validity rules, duplicate detection, and reconciliation controls.

About the company

SMX Services & Consulting, Inc. logo

SMX Services & Consulting, Inc.

IT Services & IT Consulting

SMX Services & Consulting is an information technology outsourcing (ITO) provider with more than 20 years’ applied experience providing logical solutions to emerging enterprises, in a variety of industry verticals including technology, finance, banking, real estate, insurance, and retail. Based in Miami, Florida, SMX serves private, public and institutional clients, including some Fortune 500 companies, in more than 10 countries from regional offices in Houston, San Juan, Bogotá, and Caracas.

Company details

IndustryIT Services & IT Consulting
Company size201 - 500

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Data Engineer


Hourly pay range: $44.93–$46.39
Employment type: Full-time
Schedule: Approximately 1,920 hours annually
Location: Primarily remote, with occasional onsite work in Washington, DC

Position Summary

The Data Engineer will design, develop, test, operate, document, and optimize scalable data pipelines supporting advanced fraud analytics and investigative activities.

The position will work with structured, semi-structured, and unstructured data from government, public, commercial, and other authorized sources. The Data Engineer will be responsible for reliable data ingestion, transformation, validation, lineage, governance, security, and availability within modern data-platform and Lakehouse environments.

Primary Responsibilities

  • Design, build, test, deploy, and maintain scalable ETL and ELT pipelines.
  • Ingest and transform data from flat files, JSON, XML, Excel, APIs, relational databases, graph databases, streaming sources, and other formats.
  • Develop reusable SQL and Python processes for data ingestion, validation, standardization, transformation, enrichment, and loading.
  • Load, manage, and optimize data within Databricks Unity Catalog, Microsoft SQL Server managed instances, and comparable platforms.
  • Support streaming and batch ingestion frameworks within a modern Lakehouse architecture.
  • Conduct source-system profiling and document source structures, data definitions, relationships, limitations, and quality issues.
  • Develop data mappings, transformation logic, reconciliation procedures, and exception-handling processes.
  • Implement data-quality checks, completeness checks, validity rules, duplicate detection, and reconciliation controls.
  • Develop logging, monitoring, alerting, retry, recovery, and error-handling procedures.
  • Maintain data lineage, metadata, data dictionaries, schema documentation, pipeline documentation, and operational runbooks.
  • Support entity resolution, record linkage, data matching, graph ingestion, and fraud-model feature development.
  • Optimize SQL queries, transformation processes, storage structures, and pipeline performance.
  • Implement enterprise data-management, data-governance, and data-quality standards.
  • Support role-based access, data segregation, auditability, and approved information-handling requirements.
  • Work with data owners and government stakeholders to resolve access, quality, interpretation, and integration issues.
  • Collaborate with data scientists, graph data scientists, investigative analysts, forensic accountants, and business analysts.
  • Support transition and continued operation of existing data pipelines without service interruption.
  • Maintain code and technical artifacts within government-approved repositories and version-control environments.
  • Support production releases, configuration management, testing, and operational maintenance.

Required Qualifications

  • Minimum of three years of professional experience in data engineering or a closely related discipline.
  • Demonstrated experience designing, building, and maintaining scalable ETL or ELT pipelines.
  • Strong SQL and Python skills or equivalent data-engineering capabilities.
  • Experience ingesting and transforming flat files, JSON, XML, Excel, APIs, relational data, and graph data.
  • Experience with Databricks Unity Catalog, Microsoft SQL Server managed instances, or comparable platforms.
  • Experience with streaming and batch data-ingestion frameworks.
  • Experience with modern Lakehouse, cloud-data, or distributed-data architecture.
  • Experience implementing data-quality, lineage, reliability, monitoring, and performance controls.
  • Familiarity with enterprise data management, metadata, data governance, and data-quality practices.
  • Experience documenting schemas, transformation logic, pipelines, data mappings, and operational procedures.
  • Ability to troubleshoot data, pipeline, performance, and production issues.
  • Ability to collaborate with technical, investigative, and business stakeholders.
  • Strong analytical, documentation, communication, and problem-solving skills.
  • Ability to complete federal suitability, HSPD-12/PIV credentialing, and system-access requirements.

Preferred Qualifications

  • Experience supporting fraud detection, anomaly detection, financial oversight, investigative analytics, or government-benefit programs.
  • Experience with Azure Databricks, Apache Spark, Delta Lake, Unity Catalog, Microsoft SQL Server, Neo4j, Power BI, or comparable technologies.
  • Experience with orchestration, automated testing, CI/CD, Git-based repositories, and production monitoring.
  • Experience supporting entity resolution, graph-data ingestion, machine-learning feature pipelines, or investigative analytics.
  • Degree in computer science, data engineering, information systems, software engineering, or a related discipline.

 

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Data Engineer Related jobs

Other jobs at SMX Services & Consulting, Inc.

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.