Logo for Aptus Data Labs

Sr Data Engineer- Databricks

Role overview

Qualifications

  • 5+ years of experience in data engineering roles
  • Proven expertise in Databricks, Delta Lake, and Apache Spark (PySpark preferred)
  • Deep understanding of Unity Catalog for fine-grained data governance and lineage tracking
  • Proficiency in SQL for large-scale data manipulation and analysis

Responsibilities

  • Design, implement, and optimize scalable data pipelines using Databricks and Apache Spark
  • Architect data lakes using Delta Lake, ensuring reliable and efficient data storage
  • Manage metadata, security, and lineage through Unity Catalog for governance and compliance
  • Collaborate with ML engineers and data scientists on LLM-based AI/GenAI project pipelines

Key facts

Hard skills

Other skills

  • Problem Solving
  • Communication
  • Collaboration

About the company

Aptus Data Labs logo

Aptus Data Labs

Artificial Intelligence & Machine Learning Services

Aptus Data Labs is a leading provider of data engineering and AI solutions, helping enterprises turn complex data into actionable decisions and measurable business outcomes. We partner with organizations across Pharma, Manufacturing & Supply Chain, Banking & FinTech, and Technology to enable faster, smarter decision-making in data-intensive and regulated environments. Our solutions span: > Modern data engineering and AI architectures > Agentic AI systems that automate reasoning, orchestration, and decision workflows > Enterprise-grade AI chatbots that deliver secure, contextual, and role-based insights across business functions > Data Science and Decision Science grounded in domain-specific KPIs > Proprietary Aptus Accelerators that shorten time-to-value and de-risk adoption We focus on decision intelligence, applying descriptive, predictive, and prescriptive analytics to solve high-impact use cases such as forecasting, risk identification, operational optimization, and executive decision support. Our multidisciplinary team of Data Engineers, AI Engineers, Data Scientists, Architects, and Business SMEs works closely with clients as long-term partners, building AI solutions that are practical, scalable, explainable, and ready for enterprise adoption. At Aptus Data Labs, we believe AI should augment human judgment—not replace it. We help organizations move from insight to impact with confidence.

Company details

IndustryArtificial Intelligence & Machine Learning Services
Company size201 - 500

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

Exp- 5+ yrs

Location- Remote (Preferred candidates to be in bangalore)

Notice- Looking candidates with to be joining within 30 Days

Key Responsibilities:

  • Design, implement, and optimize scalable data pipelines using Databricks and Apache Spark.

  • Architect data lakes using Delta Lake, ensuring reliable and efficient data storage.

  • Manage metadata, security, and lineage through Unity Catalog for governance and compliance.

  • Ingest and process streaming data using Apache Kafka and real-time frameworks.

  • Collaborate with ML engineers and data scientists on LLM-based AI/GenAI project pipelines.

  • Apply CI/CD and DevOps practices to automate data workflows and deployments (e.g., with GitHub Actions, Jenkins, Terraform).

  • Optimize query performance and data transformations using advanced SQL.

  • Implement and uphold data governance, quality, and access control policies.

  • Support production data pipelines and respond to issues and performance bottlenecks.

  • Contribute to architectural decisions around data strategy and platform scalability.



Requirements

Required Skills & Experience:

  • 5+ years of experience in data engineering roles.
  • Proven expertise in Databricks, Delta Lake, and Apache Spark (PySpark preferred).
  • Deep understanding of Unity Catalog for fine-grained data governance and lineage tracking.
  • Proficiency in SQL for large-scale data manipulation and analysis.
  • Hands-on experience with Kafka for real-time data streaming.
  • Solid understanding of CI/CD, infrastructure automation, and DevOps principles.
  • Experience contributing to or supporting Generative AI / LLM projects with structured or unstructured data.
  • Familiarity with cloud platforms (AWS, Azure, or GCP) and data services.
  • Strong problem-solving, debugging, and system design skills.
  • Excellent communication and collaboration abilities in cross-functional teams.



Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Data Engineer Related jobs

Other jobs at Aptus Data Labs

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.