Logo for CALYX

Data Engineer

Role overview

Qualifications

  • Bachelor's Degree in Computer Science, Data Engineering, Software Engineering, or related discipline
  • Experience in data pipeline orchestration (e.g., Airflow, Prefect, Dagster)
  • Experience with scalable data processing frameworks (Spark, Flink, Beam, Databricks)
  • Good experience in Python, SQL, and optionally 3GL technologies

Responsibilities

  • Build scalable, cloud-native data pipelines for batch and streaming workloads
  • Support the implementation of database and storage solutions across relational, NoSQL, and data lake systems
  • Implement data quality frameworks, validation rules, and automated checks
  • Develop high-throughput data processing solutions using distributed systems

Key facts

Hard skills

Other skills

  • Communication
  • Teamwork
  • Enthusiasm
  • Detail Oriented
  • Physical Flexibility

About the company

CALYX logo

CALYX

Contract Research Organizations (CRO)

Calyx and Invicro are now Perceptive! Expanding on our combined 50-year history, Perceptive provides best-in-class specialist support to global pharmaceutical, biotech, and clinical research organizations, spanning the complete R&D lifecycle, from discovery and preclinical through clinical development to post marketing. Perceptive’s offerings include imaging biomarkers and core lab services, as well as innovative technologies in randomization and trial supply management (RTSM), analytics, and software. Follow us at Perceptive Inc. and visit us at Perceptive.com

Company details

IndustryContract Research Organizations (CRO)
Company size1001 - 5000

Your match analysis

See how your profile stacks up against this role.

We compared the job requirements to your profile to show where you're strong and where you fall short.

Job description

We’re on a mission to change the future of clinical research. At Perceptive, we help the
biopharmaceutical industry bring medical treatments to the market, faster.
Our mission is to change the world but to do this, we need people like you.

As Data Engineer, Medical Imaging, you will support the design, development, and maintenance of scalable, secure, and compliant data solutions that enable analytics, reporting, and AI/ML workloads.

You will contribute to data ingestion, transformation, modelling, and orchestration activities, enabling reliable and high-quality data across the organisation with appropriate guidance.

In this role, you will work as part of a team building modern cloud data platforms, supporting data operations, and ensuring performance, reliability and compliance with industry standards.

Key Responsibilities

Cloud-Native Data Architecture & Pipeline Engineering

  • Build scalable, cloud-native data pipelines for batch and streaming workloads

  • Support the implementation of data ingestion frameworks for structured, semi-structured, and unstructured data

  • Develop APIs and integration services for data exchange between internal systems

  • Collaborating with Senior Engineers, support platform scalability and reliability across global operations

Data Modelling & Storage Design

  • Support the implementation of database and storage solutions across relational, NoSQL, and data lake systems.

  • Develop metadata-driven ingestion and transformation frameworks

  • Assist in ensuring data schemas support analytics, AI/ML, and business applications

Data Transformation, Quality & Governance

  • Implement data quality frameworks, validation rules, and automated checks

  • Build reusable transformation components to standardise data processing

  • Support the implementation of data lineage, cataloguing, and governance capabilities under guidance

  • Following established policies, ensure data privacy, protection, and compliance with regulatory requirements

Data Platform Development & Optimization

  • Develop high-throughput data processing solutions using distributed systems

  • Assist in optimising data pipelines for performance, cost efficiency, and resilience

  • Contribute to observability and monitoring activities for pipeline health, data drift, and SLA compliance

  • Build caching, partitioning, and indexing strategies to improve query performance

Analytics & AI/ML Enablement

  • Support the enablement of data scientists with curated datasets and feature pipelines

  • Develop real-time or batch-oriented feature stores

  • Integrate data workflows with ML operations (MLOps) and model deployment systems under guidance

  • Build visualisation-ready datasets for BI tools and dashboards

Security, Compliance & Risk Management

  • Apply established security practices to implement role-based access control, encryption, and secure data-sharing patterns

  • Support compliance with FDA/GxP, GDPR, HIPAA, and other applicable regulations

  • Develop audit trails, data retention, and disaster recovery solutions

Innovation

  • Stay informed on emerging data engineering, AI, and cloud technologies, developing technical capability and experience across data engineering domains

  • Share insights and drive continuous improvement across data engineering practices

Functional Competencies (Technical knowledge/Skills)

  • Working knowledge of cloud-native data platforms and serverless architectures (Azure, AWS)

  • Good understanding of CI/CD and DevOps concepts

  • Understanding of data modelling concepts

  • Familiarity with API patterns

  • Knowledge of streaming platforms (Kafka, Kinesis, Pub/Sub, Event Hubs)

  • Strong SQL skills

  • Good communication skills with ability to work across technical and business teams

  • Demonstrates willingness to learn and develop technical skills

  • A flexible attitude with respect to work assignments and new learning.

  • Ability to manage multiple and varied tasks with enthusiasm and prioritize workload with attention to detail.

  • Ability to identify and implement process improvements.

  • Proactively participates in skills improvement training and encourages their teams to participate.

Experience, Education and Certifications

  • Experience in data pipeline orchestration (e.g., Airflow, Prefect, Dagster).

  • Experience with scalable data processing frameworks (Spark, Flink, Beam, Databricks).

  • Experience working in data engineering or related technical roles.

  • Exposure to cloud-native data platforms (Azure, AWS, or GCP).

  • Exposure to modern ELT/ETL tools and distributed data processing.

  • Good experience in Python, SQL, and optionally 3GL technologies.

  • Exposure to modern data warehousing and lakehouse platforms.

  • Exposure to analytics tools (Power BI, Tableau, Looker) beneficial.

  • Bachelor's Degree in Computer Science, Data Engineering, Software Engineering, or related discipline.

  • English: Fluent.

Salary Range

  • This position attracts a salary of up to EUR 47,300 per year.

Apply once. Then go straight to the hiring manager.

After you apply, unlock the direct contact details of the people who actually make the call. A quick follow-up makes you 5x more likely to land an interview.

MR

Marcus Rivera

Chief Revenue Officer

m.rivera@company.com
linkedin.com/in/marcusrivera
Unlocked after you apply
·

Data Engineer Related jobs

Other jobs at CALYX

Premium

Reach out to the hiring manager directly.

Gain access to the contact details of the hiring managers who actually decide, and reach out to network with them directly. That, plus more when you upgrade:

  • Full match report with fit score and gaps
  • Career diagnostics on how recruiters read you
  • Curated company matches and warm intros
  • 48h early access to new roles

Cancel anytime.