Haystack
← Back to Jobs
Remote
Technology
KA

Data Analyst / Programmer

KaltechsoftUnited States🇺🇸United StatesPosted Sep 16, 2026

Quick Overview

Seniority
Mid Senior
Work mode
Remote
Location
United States
Posted
19 hours ago
SQLAWSETLMachine LearningNLPNumPyScikit-learnSnowflakeAirflowApacheAzureDatabricksGoogle CloudPandasPower BIPythondbt

Job Description

Position: Data Analyst / Programmer
Location: Remote

 

Position Summary

The Data Analyst / Programmer supports Client’s clinical research operations by designing, developing, and maintaining data pipelines, analytical models, and reporting solutions that drive operational and study-level insights. This role serves as a key technical resource for transforming clinical trial data into meaningful, actionable intelligence in partnership with Data Management, Clinical Operations, and site-level stakeholders. The successful candidate will leverage Microsoft Fabric and Power BI as primary platforms, while applying programming expertise and an understanding of clinical research workflows to deliver high-quality, compliant data solutions.

 

______________

Key Responsibilities

1. Design, develop, and maintain ETL pipelines within Microsoft Fabric (Data Factory, Lakehouse, and Dataflows) to ingest, transform, and standardize clinical trial and patient-related data from multiple source systems including CTMS platforms, and site-level datasets.

2. Build and maintain interactive Power BI dashboards and reports that communicate performance metrics, site KPIs, enrollment trends, and data quality indicators to cross-functional stakeholders including Clinical Operations, Data Management, and executive leadership.

3. Develop and execute data validation, anomaly detection, and quality control routines to ensure the integrity and regulatory compliance of clinical data assets in accordance with Google Cloud Platform, 21 CFR Part 11, and ICH guidelines.

4. Write, optimize, and maintain SQL queries and Python scripts to support data extraction, feature engineering, and automated reporting workflows across structured and unstructured datasets.

5. Support the implementation and maintenance of semantic data models within Microsoft Fabric, ensuring consistent definitions, documentation, and governance of clinical data elements across studies.

6. Train and enable end users on Power BI dashboards and self-service reporting tools; document data dictionaries, standard operating procedures, and data governance guidelines to promote data literacy organization wide.

7. Participate in the evaluation, selection, and integration of data platforms and contribute to data migration, mapping, and reconciliation activities during system implementations or upgrades.

______________

Required Qualifications

 

•             Education: Bachelor’s degree in Computer Science, Information Systems, Biostatistics, Data Science, or a related field. Master’s degree preferred.

•             Experience: Minimum 3 years of experience in data analysis, data engineering, or clinical programming. Prior experience in clinical research, pharmaceutical, or CRO environment strongly preferred. Demonstrated ability to develop and maintain ETL pipelines, data models, and BI dashboards in a production setting.

•             Licenses/Certifications: Microsoft Power BI Data Analyst Associate certification required or must be obtained within 6 months of hire. Microsoft Fabric or Azure Data Engineer Associate certification preferred.

•             Technical Skills/Systems Experience: Required: Microsoft Fabric (Data Factory, Lakehouse, Dataflows Gen2), Power BI (report development, DAX, semantic modeling), SQL (intermediate to advanced), Python (pandas, NumPy, scikit-learn).

•             Other Requirements: Ability to work effectively in a remote or hybrid environment with occasional travel to clinical sites or corporate offices as needed. Strong written and verbal communication skills with the ability to present complex data findings to non-technical audiences. Ability to manage multiple concurrent projects and deliverables in a fast-paced clinical research environment.

______________

Preferred Qualifications (Optional)

 

•             Prior experience at a Contract Research Organization (CRO), clinical site network, or pharmaceutical/biotech company.

•             Hands-on experience with NLP, machine learning model development, or RAG-based analytics applied to clinical or operational data.

•             Familiarity with Apache Airflow, dbt, Snowflake, Databricks, or other modern data orchestration and warehousing tools.

•             Familiarity with clinical data systems such as CTMS, or eTMF solutions. Working knowledge of data governance principles and regulatory standards applicable to clinical trial data (21 CFR Part 11, ICH-Google Cloud Platform, CDISC SDTM/ADaM concepts).

•             Microsoft Azure Data Scientist Associate or AWS Cloud Practitioner certification

Similar jobs