PANTHERx Rare Logo

PANTHERx Rare

Senior Data Engineer

Posted An Hour Ago
Remote
Hiring Remotely in United States
Senior level
Remote
Hiring Remotely in United States
Senior level
Builds, operates, and owns enterprise data warehouse pipelines and models on Databricks. Develops Delta Lake batch and streaming pipelines across Bronze, Silver, and Gold layers; implements data quality, reconciliation, observability, governance, and performance improvements. The role owns assigned data domains, supports event-driven ingestion, documents architecture and runbooks, responds to incidents, collaborates with QA and governance teams, and mentors data engineers.
The summary above was generated by AI

7,000 Diseases - 500 Treatments - 1 Rare Pharmacy

PANTHERx is the nation’s largest rare disease pharmacy, and we put the patient experience at the top of everything that we do.

If you are looking for a career in the healthcare field that embraces authentic dedication to patient care, you don’t need to look beyond PANTHERx. In every line of service, in every position and area of expertise, PANTHERx associates are driven to provide the highest quality outcomes for our patients.

We are seeking team members who:

Are inspired and compassionate problem solvers;

Produce high quality work;

Thrive in the excitement of the ever-challenging environment of modern medicine; and

Are committed to achieving superior health outcomes for people living with rare and devastating diseases.

At PANTHERx, we know our employees are the driving force in what we do. We cultivate talent and encourage growth within PANTHERx so that our associates can continue to explore their interests and expand their careers. Guided by our mission to provide uncompromising quality every day, we continue our strategic growth to further reach those affected by rare diseases.

Join the PANTHERx team, and define your own RxARE future in healthcare!

Location: Pittsburgh, PA (Hybrid)

Classification: Exempt

Status: Full-Time

Reports to: Manager, Data Engineering

Purpose

The Senior Data Engineer builds, operates, and owns the enterprise data warehouse and core data pipelines on PANTHERx’s Databricks platform. This role is responsible for EDW architecture, ingress patterns into Bronze, transformation logic from Bronze to Silver, and Gold harmonization rules. The Senior Data Engineer significantly contributes to the platform’s evolution to a simplified medallion architecture with event-driven, low-latency ingestion. This role serves as a senior practitioner within the Data Engineering track, setting the technical bar for pipeline quality, documentation discipline, and reliability, as well as mentoring Data Engineers as the team scales.

Responsibilities

EDW Engineering & Ownership

  • Designs, builds, and operates enterprise data warehouse pipelines and data models across the Databricks medallion architecture (Bronze, Silver, Gold), including Silver conformance logic and Gold harmonization and survivorship rules.
  • Assumes named ownership of assigned EDW domains as production systems: data model integrity, pipeline reliability, ingestion SLAs, and incident response.
  • Leads structured documentation of knowledge for assigned domains, absorbing architecture, transformation logic, and feed SQL.
  • Executes data model simplification work, consolidating legacy transformation layers into the target architecture with validated output parity.

Pipeline Development & Streaming

  • Develops and maintains production Delta Lake pipelines, evolving ingestion from batch watermark patterns to event-driven architectures using Zerobus Ingestion and Spark Declarative Pipelines for high-SLA source systems.
  • Implements row-level reconciliation, validation checkpoints, and quality gates at ingestion, transformation, and delivery layers, operationalizing governance-defined data quality dimensions in partnership with QA.
  • Builds Gold-layer tables and promotion logic that support Unity Catalog metric views, lineage capture, and access controls in partnership with Analytics Engineering and Data Governance.
  • Optimizes pipeline performance, latency, and compute cost across all layers of the platform.

Documentation, Quality & Observability

  • Documents architecture decisions, data models, pipeline logic, and operational runbooks to team standards, ensuring no critical platform capability depends on a single point of failure.
  • Operates pipeline observability tooling: monitors anomaly detection, triages data reliability incidents, and drives root cause remediation for assigned domains.
  • Implements data contract validation (ODCS or equivalent) in pipelines supporting external partner feed domains, ensuring contract failures halt delivery before reaching consumers.

Mentorship & Cross-Functional Collaboration

  • Mentors Data Engineers on Databricks development patterns, SQL and PySpark craft, and documentation discipline; reviews pull requests in Azure DevOps & Git to maintain code quality.
  • Partners with Data Governance on lineage capture and metadata standards, with QA on Tier 1 and Tier 2 pipeline test suite development, and with Partner Data Services on feed engineering requiring pipeline work.
  • Works from structured requirements and acceptance criteria entering through the Informatics intake process, and flags requirements gaps before build work begins.

Required Qualifications

  • 6+ years of progressive data engineering experience, including production ownership of an enterprise data warehouse or large-scale transformation pipelines with defined SLAs.
  • Deep, hands-on expertise with Databricks in production environments: Delta Lake, medallion architecture, Unity Catalog, and pipeline performance optimization.
  • Demonstrated data warehousing depth: dimensional and harmonized data modeling, conformance and survivorship logic, and operating a warehouse as a production system.
  • Advanced SQL and strong PySpark and Python proficiency, sufficient to build, review, and optimize complex transformation logic independently.
  • Proven experience with Azure data services: Azure Databricks, Azure Data Lake Storage, Azure Data Factory, and CI/CD practices in Azure DevOps.
  • Demonstrated ability to absorb complex, under-documented systems through structured knowledge transfer and produce documentation that makes that knowledge durable and transferable.
  • Strong communication and collaboration skills; able to work directly with QA, Governance, Informatics, and business-facing teams without an intermediary.

Preferred Qualifications

  • Healthcare, specialty pharmacy, or regulated industry experience, with exposure to clinical or operational data environments.
  • Production experience with event-driven ingestion: Azure Event Hubs, Kafka, or equivalent, consumed via Structured Streaming.
  • Familiarity with data contract standards (ODCS or equivalent) and pipeline observability platforms (Monte Carlo or similar).
  • Familiarity with enterprise data catalog tooling (Atlan, Collibra, or equivalent) and designing pipelines as catalogued, governed assets.
  • Experience leading through influence: mentoring engineers, setting standards, and driving adoption without formal management authority.
  • Bachelor’s degree in Computer Science, Data Engineering, Information Systems, or a related field, or equivalent experience.

Work Environment

This position works in a home office and professional office environment. When in-office this role routinely uses standard office equipment such as computers, phones, photocopiers, filing cabinets and fax machines, and communications via MS Teams.

Physical Demands

While performing the duties of this job, the employee is regularly required to sit, see, talk or hear. The employee frequently is required to stand; walk; use hands and fingers to handle or feel; and reach with hands and arms. Visual acuity is necessary for tasks such as reading and working with various forms of data on a screen. Reasonable accommodation may be made to enable individuals with disabilities to perform

Benefits:

Hybrid, remote and flexible on-site work schedules are available, based on the position. PANTHERx Rare Pharmacy also affords an excellent benefit package, including but not limited to medical, dental, vision, health savings and flexible spending accounts, 401K with employer matching, employer-paid life insurance and short/long term disability coverage, and an Employee Assistance Program! Generous paid time off is also available to all full-time employees. Of course we offer paid holidays too!

Equal Opportunity:

PANTHERx Rare Pharmacy is an equal opportunity employer, and does not discriminate in recruiting, hiring, promotions or any term or condition of employment based on race, age, religion, gender, ethnicity, sexual orientation, gender identity, disability, protected veteran's status, or any other characteristic protected by federal, state or local laws.

Similar Jobs

4 Days Ago
Easy Apply
Remote
United States
Easy Apply
140K-175K Annually
Senior level
140K-175K Annually
Senior level
Fintech • Insurance • Machine Learning • Analytics • Financial Services • Automation
Build and maintain production data pipelines, Airflow DAGs, Snowflake infrastructure, and Data Vault 2.0 models. Ensure data quality, observability, security, governance, reliability, and cost efficiency through automated testing and CI/CD. Lead technical projects end-to-end, collaborate with stakeholders, and support SOX-relevant reporting while developing expertise in commercial insurance data and processes.
Top Skills: Apache AirflowBigQueryCi/CdClaude CodeCursorData Vault 2.0PythonRedshiftSnowflakeSnowflake CortexSQL
7 Days Ago
Remote or Hybrid
3 Locations
110K-205K Annually
Senior level
110K-205K Annually
Senior level
Fintech • Financial Services
Designs and develops complex cloud-native data solutions, scalable batch and real-time pipelines, APIs, data models, and enterprise data platforms. Implements data governance, quality, lineage, security, observability, CI/CD, automated testing, and operational reporting capabilities. Collaborates with architects, product owners, analysts, governance, security, and platform teams on technology strategy and cloud modernization. Mentors engineers, leads technical design discussions, presents recommendations, and supports AI-ready data products and emerging development tools.
Top Skills: Ai FoundryAlationApache AirflowApi ManagementAWSAzureAzure Ai SearchAzure Container AppsAzure Data FactoryAzure DatabricksAzure DevopsAzure FunctionsAzure OpenaiAzure PurviewAzure SqlAzure StorageBigQueryCi/CdCollibraCopilot StudioCosmos DbEltErwinETLEvent HubsGCPGitGit FlowGitGithub CopilotInformatica MdmInfosphereJavaMcpMicrosoft FabricPythonScalaSnowflakeSQLTrunk-Based Development
5 Days Ago
Remote or Hybrid
92K-164K Annually
Senior level
92K-164K Annually
Senior level
Artificial Intelligence • Big Data • Healthtech • Information Technology • Machine Learning • Software • Analytics
Develop and automate tests for ETL pipelines, validating data records, content, metadata, and schema alignment. Integrate data tests into CI/CD workflows using GitHub Actions, and create release test suites to monitor performance and memory degradation. Analyze edge cases, investigate pipeline issues, and log defects while working with Databricks, Python, SQL, and ETL orchestration tools.
Top Skills: AirflowAzure Data FactoryCi/CdDatabricksETLGithub ActionsPythonSQL

What you need to know about the Charlotte Tech Scene

Ranked among the hottest tech cities in 2024 by CompTIA, Charlotte is quickly cementing its place as a major U.S. tech hub. Home to more than 90,000 tech workers, the city’s ecosystem is primed for continued growth, fueled by billions in annual funding from heavyweights like Microsoft and RevTech Labs, which has created thousands of fintech jobs and made the city a go-to for tech pros looking for their next big opportunity.

Key Facts About Charlotte Tech

  • Number of Tech Workers: 90,859; 6.5% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Lowe’s, Bank of America, TIAA, Microsoft, Honeywell
  • Key Industries: Fintech, artificial intelligence, cybersecurity, cloud computing, e-commerce
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (CED)
  • Notable Investors: Microsoft, Google, Falfurrias Management Partners, RevTech Labs Foundation
  • Research Centers and Universities: University of North Carolina at Charlotte, Northeastern University, North Carolina Research Campus

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account