Data Engineer – Databricks
Aspenview Technology Partners · Buenos Aires, Autonomous City of Buenos Aires, Argentina, Bogotá, Bogota D.C., Colombia · Remote
Posted Aug 4, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
Build the Future with AspenView Technology Partners
At AspenView, we are passionate about transforming the way organizations approach technology. We specialize in creating high-performing, nearshore IT teams to help North American clients innovate faster and more efficiently.
As we continue to grow, we’re looking for exceptional people to join our team and help drive impactful change across industries.
Why Join AspenView?
At AspenView, we’re more than a nearshore IT partner—we’re a people-first, purpose-driven company that believes great culture drives great outcomes. We’re passionate about connecting talent and technology to deliver measurable value for clients—and meaningful career paths for our people.
Here’s what you can expect:
Competitive base
Comprehensive benefits and wellness support
Flexible work model: hybrid, remote, or in-office
Real growth opportunities and leadership visibility
Inclusive, respectful culture that blends U.S. innovation with Colombian heart
A company that listens, invests in you, and celebrates wins together
About the Role
We are seeking a skilled Data Engineer – Databricks to design, build, and optimize high-throughput, production-grade data pipelines on the Databricks Lakehouse platform. This is a full-time, remote position open to mid-senior and senior candidates across Latin America.
In this role, you will implement Medallion Architecture patterns (Bronze, Silver, Gold layers) to process structured and unstructured healthcare data—including Electronic Health Records (EHR/Epic), FHIR streams, claims data, and clinical research datasets. You will work closely with data architects, clinical analysts, and cloud engineers to build secure, HIPAA-compliant ETL/ELT pipelines using PySpark, Delta Lake, and Databricks Workflows.
What You Will Do
Lakehouse Pipeline Engineering: Design, develop, and maintain automated ETL/ELT data pipelines on Databricks utilizing PySpark, Delta Lake, and Databricks Workflows or Delta Live Tables (DLT).
Data Architecture Implementation: Implement Medallion Architecture principles to clean, transform, and aggregate raw clinical/operational datasets into analytics-ready models for BI reporting and predictive analytics.
Healthcare Data Integration: Ingest and normalize healthcare data sources, including Epic EHR databases (Clarity/Caboodle), FHIR/HL7 data feeds, and third-party claims systems.
Data Governance & Security: Enforce row/column-level access control and data lineage using Databricks Unity Catalog, ensuring strict compliance with HIPAA, HITRUST, and internal data governance protocols for Protected Health Information (PHI).
Performance Tuning & Optimization: Optimize compute cluster utilization, SQL query performance, Delta table indexing (Z-Ordering/Partitioning), and memory management to control cloud infrastructure costs and lower pipeline latency.
Cross-Functional Collaboration: Partner with US-based data scientists, business…