Data Engineer
Vitol · Houston, TX, United States · On-site
Posted Sep 11, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
 
The role
We are looking for a Data Engineer to build and run the pipelines and data models behind our core data platform — the centralized data layer that serves global energy traders and analysts in real time.
This is a hands-on delivery role. You will write the Python that moves market and fundamental data from source through to the analysts, applications, and AI systems that consume it, and you will support what you ship.
Python is the core skill, but it is not the whole job. We are standing up a Snowflake warehouse, migrating legacy pipelines onto a Prefect-based framework, and consolidating how the platform stores and exposes data. The work reaches across more of the stack than pipeline code alone.
The platform is polyglot by design: streaming, caching, timeseries, relational, and warehouse technologies each doing what they are good at. You will build to the established patterns across that stack, and we expect you to understand why each technology sits where it does.
You will work alongside senior engineers who set those patterns, and directly with the analysts and desk users your work serves. The scope grows as you do — this is a role you can build a platform career from.
What You Will Do
Build, test, and operate production data pipelines in Python on our modern pipeline framework, orchestrated with Prefect
Build and maintain warehouse models in Snowflake using dbt — clean, tested, documented, and cost-aware
Migrate legacy pipelines and Oracle-based components onto current frameworks and standards, retiring technical debt as you go
Work across the platform stack — Kafka, Redis, InfluxDB, Oracle, and Snowflake — building to the established pattern for each rather than reaching for the tool you already know
Acquire data from external sources — vendor APIs, files, feeds, and web sources — and land it reliably
Implement data quality, freshness, and reconciliation checks so problems surface before users find them
Use AWS data services where our platform patterns call for them
Support what you ship — monitoring and alerting on your components, and investigating when something breaks
Contribute to making platform data AI-ready, and work with the catalog and steward teams so what you build is documented, classified, and findable
Engage directly with analysts and desk users to check that what you are building solves the actual problem
 
3+ years of hands-on data engineering experience building and operating production data pipelines
Strong Python — clean, tested, maintainable code, not scripts that happen to run. Fluency with pandas and the wider data-handling ecosystem
SQL depth — you can model, query, and tune, and you know what makes a query expensive before you run it
Snowflake or a comparable cloud warehouse — dimensional modeling, performance tuning, and cost-aware design. dbt experience is a strong plus
Breadth across data technologies — streaming, caching, timeseries, relational, and warehouse, with a…