Back to jobs

Data Engineer

Austin, Texas, United States

Pay
Salary not listed in the saved posting
Work setup
Unconfirmed
Employment
Unconfirmed
Apply at Teza-Technologies

What you’ll work on

Full posting

Teza's Data Platform team owns the data the firm trades on: every backtest, every live strategy, every portfolio decision starts with data we ingested, cleaned, stored, and served.

1 PB of raw historical vendor data, growing by ~150 GB every day

  • Design and onboard new data sources into our warehouse; improve the robustness, speed, and scalability of our systems; manage data entitlements.

  • Build automated systems for data cleansing, anomaly detection, monitoring, and alerting, bad data must never reach a strategy.

  • Evaluate new tools and technologies for organizing, querying, and streaming large datasets, and when nothing on the market fits, build it.

From the employer’s posting
Teza's Data Platform team owns the data the firm trades on: every backtest, every live strategy, every portfolio decision starts with data we ingested, cleaned, stored, and served.
1 PB of raw historical vendor data, growing by ~150 GB every day
Work directly with Portfolio Managers and Quantitative Developers: turn their requirements into datasets and pipelines, and be the person who knows every nuance of the data they trade on. Design and onboard new data sources into our warehouse; improve the robustness, speed, and scalability of our systems; manage data entitlements. Build automated systems for data cleansing, anomaly detection, monitoring, and alerting, bad data must never reach a strategy.
Design and onboard new data sources into our warehouse; improve the robustness, speed, and scalability of our systems; manage data entitlements. Build automated systems for data cleansing, anomaly detection, monitoring, and alerting, bad data must never reach a strategy. Evaluate new tools and technologies for organizing, querying, and streaming large datasets, and when nothing on the market fits, build it. That's how the bitemporal store happened.
Build automated systems for data cleansing, anomaly detection, monitoring, and alerting, bad data must never reach a strategy. Evaluate new tools and technologies for organizing, querying, and streaming large datasets, and when nothing on the market fits, build it. That's how the bitemporal store happened. Support the production data warehouse the firm depends on.

What you’ll bring

All qualifications

Core experience

  • Proficiency in Python and Unix/Linux for data manipulation, scripting, and automation.
  • Experience with on-premises data infrastructure.
  • Strong SQL, including query optimization and performance tuning, and familiarity with NoSQL.
  • Familiarity with a cloud platform (AWS or GCP).
Qualification wording
Proficiency in Python and Unix/Linux for data manipulation, scripting, and automation.
Experience with on-premises data infrastructure.
Strong SQL, including query optimization and performance tuning, and familiarity with NoSQL.
Familiarity with a cloud platform (AWS or GCP).

Tools in this posting

  • Java
  • Python
  • SQL
  • AWS
  • MongoDB
  • PostgreSQL
  • S3
  • Airflow
  • Google Cloud (GCP)
  • NoSQL
Source — Tool mentions in context
Our Stack Python and Java · Apache Airflow · Slurm · NATS · PostgreSQL · MongoDB · S3 · NFS · GitHub Actions Basic Requirements
- Financial industry experience or internships. - Java (part of our platform is written in it). - Experience with on-premises data infrastructure.
- A bitemporal data store, designed and written in-house from scratch. Every dataset answers both "what did we know then?" and "what do we know now?", which is what lets researchers trust a backtest. - A Python 3.14 migration of a large, long-lived codebase. - Adoption of the latest Apache Airflow: writing DAGs for the thousands of jobs moving off cron.
Basic Requirements - Proficiency in Python and Unix/Linux for data manipulation, scripting, and automation. - Strong SQL, including query optimization and performance tuning, and familiarity with NoSQL.
- Proficiency in Python and Unix/Linux for data manipulation, scripting, and automation. - Strong SQL, including query optimization and performance tuning, and familiarity with NoSQL. - A solid grasp of data modeling: normalization and denormalization, and the judgment to know when each applies.
- Experience with on-premises data infrastructure. - Familiarity with a cloud platform (AWS or GCP). - Apache Airflow or similar workflow orchestration tools.
- 120 TB of processed, query-ready data in historical storage - Thousands of scheduled jobs: run by cron today, actively migrating to Apache Airflow - Alternative data delivered directly into the real-time feeds of live trading strategies
- A Python 3.14 migration of a large, long-lived codebase. - Adoption of the latest Apache Airflow: writing DAGs for the thousands of jobs moving off cron. - Pipelines for market data and alternative data: everything from exchange feeds to weather.
- Familiarity with a cloud platform (AWS or GCP). - Apache Airflow or similar workflow orchestration tools. Benefits

Benefits in the posting

Full benefits wording
  • Health, visual and dental insurance
  • Flexible sick time policy

From the employer’s posting.

Job description

View original posting ↗

About the role

Teza's Data Platform team owns the data the firm trades on: every backtest, every live strategy, every portfolio decision starts with data we ingested, cleaned, stored, and served.

The scale, in plain numbers:

  • 1 PB of raw historical vendor data, growing by ~150 GB every day

  • 120 TB of processed, query-ready data in historical storage

  • Thousands of scheduled jobs: run by cron today, actively migrating to Apache Airflow

  • Alternative data delivered directly into the real-time feeds of live trading strategies

This is a hands-on position on a small team of data engineers with growth potential. The firm is looking for outstanding technical skills, strong attention to detail, and a desire to architect and build data platforms.

Location
Austin, TX / Yerevan, Armenia (in-office requirement)

Key Responsibilities

  • Work directly with Portfolio Managers and Quantitative Developers: turn their requirements into datasets and pipelines, and be the person who knows every nuance of the data they trade on.

  • Design and onboard new data sources into our warehouse; improve the robustness, speed, and scalability of our systems; manage data entitlements.

  • Build automated systems for data cleansing, anomaly detection, monitoring, and alerting, bad data must never reach a strategy.

  • Evaluate new tools and technologies for organizing, querying, and streaming large datasets, and when nothing on the market fits, build it. That's how the bitemporal store happened.

  • Support the production data warehouse the firm depends on.

  • Develop and maintain vendor relationships aligned with our business objectives.

What we're building right now

  • A bitemporal data store, designed and written in-house from scratch. Every dataset answers both "what did we know then?" and "what do we know now?", which is what lets researchers trust a backtest.

  • A Python 3.14 migration of a large, long-lived codebase.

  • Adoption of the latest Apache Airflow: writing DAGs for the thousands of jobs moving off cron.

  • Pipelines for market data and alternative data: everything from exchange feeds to weather.

  • Real-time delivery: alternative data flows straight into strategies' live feeds. Pipelines you build sit in the trading path.

  • CI/CD for all of it, in GitHub Actions.

Our Stack

Python and Java · Apache Airflow · Slurm · NATS · PostgreSQL · MongoDB · S3 · NFS · GitHub Actions

 

Basic Requirements

  • Proficiency in Python and Unix/Linux for data manipulation, scripting, and automation.

  • Strong SQL, including query optimization and performance tuning, and familiarity with NoSQL.

  • A solid grasp of data modeling: normalization and denormalization, and the judgment to know when each applies.

Nice to have

  • Financial industry experience or internships.

  • Java (part of our platform is written in it).

  • Experience with on-premises data infrastructure.

  • Familiarity with a cloud platform (AWS or GCP).

  • Apache Airflow or similar workflow orchestration tools.


Benefits

  • Health, visual and dental insurance

  • Flexible sick time policy

Your next step

  • Have your CV and examples of relevant work ready.
  • Check the listed location, eligibility and core experience before starting.
  • Ask the employer about the salary range before committing time to the process.

Complete your application on jobs.ashbyhq.com. The employer’s form will show what is required.

Already applied? Track this application

Source & posting history

View original posting ↗

Source notes

Source excerpts

Selected passages from the saved posting. Check the full description for conditions and exceptions.

Pay

No pay amount identified in the saved description.

Location & working pattern

Austin, Texas, United States

Working pattern and location restrictions need checking in the full posting.

Work authorization

No clear work-authorization passage found. Eligibility is unconfirmed.

Status in our records
Active
First seen by us
Jul 8, 2026
Recorded sightings
28
Last seen by us
Oct 7, 2026

These dates show when we found the listing. Check the employer’s website to confirm it is still accepting applications.

Report an error

See how this role fits your experience

Add your resume to compare the role’s scope, tools and requirements with your experience.

Find answers in the posting

AI
How answers work

AI selects complete passages from this posting. Check them for conditions and exceptions.

Uses this posting and your question. No profile needed.