Hiring.Camp

Senior Data Platform Engineer

Aperia

·

Today

Location
Dallas, Texas, United States
Department
Programming
Seniority
Senior
Experience
2+ years
Source
Greenhouse

Description

About Aperia Solutions 

Aperia Solutions is a leading payments SaaS company headquartered in Dallas, Texas. We build risk management, residuals accounting, compliance, and analytics platforms used by merchant banks, ISOs, payment processors, and other participants in the credit card payments ecosystem. Operating under PCI DSS 4.0 compliance across dual on-premises datacenters (Allen, TX and Chicago, IL) plus Azure cloud, we process high-volume payment data for some of the industry’s largest players. 

We’re in the middle of a significant platform modernization program: a next-generation risk analysis and scoring platform on Microsoft Fabric, an AI-powered ingestion platform, new AI-native applications, and ongoing evolution of the products our clients rely on every day. The teams executing this work are deliberately small and senior, with executive-level support and minimal bureaucracy. Every senior hire is a founding contributor to the architecture and the code. 

A

About the Role 

Aperia is building a next-generation risk analysis and scoring platform on Microsoft Fabric, and this role owns its medallion data architecture end-to-end. You'll design and build Bronze / Silver / Gold Delta tables, Fabric Spark feature pipelines, Direct Lake semantic models, and physical-layout optimizations. This is our highest-utilization engineering role on the project — central to both Phase 1 platform foundation and Phase 2 risk product delivery, working alongside a Senior Technical Lead, Senior DevOps engineer, and mid-level developer. 

The platform ingests transaction and merchant data from clients via SFTP and API, tokenizes sensitive PAN data at the ingestion boundary using our reserved-BIN HMAC-derived format-preserving scheme, and lands data in Aperia's canonical Silver on Microsoft Fabric per our internal canonical spec. You'll build the pipeline that populates canonical Silver via MERGE, consume it via OneLake shortcuts, and materialize risk-specific Gold tables that feed both custom-authored DMN rules and pre-scored Mastercard Brighterion feeds. 

What You'll Do 

  • Design and implement Bronze, canonical Silver, and Gold Delta table schemas with PAN-safe physical layout — statistics allowlist configuration, liquid clustering, deliberate retention settings, V-Order optimization. 
  • Build Fabric Spark notebooks for the feature pipeline (rolling velocity features across 24h / 7d / 30d windows, merchant-level rolling features, hierarchical rollups, cross-merchant features via pan_token joins). 
  • Author per-flavor-version Silver adapters — one function per flavor, unioned by name, tested against golden fixtures. 
  • Build the canonical MERGE DAG (Bronze → canonical Silver, idempotent on _natural_key per Aperia's canonical Fabric internal spec v1.4). 
  • Build the Gold materialization pipeline for both grains (fact_transaction_risk, fact_merchant_risk) with correct score-source stamping and idempotent merge semantics. 
  • Design and implement Direct Lake semantic models per reporting surface; performance-test against representative data volumes. 
  • Configure OneLake namespace layouts, workspace catalogs, and cross-workspace shortcut relationships with the residuals programme. 
  • Establish idempotent Delta write patterns using delta-rs from Python containers (segmented parallel writes with single atomic commit). 
  • Fabric capacity sizing (F-SKU selection, load testing, Reservation vs PAYG decisions). 
  • Partner with the DevOps engineer on Fabric IaC (Terraform capacity provisioning plus Fabric REST workspace bootstrap). 
  • Consult on ClickHouse contingency if Direct Lake latency proves insufficient for client-tier UIs. 

What You Bring 

  • 5+ years data engineering experience, with at least 2 years hands-on Spark development (PySpark preferred). 
  • Deep Delta Lake internals knowledge — transaction log format, checkpoints, retention settings (logRetentionDuration, deletedFileRetentionDuration), VACUUM, OPTIMIZE, deletion vectors, statistics collection. 
  • Parquet format expertise — footers, statistics allowlists, row-group sizing, dictionary encoding, column indexes. 
  • Microsoft Fabric hands-on experience — workspaces, capacities, notebooks, lakehouses, semantic models, Direct Lake, SQL analytics endpoint. 
  • SQL performance tuning skills for both OLAP (Fabric) and OLTP (PostgreSQL) workloads. 
  • Strong Python data engineering skills — pandas, PyArrow, delta-rs. 
  • Ability to communicate architecture clearly in writing and to work autonomously on well-scoped features. 

Nice to Have 

  • PCI-scoped data platform experience. 
  • Financial services or payments domain background. 
  • Power BI semantic model authoring (DAX). 
  • Familiarity with published data-standard governance patterns — logical model vs physical profile, SCD2 dimensions, natural-key idempotency, extension attribute stores. 
  • OneLake shortcuts and cross-workspace lakehouse patterns. 
  • ClickHouse or similar OLAP engine familiarity. 
  • Terraform for Azure data resources. 
  • Prior experience with fluid or schema-flexible ingestion patterns. 

Why Aperia 

  • Founding-member impact on a strategic platform serving a multi-hundred-million-dollar industry — your architectural decisions carry weight and ship to production. 
  • Direct working relationship with the SVP of Technology; small team; minimal bureaucracy; senior peers. 
  • Genuine engineering culture — technical correctness valued over slideware; extensive design documentation; real design reviews; ADR discipline. 
  • Explicit team commitment to AI-assisted development (Claude Code, Cursor, GitHub Copilot) — treated as a first-class productivity multiplier, not a novelty. Team standards documented and enforced. 
  • Modern tech stack: Microsoft Fabric, AKS, Azure, Airflow, Python, Delta Lake, Kogito/DMN, OpenBao, Kubernetes. 

Eligibility Requirements

  • Must be willing to submit to a background investigation and drug test as part of the selection process.

Job Type

  • Full-time

Schedule

  • Hybrid / Monday to Friday 

Work Location

  • Dallas, TX

Benefits

  • Health insurance 
  • Health savings account
  • Dental insurance
  • Vision insurance
  • 401(k) matching
  • Life insurance
  • Paid time off
  • Parental leave
  • Disability insurance
  • Childcare assistance
  • Education reimbursement
  • Fitness membership
  • Volunteer time off

 

Skills

PythonBootstrapAzureKubernetesTerraformSQLPostgreSQLSparkAirflowData EngineeringGitHubPower BIRESTDevOpsRisk ManagementCompliance

Similar Jobs

30

Senior DevOps Engineer: Data & AI Platform

Assent · Ottawa, ON, Canada · Hybrid

Today

Senior Staff Data Platform Engineer - Kafka - Apache Iceberg - Apache Spark

ServiceNow · San Diego, CALIFORNIA, United States · Hybrid

Yesterday

Sr Staff, Data Platform Architect

Danaher · POL – Krakow – Cytiva, Poland · Onsite

Yesterday

Senior/Staff Software Engineer, Data Platform

Axion · San Francisco, CA +1 · Hybrid

Yesterday

Senior Data Platform Engineer & Cloud Architect

Expleo Es En · Barcelona, CT, ES

Yesterday

Senior Data Platform Engineer

Kenvue · IN025 Embassy Manyata Business Park, India · Hybrid

Yesterday

Sr. Software Engineer, Data Platform

Coherehealth · United States

Yesterday

Sr. Software Engineer, Data Platform (Starlink)

Spacex · Redmond, WA +2

Yesterday

Senior Analytics Engineer, Data Platform

The New York Times · New York, NY

2 days ago

Senior Data Platform Engineer

Bondora · Tallinn, Harju, Estonia +1

2 days ago

Engineering - Data Platform - Senior Software Engineer

LightBox · United States

2 days ago

Engineering - Data Platform - Senior Software Engineer

LightBox · Remote

2 days ago

Senior Associate Product Management - Authorizations Data Platform

Capitalone · McLean, VA, United States of America +2

2 days ago

Sr. Scientific Data Engineer, R&D Data Platform

Abbott · United States of America : Remote · Onsite, Remote

2 days ago

Sr. Scientific Data Engineer, R&D Data Platform

Abbott · United States of America : Remote · Onsite, Remote

2 days ago

Senior Platform Engineer (Data Platform / DevOps)

Job Listings · Bangalore, India

2 days ago

Senior Data Platform Engineer

Job Listings · Bangalore, India

2 days ago

Senior Backend / Data Platform Engineer – Customer Data

Ing · Bruxelles Avenue Marnix (ING), Belgium

2 days ago

Senior Field Specialist - Data Platform

Cloudera Careers · UK-Remote, United Kingdom · Remote

2 days ago

Senior Data Analyst - Wise Platform Pricing

WISE · London, United Kingdom · Hybrid

2 days ago

Senior Software Engineer - Big Data Platform

Walkme · Tel Aviv · Hybrid

3 days ago

Senior Data Platform Engineer

Smartlyio · Berlin, Berlin, Germany; Helsinki, Uusimaa, Finland +1

3 days ago

Senior Software Engineer - Data Platform - Kubernetes - Federal

ServiceNow · San Diego, CALIFORNIA, United States · Hybrid

3 days ago

Senior Software Engineer, Data Platform

Harvey · New York · Hybrid

3 days ago

Senior Software Engineer, Data Platform

Harvey · San Francisco · Hybrid

3 days ago

Senior Java Data Platform Engineer - Clearance Required

Bti36021 · Herndon, VA +1

3 days ago

Senior Manager, Data & AI Platform

Dolby · Bengaluru, KA,IN, IN

4 days ago

Senior Data Platform Engineer

Clio · Calgary, Canada +3 · Remote

4 days ago

Senior Director, Semantic Data Platform

WEX · California - Remote Office, United States of America · Remote

4 days ago

Senior Middleware Platform Engineer - Data Transport - Manager

Statestreet · Bangalore, India · Hybrid

4 days ago
Senior Data Platform Engineer at Aperia | Hiring.Camp