JobHabor

Data Engineer

Software Guidance & Assistance, Inc. (SGA)

Location
Austin
Workplace
Hybrid
Employment
Contract
Salary
Apply on the employer’s site

Posted 1mo ago

The employer’s full description could not be read from their board. This is a summary of the posting — follow the apply link for the original.

Responsibilities

  • Design, build, and maintain data pipelines from Salesforce, Dynamics, Platform.sh, Snowplow, and cloud sources
  • Develop and optimize ETL/ELT workflows using Python and PySpark
  • Build and maintain data models (fact/dimension tables) for customers, deals, margin, usage, and migration data
  • Ensure data reliability, quality, and monitoring across ingestion and transformation
  • Design for scalability, fault tolerance, and performance
  • Build orchestration, alerting, and error-handling into pipelines
  • Manage schema evolution, versioning, and backfills
  • Translate data requirements from Analytics, Finance, Product, and Engineering into pipeline and schema designs
  • Support pre- and post-launch data needs for ACCS and ACO
  • Ensure data is accurate, timely, and well-documented so downstream teams trust it

Requirements

  • Design, build, and maintain data pipelines from Salesforce, Dynamics, Platform.sh, Snowplow, and cloud sources
  • Develop and optimize ETL/ELT workflows using Python and PySpark for large-scale batch and incremental processing
  • Build data models (fact/dimension tables) for customers, deals, margin, usage, and migration data
  • Own data quality, validation, and monitoring across ingestion and transformation layers
  • Design for scalability, fault tolerance, and performance across growing data volumes
  • Build orchestration, alerting, and error-handling into pipelines
  • Manage schema evolution, versioning, and backfills
  • Translate data requirements from Analytics, Finance, Product, and Engineering into pipeline and schema designs
  • Support pre- and post-launch data needs for ACCS and ACO
  • Ensure data is accurate, timely, and well-documented
  • Partner with Analytics and BI teams to ensure pipeline outputs meet modeling needs

Skills

  • Python
  • PySpark
  • AWS
  • Azure
  • Google Cloud Platform
  • Amazon Redshift
  • Snowflake
  • PowerBI
  • Salesforce
  • Dynamics
  • Platform.sh
  • Snowplow
  • AWS Glue

Similar roles