About this role
WHAT WE DO
Founded in 2007, Growth Acceleration Partners (GAP) is a consulting and technology services company. We consult, design, build and modernize revenue-generating software and data engineering solutions for clients. With modernization services and AI tools, we help businesses achieve a competitive advantage through technology. GAP’s remote, integrated engineering teams use end-to-end solutions to innovate and align with your business goals. We have 600+ English-speaking engineers based in Latin America and approximately 20 U.S.-based engineers. With some of the highest customer satisfaction scores in the industry, GAP’s focus is customer and employee success.
GAP is a woman-owned company headquartered in Austin Texas. We are a values-based company focused on growing our people by investing in education, onsite English classes and training in the latest technologies, including AI, data analytics and machine learning. Our goal is to provide solutions for our customers that help them achieve critical business outcomes, while enabling our GAPSters and our communities to attain long-term success.
Summary
We are seeking a Senior Data Engineer to take technical ownership of data quality, identity resolution workflows, and data governance within a large-scale enterprise analytics platform. In this client-facing role, you will bridge identity data processing with downstream analytical layers, ensuring data integrity, source-to-destination validation, and seamless data contract management. Operating with technical autonomy, you will collaborate with data analysts, platform engineers, and cross-functional stakeholders to establish repeatable data governance standards and troubleshoot complex data flows across medallion architectures.
Education
Bachelor’s Degree in Computer Science, Data Engineering, Information Systems, or a related technical field.
Professional Experience
5+ years of professional software and data engineering experience building, profiling, and maintaining enterprise data pipelines.
Proven track record in identity resolution, entity matching, and data lineage validation within complex first-party data ecosystems.
Strong experience with modern cloud data architectures (Medallion architecture: Bronze, Silver, Gold layers) and enterprise data sharing protocols (e.g., Delta Share).
Demonstrated expertise writing advanced SQL, Python, and PySpark scripts to profile, transform, and validate large-scale datasets.
Hands-on experience developing data contracts, usage guidelines, and governance documentation using tools like Atlan and Confluence.
Key Responsibilities
Profile and validate digital identity records, verifying correct table joins, identity mapping logic, and downstream propagation across data layers.
Establish and maintain formal data contracts, join patterns, usage guidelines, and technical documentation in central governance repositories.
Review high-impact business analyses prior to stakeholder delivery to identify methodological errors, data quality anomalies, or improper join logic.
Perform root-cause analysis on pipeline failures, missing data transformations, and downstream propagation breaks, coordinating fixes through engineering backlogs.
Serve as the technical consultative link between identity platform capabilities, core data engineering teams, and enterprise business analysts.
Define, track, and automate repeatable quality assurance checks across cloud data lakes and distributed data sharing feeds.
Required Technical Skills
Strong experience in several of the following areas:
SQL & Data Profiling: Advanced SQL mastery for complex join logic, data lineage tracing, data validation, and source-to-visual reconciliation.
Programming & Processing: Python and PySpark for data transformation, pipeline troubleshooting, and automated quality validation.
Data Architecture & Governance: Medallion architecture (Bronze/Silver/Gold), Delta Lake / Delta Share, Data Contracts, and Data Lineage tooling.
Pipeline Diagnostics: Diagnostic troubleshooting across cloud data warehouses, ETL/ELT pipelines, and distributed data environments.
Documentation & Tooling: Confluence, Atlan, Jira, or enterprise data catalog and data dictionary management platforms.
Nice to Have
Prior experience in automotive, digital marketing, or customer data platform (CDP) environments.
Exposure to applied data science methodologies, predictive analytics, or identity resolution graph algorithms.
Experience working in client-facing technical delivery or data consulting roles within enterprise accounts.
Soft Skills
Advanced English proficiency (written and verbal).
Exceptional stakeholder communication skills with the ability to explain complex technical data anomalies to non-technical business leaders.
Meticulous attention to detail and a high standard for data accuracy and reliability.
Strong analytical problem-solving mindset with high technical curiosity and self-direction.
Collaborative mindset capable of bridging gaps between engineering, product, and business analytics teams.
At Growth Acceleration Partners, we're an equal opportunity employer committed to building a diverse and inclusive team. We value everyone's unique background, and we provide equal opportunities regardless of race, color, creed, religion, sexual orientation, gender identity, age, national origin, disability, marital status, veteran status or any other personal right protected by law. We foster a culture of belonging and strive to provide a welcoming environment where everyone feels safe to contribute and grow.
Tired of cold applications?
Sign up with Clera and we'll reach out the moment a role actually fits you — no more spraying applications into the void.
Know someone who'd be great for this?