Role Overview
We are seeking a Senior Data Engineer who will play a lead role in governing code quality, reviewing data pipelines, and ensuring scalable data product engineering.
The candidate will work closely with distributed data teams to enforce best practices across Snowflake, Redshift, AWS Athena, and Python-based data pipelines, while driving robust data modeling and architecture standards.
Key Responsibilities
Code Review & Engineering Governance (Primary Focus)
Lead code reviews across data engineering teams (SQL, Python, ETL pipelines) to ensure:
Performance optimization (query tuning, partitioning, clustering)
Maintainability and readability of code
Adherence to enterprise coding standards and design principles
Establish and enforce coding standards, review frameworks, and best practices
Identify anti-patterns in data pipelines and recommend improvements
Drive adoption of CI/CD, version control, and automated quality checks
Data Platform Engineering (Snowflake / AWS / Redshift / Athena)
Design, review, and optimize data pipelines and transformations across:
Snowflake (EDW / Data Lakehouse)
Amazon Redshift & AWS Athena
Ensure efficient handling of large-scale structured and semi-structured datasets
Optimize data ingestion, transformation, and consumption layers
Guide teams in building scalable, cost-efficient cloud data solutions leveraging AWS services
Data Modeling & Architecture
Define and review logical and physical data models (dimensional, normalized, data vault, etc.)
Ensure alignment with data product architecture and consumption patterns
Drive data consistency, reusability, and governance standards
Implement best practices for data lineage, metadata management, and documentation
Data Quality & Validation
Establish frameworks for:
Data validation and reconciliation
Anomaly detection and monitoring
Ensure reliability of pipelines through robust testing strategies (unit/integration)
Stakeholder Collaboration & Mentorship
Work with data engineers, architects, and product owners to translate requirements into scalable solutions
Provide technical leadership and mentoring to junior engineers
Act as a gatekeeper for production-ready data assets and pipelines
Required Skills & Experience
8–12+ years in Data Engineering with strong hands-on coding and review experience
Strong expertise in:
Snowflake, Redshift, AWS Athena
Python (data processing, scripting, frameworks)
Advanced SQL & query optimization
Proven experience in:
Data modeling techniques (Star schema, Snowflake schema, Data Vault)
ETL/ELT pipeline design and optimization
Working with AWS data services (S3, Glue, etc.)
Familiarity with:
CI/CD pipelines, Git-based workflows
Orchestration tools (Airflow, etc.)
Data quality and governance frameworks