We are seeking a Senior Data Software Engineer to build reusable, governed data-sharing adapters that connect a cloud lakehouse and analytics warehouse to external data platforms while meeting strict access controls and SLAs. You will design lakehouse patterns, implement integrations, and strengthen lineage and governance.
Responsibilities
-
Design a UniForm write layer with one physical dataset and dual-format metadata for Delta and Iceberg consumers
-
Build and validate GCS-to-BigQuery ingestion pipeline patterns for structured operational data
-
Implement Kafka-based CDC patterns for real-time and near-real-time movement into the lakehouse
-
Develop dependency-aware bookkeeping and data lineage tracking patterns across pipelines
-
Engineer modular, version-controlled adapter code designed for reuse across new integrations
-
Configure Iceberg external table definitions in Snowflake using Horizon Catalog governance features
-
Validate zero-copy read access from Snowflake to Iceberg and Delta tables without data movement
-
Implement tenant-scoped access controls aligned with Snowflake metadata governance requirements
-
Implement and certify a Delta Sharing adapter for live, zero-copy sharing from Delta Lake to Databricks consumers
-
Configure Delta Sharing endpoints and manage sharing agreements for Databricks access
-
Validate Databricks read access via Delta Sharing for Spark, Pandas, and compatible consumers
-
Test end-to-end freshness and sharing latency to meet agreed SLA targets
-
Register connector types and implement RBAC plus tenant-scoped authorization for all data-out paths
-
Implement metering hooks compatible with billing requirements for governed data-out flows
Requirements
-
3+ years of data engineering experience with Python and cloud data platforms
-
Experience with Google Cloud BigQuery in advanced analytics and data modeling
-
Experience with lakehouse table formats including Apache Iceberg and Delta Lake
-
Strong leadership skills to drive integration designs, technical decisions, and delivery ownership
-
Proven project execution skills delivering data pipelines and adapters that meet SLAs
-
Advanced hard skills in Kafka/CDC patterns for real-time and near-real-time data movement
-
Strong architecture skills in data lake and lakehouse ingestion patterns on object storage
-
Excellent collaboration skills to work across platform, governance, and downstream consumer needs
-
Upper-Intermediate English proficiency (B2) for technical discussions and documentation
Nice to have
-
Apache Spark experience for validation and consumption testing
-
Databricks Unity Catalog knowledge for governed access patterns
-
Delta Lake expertise including Delta Sharing configuration and troubleshooting
-
Gen AI Assisted Development proficiency with tools such as Claude Code, GitHub Copilot, or Cursor
-
Snowflake Horizon Catalog experience for metadata governance and external table management
We offer
-
International projects with top brands
-
Work with global teams of highly skilled, diverse peers
-
Healthcare benefits
-
Employee financial programs
-
Paid time off and sick leave
-
Upskilling, reskilling and certification courses
-
Unlimited access to the LinkedIn Learning library and 22,000+ courses
-
Global career opportunities
-
Volunteer and community involvement opportunities
-
EPAM Employee Groups
-
Award-winning culture recognized by Glassdoor, Newsweek and LinkedIn
EPAM is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, disability, protected veteran status, or any other characteristic protected by applicable law.