TU PRÓXIMO CAPÍTULO
Senior Data Engineer, Data Lake
Sobre el puesto
• Develop, test, and maintain data ingestion, transformation, and quality-check pipelines using PySpark/SparkSQL on Databricks
• Build and modify Airflow DAGs (MWAA) for pipeline orchestration
• Write and optimize SQL queries for data transformation and validation
• Implement and monitor data quality checks using Great Expectations or equivalent
• Participate in code reviews
• Investigate pipeline failures and data quality issues with guidance from senior engineers
• Write and maintain documentation for datasets and pipelines
• Participate in sprint planning, estimation, and retrospectives
• Follow team-defined CI/CD, testing, and governance standards
• Contribute to the reliability and quality of Visa's corporate data lake
• Must be based in Brazil
• Proficiency in English at B2 level or above (Upper-Intermediate)
• 3+ years of relevant work experience with a Bachelor's or Associate’s Degree OR 5+ years of relevant work experience
• Python for automation and data processing
• Data pipeline concepts (batch, ETL/ELT patterns)
• Apache Spark basics (PySpark or SparkSQL)
• Databricks (jobs, workflows, cluster management, tuning)
• SQL (intermediate–advanced: joins, window functions, CTEs)
• Amazon S3 data lake design (partitioning, layout, lifecycle)
• Basic cloud concepts (AWS: S3, IAM, CloudWatch)
• Data modeling (dimensional / Kimball, medallion layers)
• Git/GitHub workflow and code review
• Preferred: Delta Lake, Apache Airflow/MWAA, CDC patterns, data quality concepts, Terraform basics, BI integration (Superset, dashboarding))
• Remote work arrangement
• Opportunity to create impact at scale
• Skill growth and professional development through meaningful work
• Scheduled-notice Visa office presence may be available for remote positions