About this role
<div style="font-family:Arial;font-size:1.0em"> <p>At EY, we’re all in to shape your future with confidence. </p> <p>We’ll help you succeed in a globally connected powerhouse of diverse teams and take your career wherever you want it to go. </p> <p>Join EY and help to build a better working world. </p> </div> <div style="font-family:Arial;font-size:1.0em"> </div><p><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><strong>Staff Data Engineer</strong></span></p> <p> </p> <p> </p> <p><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><strong>The Opportunity:</strong></span></p> <p> </p> <p><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Join our dynamic team as a Senior Data Engineer, where you will play a critical role in designing, building, and optimizing scalable data solutions that drive business insights and innovation. You will work with cutting-edge technologies in a collaborative Agile environment, contributing to the full data lifecycle from ingestion to transformation and deployment.</span></p> <p> </p> <p> </p> <p><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><strong>Key Responsibilities:</strong></span></p> <p> </p> <ul> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Design, develop, and manage robust CI/CD pipelines to streamline the product development lifecycle and ensure seamless deployment of data solutions.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Perform data mapping, and integration between diverse source and target systems to support analytics and business intelligence needs.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Build and optimize Data & AI solutions, including data engineering pipelines, data modeling, and performance tuning, with 3-5 years of relevant experience.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Develop high-quality, production-level code primarily in Python, PySpark, and SQL, ensuring maintainability and scalability.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Implement enterprise-scale production deployments, applying DevOps best practices, with hands-on experience in Git version control.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Collaborate effectively within Agile teams, contributing to sprint planning, reviews, and continuous improvement.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Apply strong knowledge of Data Governance and Data Quality standards to maintain data integrity and compliance.</span></li> </ul> <p> </p> <p> </p> <p><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><strong>Technical Skills & Experience:</strong></span></p> <p> </p> <ul> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Extensive experience in building data ingestion frameworks and scalable data pipelines.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Proven expertise in AWS cloud services including S3, EC2, Glue, Lambda, and Secrets Manager for secure and efficient data processing.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Hands-on experience with Databricks for unified analytics and collaborative data engineering.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Proficient in DBT (Data Build Tool) for data transformation and modeling within modern data stacks.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Skilled in orchestrating workflows using Apache Airflow to automate complex data pipelines.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Strong programming skills in Python, PySpark, Spark, and Scala for big data processing.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Experience with software version control and CI/CD tools such as Git, Jenkins, and Apache Subversion.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">AWS / Databricks certifications or equivalent professional technical certifications are highly desirable.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Familiarity with cloud and enterprise integration technologies.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Demonstrated ability to write efficient Spark jobs and optimize performance in large-scale environments.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Excellent analytical skills with deep knowledge of SQL and data querying.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Minimum 3 years of experience working in very large data warehousing environments.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Strong communication skills, both written and verbal, to effectively collaborate with cross-functional teams.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">At least 3 years of experience with data warehouse architectures, ETL/ELT processes, and reporting/analytics tools.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Experience in Python and/or Java development in data engineering contexts.</span></li> <li style="font-family:arial, helvetica, sans-serif;font-size:10.0pt"><span style="font-family:arial, helvetica, sans-serif;font-size:10.0pt">Familiarity with Big Data ecosystems including EMR, Hadoop, Databricks, Hive, and Pyspark.</span></li> </ul> <p> </p><div style="font-family:Arial;font-size:1.0em"> <p><b>EY | Building a better working world </b></p> <p>EY is building a better working world by creating new value for clients, people, society and the planet, while building trust in capital markets.</p> <p>Enabled by data, AI and advanced technology, EY teams help clients shape the future with confidence and develop answers for the most pressing issues of today and tomorrow.</p> <p>EY teams work across a full spectrum of services in assurance, consulting, tax, strategy and transactions. Fueled by sector insights, a globally connected, multi-disciplinary network and diverse ecosystem partners, EY teams can provide services in more than 150 countries and territories.</p> </div>