Skip to main content
MetLife logo
MetLifeCompany application

Data Engineer II

Cary, United StatesFull-timePosted 54 days ago0 applicants
On-siteTechnology$90,000 – $110,000
Accepting applications
$90,000 – $110,000/ year
Type
Full-time
Mode
On-site
Level
Open

About the role

  • Description and Requirements The Team You Will Join At MetLife, data isn’t just a tool - it is a catalyst for growth. As part of our Data & Analytics organization, you’ll unlock trusted insights that drive bold decisions, power personalized customer experiences, and deliver lasting business impact. We’re building the future of data - one that’s governed responsibly, engineered for scalability, and designed for growth. When you join us, you’re not just supporting the business - you’re empowering it. Let’s transform insight into impact and data into action, together. The Opportunity The MetLife Corporate Functions Data Office is part of the Data and Analytics Organization (D&A) within GTO. Our mission is to implement scalable data solutions for our stakeholders to generate actionable insights. We achieve this by partnering with our D&A teams, Technology, and our Business and functional partners to build and deploy next generation data solutions for MetLife. As a Data Engineer II at MetLife, you will build, test, monitor, validate, and support data pipelines using Azure Data Factory, Databricks, PySpark, Spark SQL, SQL, and Python. The role requires strong hands-on engineering experience and significant Python expertise to develop, automate, validate, and operationalize data solutions supporting reporting, analytics, and operational decision-making. This role will initially focus on validation activities, including source-to-target validation, data profiling, reconciliation, anomaly detection, test automation, defect analysis, and data quality controls. Over time, the role is expected to contribute more broadly to pipeline development, optimization, deployment, and operational support within Azure and Databricks environments. Working with minimal supervision, you will perform intermediate to complex data engineering, data preparation, validation, evaluation, deployment, and operational support activities. You will partner with engineering, analytics, business, and operations teams to deliver reliable, analytics-ready data assets while ensuring data quality, performance, scalability, security, and compliance requirements are met. You may lead small project teams and contribute to engineering standards, reusable validation frameworks, and platform improvements. Key Responsibilities Perform data validation activities for enterprise data assets, including source-to-target validation, reconciliation, profiling, anomaly detection, and defect analysis.
  • Develop automated validation, testing, and data quality controls using Python, PySpark, Spark SQL, SQL, and related frameworks to ensure the accuracy, completeness, consistency, and timeliness of enterprise data.
  • Build, enhance, and support ETL/ELT pipelines using Azure Data Factory, Databricks, Python, PySpark, and Spark SQL, with an initial focus on validation and quality engineering use cases. Troubleshoot data issues, analyze root causes, document findings, and partner with Business, Technology, Operations, and Data & Analytics teams to resolve defects and improve data reliability.
  • Design, develop, and optimize scalable data processing and validation solutions that support reporting, analytics, and operational decision-making.
  • Implement and support Delta Lake and Lakehouse architecture patterns to enable reliable, scalable, and efficient data processing.
  • Ensure data processing and validation solutions comply with established security, quality, and operational standards.
  • Contribute to reusable validation frameworks, engineering standards, operational playbooks, and continuous improvement initiatives across the data platform.

Required Qualifications

  • Bachelor's or master's degree in computer science, Engineering, Information Systems, Mathematics, Statistics, Operations Research, or a related quantitative field, or equivalent experience. 3-5 years of experience in data engineering, analytics engineering, data validation engineering, data platform development, or related disciplines.
  • Strong hands-on experience developing data engineering and validation solutions using Python, PySpark, Spark SQL, and/or SQL.
  • Experience building, testing, validating, and supporting ETL/ELT pipelines using Azure Data Factory, Databricks, and Delta Lake architectures.
  • Experience developing, troubleshooting, and optimizing scalable cloud-based data solutions in Azure, including data quality, reconciliation, and validation activities.

Preferred Qualifications

Experience with data quality testing, data profiling, source-to-target validation, reconciliation, anomaly detection, and validation frameworks. Experience developing automated testing solutions using pytest or similar frameworks. Experience supporting production data pipelines, monitoring, observability practices, and incident or defect resolution. Experience with Git, Azure DevOps, and CI/CD pipelines. Experience with Attacama or similar data quality platforms and exposure to GenAI technologies. #LI-WRAPJOB

Location Expectation

This is a hybrid role requiring a minimum of 3 days per week in office. The expected salary range for this position is $90 ,000 - $110,000 . This role may also be eligible for annual short-term incentive compensation. All incentives and benefits are subject to the applicable plan terms.

Ready to apply?

Take the next step.
It takes 90 seconds.

Applications are reviewed directly by the MetLife hiring team. You will be redirected to their careers page.

0applicants so far
Full-timerole type
On-sitework mode

You can return to this role from saved jobs any time.