Senior DevOps Engineer, IT Data Operations Engineering
Job Summary
Innovation starts from the heart.At Edwards Lifesciences we are dedicated to developing ground-breaking technologies with a genuine impact on patients lives. At the core of this commitment is our investment incutting-edgeinformation technology enabling our global teams to collaborateoperateefficiently and advance our patient-focused mission.
As part of our IT team yourexpertiseand commitment will help build andoperatethe platforms that power Edwards global data and analytics solutions. We are seeking a Senior DevOps Engineer to support Data Operations Engineering and own the operational backbone of our data platforms including infrastructure as code continuous integration and delivery data quality monitoring and incident response observability security operations and cloud cost optimization. Our multi-cloud estate includes Databricks and Snowflake with data services moving to Azure while AWS continues to host corporate source systems.
How you will make an impact
Build andmaintainmodular Terraform components that provision cloud and Databricks platform capabilities including workspaces cluster policies secret scopes networking and Unity Catalog storage credentials and external locations consistently across environments and clouds.
Design andoperateautomated GitLab CI/CD pipelines that deploy data platform code and infrastructure through development staging and production with approval gates and controlled promotion.
Monitor data quality and pipeline health own first-line response to production alerts resolve issues directly where possible and escalate only when the root cause requires deeper pipeline or platform change.
Build test and continuously improve runbooks for recurring failure modes automating remediation when the response can be made safe and repeatable.
Apply AI- and agent-assisted operations capabilities to accelerate detection triage and remediation as those capabilities mature.
Build end-to-end observability across Databricks Snowflake and orchestration services by bringing platform and execution metadata into unified dashboards alerting and pipeline service-level monitoring.
Containerize andoperateplatform workloads orchestration and CI runners on Docker and Kubernetes across EKS and AKS.
Operate and re-host Apache Airflowconsolidatingself-managed instances onto Kubernetes or a managed runtime while supporting migration of legacy build tooling to GitLab CI.
Embed security operations into the platform lifecycle throughsecretsmanagement short-lived credentials least-privilege access encryption key policies audit logging and vulnerability remediation.
Drive cloud cost optimization through cost attribution policy enforcement right-sizing tagging standards and anomaly alerting.
Support secure cross-cloud connectivity and data movement between AWS-resident source systems and the Azure data platform.
Identifyopportunities for automation reusable standards reliability improvement and technical debt reduction andprovidetechnical guidance to team members.
What you will need (Required)
Bachelors Degree in Computer Science Engineering Information Technology or a related field and a minimum of 5 years of relevant experience in DevOps platform engineering site reliability engineering or cloud infrastructure roles.
Strong hands-on experience administering and automating workloads on both Amazon Web Services and Microsoft Azure including identity networking and storage services.
Demonstratedexpertisebuilding modular reusable infrastructure with Terraform across multiple environments.
Proven experience designing andmaintainingCI/CD pipelines in GitLab CI or an equivalent platform including environment promotion approval gates andsecretshandling.
Strongproficiencywith Docker and Kubernetes including EKS or AKS and running CI/CD or orchestration workloads on Kubernetes.
Hands-on experience operating Apache Airflow including DAG deployment scheduling retries backfills pools and failure recovery.
Solid Linux systems administration shell scripting and troubleshooting skills.
Hands-on experience implementing metrics logging dashboards and alerting using tools such as Datadog or Grafana.
Demonstrated experience monitoring pipeline and data quality owning first-line triage of production alerts and writing and executing runbooks for recurring failures.
Working knowledge ofsecretsmanagement least-privilege access control encryption key management audit logging and vulnerability remediation.
Working knowledge of cloud cost monitoring attribution tagging standards and right-sizing practices.
Proficiencyin Python or another scripting language used for automation and tooling.
Excellent analytical troubleshooting communication and organizational skills with the ability to work across cross-functional and geographically distributed teams.
Ability to manage competing priorities in a fast-paced environment whilemaintainingattention to quality controls and long-term outcomes.
Adhere to all company rules and requirements including Environmental Health & Safety rules and take adequate control measures to prevent injuries protect the environment and prevent pollution under their span of influence or control.
What else we look for (Preferred)
Experience operating Databricks or Snowflake including platform administration policy management and service principal or access management.
Experience with Databricks Asset Bundles Unity Catalog the Databricks Terraform provider or managed Terraform state and gated production applies.
Experience with keyless authentication from CI such as OIDC-based role assumption or IRSA.
Experience with data observability or AI-assisted operations capabilities such as Monte Carlo Datadog AI capabilities or Databricks operations automation.
Experience with Astronomer MWAA or Airflow on Kubernetes usingKubernetesExecutor.
Experience with Azure Data Factory ADLS Gen2 Key Vault Azure Monitor and Entra ID or AWS Glue Athena Lambda and Secrets Manager.
Experience supporting hybrid or cross-cloud architectures and private connectivity among clouds and on-premises environments.
Experience migrating CI/CD from Jenkins or another legacy build platform.
Experience withstreaming orchangedata capture technologies such as AWS DMS Amazon MSK or Kafka.
Experience with SALT Ansible Azure DevOps or related configuration and delivery tooling.
Relevant experience in medicaldevice pharmaceuticals or another highly regulated environment including validation and change control expectations.
Relevant AWS Azure Kubernetes or Terraform certifications.
Additionalinformation
Develops solutions to a variety of complex platform and operational problems and exercises judgment in selecting methods and techniques after considering risk and alternatives.
Works with general direction focused on end results and mayprovidetechnical guidance to lower-level personnel.
Participation in an on-call rotation may beto support platform reliability.
Travel may beup to 5%.
Benefits and equal opportunity
Aligning our overall businessobjectiveswith performance we offer competitive salaries performance-based incentives and a wide variety of benefits programs to address the diverse individual needs of our employees and their families.
Edwards Edwards Lifesciences and the stylized E logo are trademarks of Edwards Lifesciences Corporation or its affiliates. All other trademarks are the property of their respective owners.
Recruiting scam alert: Read our notice about potential recruiting scams.
Required Experience:
Senior IC
About Company
Edwards Lifesciences (NYSE: EW), is the global leader of patient-focused medical innovations for structural heart disease and critical care monitoring. We are driven by a passion for patients, dedicated to improving and enhancing lives through partnerships with clinicians and stakehol ... View more