Start Your Search Here

push notification bell

Would you like to receive notifications about jobs in West Hartford?

push notification bell

You have blocked notifications

Oops! You have blocked notifications. Click here for more info

You have blocked notifications, please check your browser settings.

push notification bell

You're currently subscribed to job notifications

Want to change your notifications for job alerts?

push notification bell

Subscribe to notifications

You will no longer receive notifications

Job Search

Diverse Lynx

West Hartford / Global

Platform Reliability Engineer

Job Description

Platform Reliability EngineerLocation: West Hartford, CT 06110, United States (100% Remote) Duration: 12 Month Contract Rate: $50/hrQualifications: 8+ years of relevant experience in platform engineering, infrastructure engineering, DevOps, production engineering, or site reliability engineering, including experience operating at senior or lead level. Demonstrated experience establishing observability and reliability practices for distributed enterprise applications — building operational capability from low or zero baseline, not solely operating within an already-mature environment. Hands-on experience with observability architecture and implementation — logging, metrics, distributed tracing, dashboards, alerting, and production monitoring using enterprise observability platforms such as Datadog or Dynatrace. Strong CI/CD pipeline ownership and build/release engineering experience, including CI pipeline monitoring and optimization using tools such as Datadog CI Visibility or equivalent. SAP Commerce Cloud (CCv2) operational experience strongly preferred, including familiarity with its deployment model, build process, monitoring, performance characteristics, and platform-specific constraints; comparable experience operating large-scale Java commerce or enterprise application platforms considered. Vercel or comparable cloud/edge platform experience at production scale. Experience with environment management, configuration governance, and provisioning across heterogeneous deployment targets. Strong experience with infrastructure-as-code and configuration-as-code practices for repeatable environment and platform configuration, using Terraform or comparable tooling where supported by the underlying platforms. Strong scripting and automation capability using languages such as Python, TypeScript/JavaScript, Bash, or PowerShell. Experience with performance testing infrastructure and environment provision for load testing; familiarity with tools such as k6 or comparable. Experience defining and operationalizing service level indicators (SLIs), service level objectives (SLOs), and other measures of production health. Working knowledge of automated testing practices and how unit, integration, contract, and end-to-end testing fit into CI/CD and release controls. Experience implementing dependency scanning and security scanning in CI/CD pipelines (Snyk, SonarCloud, or equivalent). Strong incident management, production troubleshooting, and post-incident review experience. Demonstrated ability to establish reliability, observability, and operational standards across multiple engineering teams. Strong ability to influence engineering teams, establish operational practices across organizational boundaries, and translate reliability risks into actionable technical and business decisions.
Apply Now

Similar Opportunities

View all jobs

Get Job Alerts

Don't miss the perfect fit. Get Daily curated job alerts.

Job Title or Keyword(s)
Location