Prepin
Log in
Colonial Life

engineering opportunity

Site Reliability Engineer

Design, build, and maintain observability and monitoring capabilities for digital platforms while troubleshooting distributed systems. Partner with engineering teams to improve service reliability, automate health checks, and manage CI/CD pipeline quality.

Atlanta, Georgia, United StateshybridFULL_TIME

Posted

About the role

What will you do at Colonial Life?

Job Posting End Date: August 14

Our Fortune 500 company is driving a digital transformation and looking for forward-thinking innovators to disrupt how our industry thinks about and uses technology. As one of the world's leading employee benefits providers, we help millions of people gain affordable access to benefits that help them protect their families, their finances and their futures.

Are you an asker of questions, a solver of problems, and a challenger of the status quo? Our mission is to provide a differentiated customer experience and exceed the expectations people have of technology at any company — not just insurers.

We are seeking individuals to join our team of talented IT professionals who share never-ending passion and an unwavering focus on our customer experience. Team members comfortable working in an agile, fast-paced, and delivery-focused environment thrive in our environment where we value an entrepreneurial spirit and those who challenge the status-quo.

Unum is changing, and we’re excited about what’s next. Join us.

General Summary:

Unum Group seeks Site Reliability Engineers in Atlanta, GA.

Applicants who are interested in this position may apply at www.jobpostingtoday.com (Ref #66753) for consideration.

Design, build, and maintain observability, monitoring, and alerting capabilities across consumer and client-facing digital platforms

Develop and maintain dashboards that measure availability, latency, error rate, throughput, capacity, MTTR/MBTI, and other reliability metrics

Diagnose and troubleshoot distributed systems issues across cloud-based and on-prem services

Partner with engineering teams to improve service reliability, reduce operational toil, and mature incident response practices

Implement automation for service health checks, performance monitoring, and remediation

Manage CI/CD pipeline reliability and deployment quality controls

Conduct root-cause analysis and drive long-term corrective actions

Collaborate with Run teams to transition monitoring, dashboards, and operational insights into production support processes

Provide guidance on service re-platforming, performance improvements, and architectural decisions based on reliability data

Requires a Bachelor’s degree in Computer Science, Engineering, or related field plus 5 years of experience.

Requires 5 years of experience with the following: Observability and monitoring platforms used to monitor application performance and system health using Dynatrace, AWS CloudWatch, Datadog, Grafana, or Amplitude; Working with containerized and cloud-native architectures, including deployment, configuration, and operational support in cloud environments, using AWS; Supporting incident response processes, including participation in on-call rotations, post-incident reviews, and implementation of service-level objectives (SLOs), service-level indicators (SLIs), or service-level agreements (SLAs); Developing scripts or automation to improve system reliability or operational efficiency using Python, Bash, or PowerShell; Troubleshooting distributed systems and analyzing performance bottlenecks across multi-tier or microservices-based architectures; Collaborating with cross-functional engineering teams, including software engineering, platform, infrastructure, or operations teams, within a DevOps or reliability-focused environment; working with version control systems and collaborative development workflows using GitHub, GitLab, or Bitbucket.

Requires 4 years of experience with the following: Designing, implementing, or maintaining logging, metrics, and distributed tracing pipelines for enterprise or cloud-based systems; Hands-on experience with continuous integration and continuous deployment (CI/CD) tools and pipelines, using GitHub Actions, Jenkins, or Azure DevOps.

Requires 3 years of experience with using infrastructure-as-code or configuration management tools to provision, manage, or maintain environments, including Terraform, AWS CloudFormation, or Ansible. Telecommuting w/i worksite. Up to 5% domestic travel.

40 hours/week; $152,131 - $162,131 per year. This wage range supersedes the base salary range listed below, due to the salary range below reflecting a national range.

#LI-TS1

~IN1

Our company is built on helping individuals and families, and this starts with our employees. We want employees to maintain a positive balance, which is why we provide access to the benefits and resources they need to invest in themselves. From our onsite fitness facilities and generous paid time off to employee professional development programs, we are committed to helping employees live and work their best – both inside and outside the office.

Unum is an equal opportunity employer, considering all qualified applicants and employees for hiring, placement, and advancement, without regard to a person's race, color, religion, national origin, age, genetic information, military status, gender, sexual orientation, gender identity or expression, disability, or protected veteran status.

The base salary range for applicants for this position is listed below. Unless actual salary is indicated above in the job description, actual pay will be based on skill, geographical location and experience.

$98,340.00-$201,900.00

Additionally, Unum offers a portfolio of benefits and rewards that are competitive and comprehensive including healthcare benefits (health, vision, dental), insurance benefits (short & long-term disability), performance-based incentive plans, paid time off, and a 401(k) retirement plan with an employer match up to 5% and an additional 4.5% contribution whether you contribute to the plan or not. All benefits are subject to the terms and conditions of individual Plans.

Company:

Unum

Which skills does this role require?

Site Reliability EngineeringCI/CDTerraformAnsibleSite Reliability EngineerAlertingSLOSLISLAPerformance monitoringRoot-cause analysisInfrastructure-as-codeDigital transformationAzureAgile

Make your next move

Build a shortlist and prepare

Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.

Review the responsibilities and requirements before adding an opening to your shortlist.

Role information can change. Confirm current details on the original application page.

Product

AI Candidate AgentCompaniesBrowse JobsDeep ProfileSkill AssessmentOpportunity Matching
Prepin.ai

© 2026 Prepin | All rights reserved.