About the role
What will you do at Saviynt?
PLATFORM SUPPORT ENGINEER
As a Platform Support Engineer in our SRE Operations team, you will be a key
player in ensuring the 24x7x365 smooth operation of Saviynt’s Enterprise
Identity Cloud. This role focuses on maintaining the stability, performance, and
reliability of our platform with a strong emphasis on application layer support
and operational ownership. You will be working closely with other operations
team members, development, and engineering to resolve issues, implement
improvements, and provide exceptional support. This is an opportunity for
someone who enjoys operational challenges and problem-solving in a dynamic cloud
environment and wants to see their work through to completion.
WHAT YOU WILL BE DOING
* Strong pod-level troubleshooting skills in AKS/EKS (not just restarting
pods).
* Analyze application and DB (RDS, MySQL) performance issues.
* Deeply investigate and analyze application performance issues (Java, Grails,
Hibernate), identifying root causes and implementing solutions.
* Oversee the monitoring of our SaaS applications and underlying infrastructure
(Kubernetes on AWS and Azure, VPN connections, customer applications, Elastic
Search, MySQL) for alerts and performance issues.
* Strong understanding of basic computing concepts like DNS, IP addressing,
Networking, and LDAP.
* Effectively participate and contribute in on-call escalations with a strong
operational mindset and provide technical guidance during critical incidents.
* Proactively communicate with customers on technical issues when required.
* Ability to guide junior engineers when needed technically.
* Manage the full lifecycle of alerts, incidents, and service requests reported
through FreshService, ensuring timely and accurate logging, prioritization,
resolution, and escalation.
* Develop, implement, and maintain operational procedures, runbooks, and
knowledge base articles to standardize incident resolution and service
request fulfillment.
* Drive continuous improvement initiatives to optimize operational efficiency,
reduce incident rates, and improve service request turnaround times via
engineering automation to reduce toil and waste.
* Collaborate with backend engineering and development teams to troubleshoot
complex issues, identify root causes, and implement preventative measures.
* Ensure adherence to defined SLAs (Service Level Agreements) and KPIs (Key
Performance Indicators) for operational performance.
* Maintain operational documentation, including system diagrams, contact lists,
and escalation paths.
* Ensure compliance with relevant security and compliance policies.
* Plan and coordinate scheduled maintenance activities with minimal impact to
service availability.
WHAT YOU BRING
* Bachelor's degree in Computer Science, Information Technology, Engineering,
or a related field.
* Minimum of 2-5 years of experience in IT/Cloud operations and application
support (specifically Java apps), with knowledge of cloud infrastructure (AWS
and Azure).
* Kubernetes Certification or 3+ years of experience supporting Kubernetes.
* Strong experience with application support (Java, Grails, Hibernate) and
performance analysis in a production environment, able to pinpoint a
performance degradation through analysis.
* Strong understanding of cloud computing concepts, architectures, and services
on both AWS and Azure platforms.
* Working knowledge of containerization and orchestration technologies,
specifically Kubernetes.
* End-to-end technical accountability and operational ownership.
* Willingness to work in a 24/7 operating model (including Night Shift).
* Experience managing and troubleshooting network connectivity, including VPNs
and connections to external networks.
* Familiarity with monitoring tools and practices, with experience in setting
up and responding to alerts.
* Hands-on experience with log management and analysis tools, preferably
Elastic Search.
* Working knowledge of database systems, preferably MySQL, including L2
troubleshooting and performance monitoring.
* Experience with ITSM (IT Service Management) systems, preferably
FreshService, including incident, problem, and service request management
processes.
* Excellent problem-solving, analytical, and troubleshooting skills with a
data-driven approach.
* Experience with Grafana systems and dashboards is a plus.
* Strong communication (written and verbal), interpersonal, and presentation
skills.
* Ability to work effectively under pressure and manage multiple priorities in
a fast-paced environment.
* Experience in developing and documenting operational procedures and runbooks.
* Experience with automation tools and scripting languages (e.g., Python, Bash)
is a plus.
* Experience working in a SaaS environment is highly desirable
Must Have Requirements
* Working on FedRamp (AWS Gov/ Azure Government)
* US - Citizen
* Kubernetes Certification (CKA)
We offer you a competitive total rewards package, learning and tremendous
opportunities to grow and advance in your career. At Saviynt, it is not typical
for an individual to be hired at or near the top of the range for their role and
final compensation decisions are dependent on many factors including, but are
not limited to location; skill sets; experience and training; licensure and
certifications; and other relevant business and organizational needs.
Saviynt is an amazing place to work. We are a high-growth, Platform as a Service
company focused on Identity Authority to power and protect the world at work.
You will experience tremendous growth and learning opportunities through
challenging yet rewarding work that directly impacts our customers, all within a
welcoming and positive work environment. If you're resilient and enjoy working
in a dynamic environment you belong with us!
Saviynt is an equal opportunity employer and we welcome everyone to our team.
All qualified applicants will receive consideration for employment without
regard to race, color, religion, sex, sexual orientation, gender identity,
national origin, disability, or veteran status.
\n
\n
Which skills does this role require?
Make your next move
Build a shortlist and prepare
Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.
- Build a focused shortlist before you applyCompare role requirements with your experience and give each application a clear reason.
- Practice explaining your experience in an interviewRehearse your answers before meeting the hiring team.
Other roles to compare
Review the responsibilities and requirements before adding an opening to your shortlist.
AI Outcome Customer Engineer, Forward Deployed Engineering
Google · Atlanta, Georgia, United States
AI Workflow Engineer
Scout Motors Inc. · Charlotte, North Carolina, United States
AI Integration Engineer (Hybrid)
DXC Technology · Arlington, Virginia, United States
Analytics Sr Software Engineer (US Federal)
Workday · Reston, Virginia, United States
AI Solutions Engineer
Superior Essex · Sandy Springs, Georgia, United States
Automation Engineer, CGIC & BD AI Automation
Quantum Sky · Reston, Virginia, United States
Role information can change. Confirm current details on the original application page.
