About the role
What will you do at SchoolsFirst Federal Credit Union?
We’re always looking for diverse, talented, service-oriented people to join our exceptional team. Splunk Administrator (Site Reliability Engineer) The pay range for this position is listed below. Our pay ranges are built to allow for candidates with various levels of skill and experience to be considered, as well as for room for growth and tenure achieved in a role over time.
Typical new hire salary offers fall within the minimum to midpoint of a pay range for many candidates. Any offer extended to a candidate will be based upon their unique set of knowledge, skills, education, and experience as well as internal equity. Pay Range: $46.90 - $75.04 Scheduled Weekly Hours: 40 What You’ll Be Doing Responsible for deploying, managing and optimizing.
Onboard machine data, build queries using Search Processing Language (SPL) and design dashboards to power IT Enterprise Applications, Security Operations and observability practices. Proactively monitor and report on the environmental health of applications, as well as engage in critical system outages (triage) resolution efficiently to identify, review, analyze, debug and resolve issues with a team of Site Reliability Engineers. Serve as a subject matter expert for the Splunk Platform. .
Ensure enterprise applications and supporting platforms are fully functional, reliable, available, performant, and secure through frequent monitoring, operational checks, and timely restoration of service in the event of outages, ingestion issues, or system failures.
Ingest and onboard new operational data sources, including server, application, infrastructure, network, security, API, and system logs; ensure data is reliable, appropriately normalized, accurately timestamped, correctly categorized, and aligned to the Splunk Common Information Model where applicable.
Monitor and maintain Splunk platform health, including indexers, search heads, data flows, storage utilization, clustered or distributed components, internal logs, capacity trends, and service availability; resolve ingestion bottlenecks, performance issues, and platform outages.
Create, maintain, and improve real-time Splunk dashboards, visualizations, alerts, reports, metrics, logs, traces, and event views to provide actionable insight into system health, service levels, user experience, operational risk, network activity, security events, and threat indicators.
Craft, tune, and optimize Splunk Processing Language (SPL) searches, scheduled searches, alert logic, reports, dashboards, and resource-intensive queries to improve performance, reduce noise, support accurate KPI/SLA reporting, and enable analysis across large datasets. Configure automated alerts and triggers for anomaly detection, system downtime, performance degradation, ingestion issues, and cyber or operational threats so critical issues are flagged immediately and routed for response.
Alert and accurately report KPIs on systems status with tuning recommendations at regular intervals; provide detailed analysis using APM, Splunk, and other monitoring or observability solutions. Support a 24x7 production environment with a team of experienced engineers, including on-call rotation, deployment support, and timely response to critical incidents as needed.
Respond quickly to incidents, investigate triggered alerts, isolate performance bottlenecks, fix issues, and work with other engineers to ensure enterprise applications and platforms remain fully functional with a Member-first focus.
Serve as an escalation point for complex application, infrastructure, observability, security, and platform issues; coordinate resolution across infrastructure engineering, software engineering, software quality assurance, cybersecurity, operations, vendors, and other organizational teams. Support incident triage, root cause analysis, post-incident reviews, postmortems, and corrective action planning; identify recurring issues and recommend preventive improvements.
Develop solutions to meet technical and business requirements; create detailed design documents and associated solutions built around current and new technologies. Demonstrate application environment tuning abilities and provide solutions to capacity requirements, including JVM, web containers, database connections, HTTP servers, and related enterprise application components.
Contribute to automation and efficiency improvements for application and infrastructure buildout components, operational tasks, monitoring coverage, and deployment readiness. Ensure effective release management, change management, quality assurance, rollback planning, peer review, and deployment readiness in project lifecycles using an ITIL-centric approach.
Support production and non-production environments, account for differences in each environment for deployments and testing, and peer review changes within the Enterprise Applications team. Maintain and improve configuration documentation, diagrams, SOPs, runbooks, work/incident ticketing systems, knowledge articles, training materials, and other supporting documentation.
Ensure adherence to configuration, operating, security, compliance, and governance best practices and standards, including approved capacity, licensing, access control, and retention requirements. Deploy and troubleshoot applications on both Java and . NET platforms as an expert for infrastructure and development teams.
Ensure the highest levels of service for Members and Team Members. Additional Job Functions Performs other duties as assigned Complies with regulatory compliance and assigned training requirements including but not limited to BSA regulations corresponding to their specific job duties. Failure to do so may result in disciplinary and other employment related actions.
Qualifications
- High School Diploma or GED required Bachelor's Degree in a related field or equivalent years of experience required 5-7 years of prior relevant experience required Knowledge, Skills, and Abilities Knowledge of the following: Java Enterprise Edition and .
- NET frameworks.
- Standard java/.
- NET tools to debug application and system performance issues.
- Tuning application serving environments.
- Software release management and the SDLC.
- Redhat Linux and Microsoft Windows Server 2008/2012.
- Security best practices for multi-tiered operating systems as well as application security practices.
- Disaster recovery experience Preferred Application Performance Management, such as Application Dynamics, Splunk and/or SNMP monitoring tools IBM Security Access Manager (ISAM), IBM/Apache HTTP Server, IBM Secure Directory Server Preferred Scripting skills for automating tasks Experience with Docker, Kubernetes, Openshift, Python, Jenkins, Git/Gitflow, Ansible, Quay and Artifactory Preferred Managing IBM AIX, Redhat, Linux and/or Microsoft Windows Servers, process development, support and implementation of systems architecture Must possess solid oral and written communication, project management, and interpersonal skills.
- SchoolsFirst FCU is committed to Diverse, Equitable, and Inclusive Hiring At SchoolsFirst FCU we are dedicated to building and growing a diverse, inclusive, and authentic Dream Team, so if you’re excited about a position or wanting to make a career change but your past experience doesn’t align perfectly with every qualification in the job description, we encourage you to apply anyway.
- Many skills are transferrable and you may be just the right candidate for the position, or for other roles we are working on.
- SchoolsFirst Federal Credit Union is committed to fostering, cultivating, and preserving a culture of diversity and inclusion.
- SchoolsFirst FCU is an equal opportunity employer and prohibits discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities and prohibits discrimination against all individuals based on their race, color, religion, sex, national origin, age, sexual orientation, gender identity or expression, political affiliation, or genetic information.
- This organization participates in E-Verify.
- We're passionate about living our mission of providing Members with World-Class Personal Service and financial security.
- There's a reason we love coming to work every day.
- We have the privilege of serving those who build the future: school employees, and the families who make their work possible.
- Our success is built on our relationships with each other, our Members and the community.
- We value teamwork and face-to-face interactions and believe we’re at our best when we’re collaborating in-person.
- We also know the flexibility to work remotely benefits the wellbeing of our team.
- View the work location information below on the job listing that interests you. 100% on-site, 0% remote.
- Including those serving Members in-person at our branches.
- Three days a week in the office and up to two days remote.
- Hybrid teams coordinate on-site days to collaborate in-person.
- Up to 100% remote.
- Select individual contributor roles may be given the option to work fully remote.
- Not finding the right fit?
- Let us know you are interested in a future opportunity below, or create an account by clicking 'Sign In' at the top of the page to set up email alerts as new job postings become available that meet your interest!
Which skills does this role require?
Make your next move
Build a shortlist and prepare
Identify the requirements you can demonstrate, then choose examples from your work to discuss with the hiring team.
- Build a focused shortlist before you applyCompare role requirements with your experience and give each application a clear reason.
- Platform engineering jobsCompare current openings and review what to look for in this role.
- Practice explaining your experience in an interviewRehearse your answers before meeting the hiring team.
Other roles to compare
Review the responsibilities and requirements before adding an opening to your shortlist.
Adaptive Network Infrastructure Security Engineer
V2X Inc · United States
Infrastructure Engineer - Security
Milliman MedInsight · Dallas, Texas, United States
Sr. Network Engineer, AI Infrastructure (Starshield)
SpaceX · Palo Alto, California, United States
Infrastructure Engineer Sr. (Azure Cloud Product Team)
PNC · Pittsburgh, Pennsylvania, United States
AI Infrastructure Platforms - Senior Product Manager
JPMorganChase · Palo Alto, California, United States
Hybrid Infrastructure Engineer (Cloud Migration Lead)
Paramount Acceptance · Holladay, Utah, United States
Role information can change. Confirm current details on the original application page.
