We use essential cookies

Please Accept our Privacy Policy

Site Reliability Engineer (Hybrid)

Unlimited Systems

Cincinnati, OH • 9/12/2026

Job Description

Job Description

The Company

Unlimited Systems authors the category-leading Unlimited Financials practice management system focused on the unique revenue cycle requirements of specialty healthcare providers. UnlimitedSystemscustomers enjoy streamlined business office workflows, reduced claim denial rates, and accelerated and amplified revenuestreams. Unlimited Systems is committed to ensuring that specialty healthcare providers thrive in a dynamic reimbursement environment.

The Challenge

Unlimited Financialsoperatesacross a sophisticated cloud-native environment Azure Kubernetes Service, event-sourced microservices, healthcare integrations, and a growing network of downstream data pipelines all of which requires dedicated technical support to remain stable, observable, and performant for our customers. We are building the reliability practice this platform deserves. As we scale into new specialty healthcare markets and take on greater complexity, we need someone who thinks proactively not just responding to incidents but engineering the systems and signals that prevent them. This is not a run-and-maintain role. It is a “build-and-shape-the-future" role.


Our Site Reliability Engineer will own the observability strategy across our observability applications,defineand drive meaningful SLOs and SLIs for our critical services, and work shoulder-to-shoulder with DevOps and Software Engineering to make reliability a first-class concern not an afterthought. You will bring structure to chaos, clarity to ambiguity, and measurable improvement to customer experience.

The Position

A full-time, hybrid role reporting to our Technical Support Manager, our Site Reliability Engineer will workfrequentlyfrom our headquarters in Cincinnati, Ohio. Residency in the state of Ohio is preferred.

As Site Reliability Engineer, you will own the reliability, scalability, and performance of our production systems hosted on Azure Kubernetes Service (AKS), working at the intersection of software engineering and operations to build the tooling, processes, and culture that keep our services running at scale.You will be a key contributor to our observability practice using Splunk for log analytics and alerting,Instanafor APM and distributed tracing, and native Azure tools including Azure Monitor, Log Analytics, and Application Insights to provide a comprehensive, real-time view of system health.

This position will be accountable for:

Reliability & Incident Management

  • Defining, tracking, andreportingon SLIs, SLOs, and error budgets for all critical services.

  • Designing andmaintainingrunbooks, escalation paths, and on-call rotation schedules.

  • Designing chaos engineering practices to proactively surface reliability weaknesses before theyimpactcustomers.

Observability Splunk,Instana& Azure

  • Building and maintaining Splunk searches, dashboards, and alert policies covering application and infrastructure logs.

  • Developing KPIs and unified service-healthviewsfor engineering and leadership.

  • Configuring and extending instrumentation across microservices for distributed tracing and real-time baselining.

  • Creating smart alerts integrated with on call and ticketing systems for automated incident routing.

  • Maintaining Azure Monitor alert rules, action groups, and workbooks across Azure subscriptions.

  • UtilizingLog Analytics workspaces including data-retention policies and ingestion cost governance using KQL.

  • Leveraging Application Insights forAPM,availability testing.

  • Driving convergence of Splunk,Instana, and Azure signals into a unified observability strategy.

  • Building internal tooling in Python, Go, or Bash toeliminatetoil and accelerate incident response.

  • Participating in security reviews and threat-modeling sessions for new platform capabilities.

The Person

The opportunity to cast a vision for success, to blend artand sciencewithproactive strategyandtactical execution,will bea key to success in this role.

  • You’rea self-starter.In many ways,you’llhelpdefine success metrics for this position. Youintuitively understand what will berequiredto excel at this work andwon’texpect others to produce the blueprintfor you.You’llhit the ground running as the architect for thisfunction,helping toideateandinform where, and how, with what, and by whomsite reliabilitywill be achieved.

  • You enjoy wearingmultiplehats. At this stage inUnlimited’sgrowth,our thought leaders have an opportunity to speak into and influence our technology, operations,and processes.Having a green field to work in will mean as you build your role,you’lltouch on others’work andplay in others’ courts.Ifwe agreeon successfuloutcomes, how we get there will be acollectivestrategy andeffort.

  • You’recomfortable withshiftingpriorities, and some might sayyou’rean expert at context shifting. Not everyone can sustainthe exercise ofplanningstrategically butbuilding reactively.Not every day will present the requirement to shift gears, but some days will present the need to be tactically agile. When that moment comes,you’llnot only be prepared, butyou’llbe alsoready.

Qualifications

Required

  • 5+ years in an SRE, DevOps, or Platform Engineering role in a production cloud environment.

  • Hands-on AKS experience: cluster provisioning, upgrades, CNI networking, and workload lifecycle management.

  • Proficiencywith Splunk: SPL authoring, dashboard creation, alert configuration, and data onboarding.

  • Working knowledge ofInstanaAPM: agent deployment, custom tracing, alerting, and performance analysis.

  • Solid command of Azure Monitor, Log Analytics (KQL), and Application Insights.

  • Infrastructure-as-Code experience with Terraform; familiarity with Helm andKustomize.

  • Scripting or developmentproficiencyin at least one of: Python, Go, or Bash.

  • Demonstrated ability to drive incident resolution andleadblameless post-mortems.

  • Clear communicator able to translate complex technical topics for non-technical stakeholders.

Preferred

  • Microsoft Certified: Azure Administrator Associate (AZ-104) or Azure DevOps Engineer Expert (AZ-400).

  • Certified Kubernetes Administrator (CKA) or Certified Kubernetes Application Developer (CKAD).

  • Splunk Certified Power User or Splunk Enterprise Certified Architect.

  • Experience with service mesh technologies (Istio,Linkerd) on AKS.

  • Familiarity with FinOps practices and Azure cost-management tooling.

  • Experience in a regulated environment (SOC 2, ISO 27001, PCI-DSS, or HIPAA).

Our Process

Ifyou’vemade it this far in the job description,we’dlove to hear from you. We will try to ensure ourselectionprocess is efficient, transparent, andfriendly. (Ifit’snot, please let us know.)

What you can expect from here:

  • If we feel your experience aligns with our search,we’llreach out for a discovery call.We’lltalkabout our companyculture, Unlimited System’s Core Values,thisposition, andwe’lllearn more about your background.

  • From there,we’llasktoviewsome of yourpreviousworkas wellas request thatyoucompletean assessment.Thislets us know whether our expectationsfor the roleactuallyalignwith youruniquestrengths.

  • Ifwe’restill feeling good about each other,we’llpass you on to our Hiring Manager and then to a panel interview so you can meet more of the faces of Unlimited Systems before we make an offer decision.

EEO Statement

Unlimited Systems is an equal opportunity employer. We do not discriminate on the basis of race, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, veteran status, or any other protected characteristic under applicable law.

Disclaimer

Applicants may be subject to a background check. Employees in this position must be able to satisfactorily perform the essential functions of the position. If requested, Unlimited Systems will make every effort to provide reasonable accommodation to enable employees with disabilities to perform the position’s essential job duties. As markets change and the Organization grows, job descriptions may change over time as requirements and employee skill levels evolve. With this understanding, Unlimited Systems retains the right to change or assign other duties to this Senior Infrastructure Administrator position.

Powered by JazzHR

zNeY98sEHk