NextPath Jobs

Sr. Site Reliability Engineer

CentralReach · Holmdel, New Jersey

Posted 2026-09-15 · Verified live 2026-09-28

See how your real experience scores against this role — our AI drafts an honest, verifiable resume tailor. Nothing invented, ever. You approve everything.

Get matched — join the waitlist Apply on company site ↗

About this role

<p>CentralReach is a leading provider of autism and IDD care software for Applied Behavior Analysis (ABA), multidisciplinary therapy, and special education. Trusted by more than 200,000 users, we enable therapy providers, educators, and employers to scale the way they deliver ABA and related therapies with innovative technology, market-leading industry expertise, and world-class customer satisfaction.&nbsp;</p><p><em>The&nbsp;Platform Engineering&nbsp;group at&nbsp;CentralReach&nbsp;builds the underlying technologies that power our Public and Private Cloud Platforms worldwide. The group is responsible for storage, data infrastructure, IT, observability systems, DevOps, SRE, provisioning, compute, orchestration platform, internal tools, internal platforms (laptops, networks, systems etc.) and services - all the components that make up the&nbsp;CentralReach&nbsp;Platform.</em>&nbsp;</p>

<p><em>If you have a passion for the future, enjoy and thrive in an agile, fast-moving, ever-changing startup environment, welcome and take on technical challenges of all shapes and sizes, have excellent interpersonal skill and sense of humor and enjoy rolling up your sleeves and jumping in, then read on!&nbsp;</em>&nbsp;</p>

<p><em>As a Sr. SRE, you will work closely with the key stakeholders in Software Engineering to drive adoption of modern reliability practices like SLOs, error budget policies, actionable alerts, incident retrospectives, chaos testing, and end-to-end ownership.</em>&nbsp;</p>

<p><strong>Key Accountabilities:</strong>&nbsp;</p>

<ul>

<li>Own production reliability, including availability, latency, performance, capacity planning, monitoring, emergency response, and uptime for production environments.&nbsp;</li>

</ul>

<ul>

<li>Define,&nbsp;maintain, and improve SLOs, SLIs, error budgets, actionable dashboards, and observability practices.&nbsp;</li>

</ul>

<ul>

<li>Analyze, troubleshoot, and resolve operational issues that affect service reliability and SLO performance.&nbsp;</li>

</ul>

<ul>

<li>Build and automate multi-environment observability capabilities, including capacity forecasting based on usage patterns.&nbsp;</li>

</ul>

<ul>

<li>Reduce toil and increase development velocity through automation and continuous improvement.&nbsp;</li>

</ul>

<ul>

<li>Provide production support, including incident, change, and problem management; root cause analysis; service restoration; runbooks; and standard operating procedures.&nbsp;</li>

</ul>

<ul>

<li>Identify&nbsp;data-driven opportunities to improve system architecture, availability, performance, and reliability.&nbsp;</li>

</ul>

<ul>

<li>Collaborate with software engineering teams on release management, roadmap planning, and operational readiness.&nbsp;</li>

</ul>

<ul>

<li>Implement and manage reliability and observability tools such as Datadog, Prometheus, and Grafana.&nbsp;</li>

</ul>

<p>&nbsp;</p>

<p><strong>Desired Skills and Experience:</strong>&nbsp;</p>

<ul>

<li>Experience with monitoring, APM, and observability tools such as Splunk, Prometheus, Datadog, and&nbsp;OpenTelemetry.&nbsp;</li>

</ul>

<ul>

<li>Experience implementing observability strategies for logs, metrics, and traces.&nbsp;</li>

</ul>

<ul>

<li>Strong understanding of CI/CD practices and tools such as Jenkins, GitHub Actions, GitLab, Argo, and Kargo.&nbsp;</li>

</ul>

<ul>

<li>Strong understanding of major cloud providers, preferably AWS, and cloud-native infrastructure concepts.&nbsp;</li>

</ul>

<ul>

<li>Strong understanding of containerization technologies, including Kubernetes and Helm.&nbsp;</li>

</ul>

<ul>

<li>Experience with one or more programming languages, such as Java, Python, or Go, and familiarity with .NET application development.&nbsp;</li>

</ul>

<ul>

</ul>

One of thousands of fresh listings refreshed nightly, built for cleared & defense careers.

Browse all jobs