Senior Site Reliability Engineer
SYNAPSE HEALTH · Remote
📍 Remotevia greenhousePosted 2026-09-21
Apply on company site ↗
Career Moonshot pulls this listing straight from the employer's hiring system — no recruiter middleman, no reposts. Applying takes you directly to SYNAPSE HEALTH.
Who We Are :
At Synapse Health, we're streamlining the durable medical equipment (DME) process. We manage intake, documentation, routing, claims, billing, and patient support. Our model reshapes how DME is delivered and experienced.
Since 2016, with decades of industry and leadership experience, we've delivered tech-based solutions that help our partners to modernize operations, improve coordination, and reduce administrative burdens. By taking on operational and financial complexity, we're redefining how DME works for providers, prescribers, and patients. We are proud to offer work that matters, on a mission that matter s .
Learn more at SynapseHealth.com and on Synapse Health’s LinkedIn .
What We Need :
We’re at a pivotal moment in both our company growth and platform evolution. As we scale, we are actively transforming our infrastructure to support a more modern, containerized, and highly scalable architecture.
We’re seeking a Sr. Site Reliability Engineer (SRE) to help lead that transformation. This role will play a critical part in evolving our platform from legacy Azure-based services toward a Kubernetes-driven, microservices-oriented environment.
As a senior member of the team, you will take ownership of complex, ambiguous infrastructure challenges and drive them through to practical, scalable solutions. You’ll partner closely with engineering and data teams to ensure reliability, performance, and scalability are built into our systems from the ground up.
This is an ideal opportunity for an engineer who thrives in fast-paced environments, enjoys solving real infrastructure problems, and wants to have a direct impact on the technical direction of a growing healthcare platform.
What You Will Do :
Platform & Infrastructure Evolution
Contribute to the migration from legacy Azure services and function-based architectures to containerized, microservices-based systems
Help design, build, and scale Kubernetes-based infrastructure and supporting tooling
Partner with engineering teams to ensure new systems are designed for reliability, scalability, and operational efficiency from day one
Drive standardization across infrastructure to reduce silos and enable broader team ownership
Cost Optimization: Monitor cloud usage and spending, identify inefficiencies, and recommend and implement cost optimization strategies
Networking & Connectivity: Design, deploy, and manage secure networking, including public and private endpoints, environment segmentation, site-to-site and point-to-site VPNs, and inter-environment connectivity.
Reliability & Observability
Design and maintain highly available , resilient systems in a cloud-native environment
Implement and evolve observability practices including monitoring, alerting, and logging (e.g., Datadog, Prometheus, Grafana)
Define and manage SLIs, SLOs, and SLAs aligned to system performance and user experience
Lead incident response efforts and drive root cause analysis and long-term improvements
Automation & Developer Enablement
Build and optimize CI/CD pipelines to support fast, safe, and repeatable deployments
Champion Infrastructure-as-Code practices using tools such as Terraform to eliminate manual processes
Leverage scripting (Python, Bash, or similar) to solve problems, automate workflows, and reduce operational toil
Performance, Scalability & Strategy
Drive capacity planning, performance tuning, and infrastructure improvements to support rapid growth
Proactively identify system risks and scalability bottlenecks before they impact customers
Contribute to infrastructure strategy and help shape how the platform evolves as the business scales
Knowledge Sharing & Team Enablement
Document systems, processes, and best practices to improve team-wide reliability and reduce single points of failure
Contribute to cross-training efforts as the team moves toward broader ownership and standardization
Share knowledge and elevate the team through mentorship and collaboration
Note: These responsibilities reflect the general nature and scope of the role but are not exhaustive. Responsibilities may evolve to meet changing business needs.
What You Have :
At Synapse Health, we’ve intentionally built a culture rooted in kindness, collaboration, and creativity, qualities we consider essential for every team member. Additional requirements include:
5+ years of experience in Site Reliability Engineering, DevOps, or Infrastructure Engineering
Hands-on experience working in cloud environments (Azure preferred, AWS or GCP acceptable)
Strong experience with Kubernetes in production environments
Experience deploying and managing applications on Kubernetes using Helm
Proficiency with Infrastructure-as-Code tools such as Terraform
Strong scripting skills (Python, Bash, or similar) used to automate and solve infrastructure challenges
Experience with observability, monitoring, and incident response in production environments
Experience building or supporting CI/CD pipelines ( GitHub Actions and/or GitLab CI/CD; experience with CI/CD migrations is a plus )
Solid understanding of networking fundamentals, system design, and cloud infrastructure components
Familiarity with Azure Entra ID, app registrations, federated identity credentials, and workload identity
Proven ability to take loosely defined problems and drive them to practical, scalable solutions
Strong communication skills and ability to collaborate across engineering and non-technical stakeholders
Comfort operating in fast-paced environments with evolving priorities
Understanding of PHI handling requirements, access control patterns, and audit controls in a healthcare environment.
What Sets Y
More Remote jobs
Remote jobs · Browse all locations