Staff Site Reliability Engineer
FinalsiteWHO WE ARE
Finalsite is the most valued partner for K–12 schools to build trust, strengthen community, and grow enrollment. Ranked among the best EdTech Companies in America, Finalsite supports more than 7,000 schools and districts worldwide with an integrated platform for websites, communications, mobile apps, enrollment, and marketing services.
Headquartered in Glastonbury, Connecticut, Finalsite is a global company with employees working remotely across nearly every U.S. state, as well as throughout Europe, South America, and Asia.
We believe people do their best work when they feel supported, connected, and empowered to grow. That’s why we invest in our employees through competitive benefits, professional development opportunities, and a collaborative culture built on partnership and purpose. Whether you’re looking to expand your skills, take on new challenges, or make a meaningful impact in education, Finalsite offers the opportunity to grow your career while helping schools thrive.
At Finalsite, every interaction matters — with our clients, with each other, and with the schools and families we serve. Join us and help shape stronger school communities around the world
SUMMARY
As a Staff Site Reliability Engineer, you'll focus on the platform behind Finalsite's Composer CMS, while also working cross-team to shape platform-wide standards and integrations. You'll help set the technical direction for how we build, scale, and operate our platform in a GCP-primary, multi-cloud environment. You'll be the person other engineers come to when a system needs to be rethought, not just repaired, and you'll play a key role in growing the senior engineers around you. This is a role for someone who wants to know exactly how the systems work, how they will fail, and how to build the guardrails that keep everyone else from finding out the hard way.
LOCATION
100% Remote - Anywhere within the US
WHAT WE LOOK FOR
- Cloud Architecture & Scalability. You guide our cloud architecture strategy with a focus on scalability, maintainability, and cost efficiency, and you lead capacity planning so we scale ahead of demand instead of reacting to it.
- Kubernetes & Container Platform Ownership. You own the health of our Kubernetes (GKE) platform, from cluster architecture to workload reliability at scale.
- Network & Edge Architecture. You design and evolve our network architecture, including global routing, connectivity, and edge strategy with Cloudflare.
- Infrastructure as Code. You drive IaC standards across the team, building reusable Terraform/Terragrunt modules that other teams can adopt without reinventing them.
- Observability Mindset. You build and mature our observability practice, defining what we monitor, how we alert, and how we define and hold ourselves to SLOs.
- Focus on Reliability. You lead disaster recovery planning for the platforms you own, including backup design, failover planning, and clear recovery objectives (RTO/RPO), and you architect highly available, fault-tolerant systems.
- Security-First Engineering. You bring a security-first mindset to everything you build, treating it as a design input from day one, not a review gate at the end.
- Cost Optimization. You keep a close eye on cloud cost and help teams make smart tradeoffs between performance, resilience, and spend.
- Code Literacy. You can read, understand, and write basic code when needed, across languages and runtimes such as Ruby/Rails, Java, Python, for example. You're comfortable in application code to spot reliability, performance, or architecture issues.
HOW YOU'LL WORK
- Team First. You mentor senior engineers, helping them grow their technical judgment and take on bigger calls of their own.
- Developer Enablement. You build tools and patterns that let other teams move independently, turning one-off solutions into lasting, self-serve practice, so SRE isn't the bottleneck standing between people and the answer.
- Collaborative Planning. You represent infrastructure and reliability concerns in planning conversations, weighing in early enough to shape decisions, not just implement them.
- Change Management. You champion strong change management practices, including peer review, staged rollouts, and go/no-go gates for high-risk changes.
- Incident Response. You lead response for high-severity incidents, participate in our on-call rotation, and turn every incident into a lasting improvement.
WE'D LIKE YOU TO HAVE
- GCP & Infrastructure Expertise. Expert-level knowledge of GCP and infrastructure broadly, comfortable operating across the full stack rather than one layer of it.
- Infrastructure as Code. Expertise in IaC, including Terraform and Terragrunt.
- CI/CD Pipelines. Experience with GitLab (or similar) and CI/CD pipeline design and operation.
- Kubernetes & Containerization. Expertise in Kubernetes and containerization, including Helm.
- Network Architecture. Strong experience with network architecture, including routing and connectivity within GCP and across cloud providers.
- Edge Security. Strong experience with edge security and traffic management, using solutions like Cloudflare.
- Observability. Deep familiarity with observability tooling and monitoring strategy.
- Incident Leadership. Real experience leading incident management and response, not just participating in it.
- Staff-Level L6 Track Record. A track record of operating at a staff level, including mentoring senior engineers.
- High Availability & Fault Tolerance. Experience architecting highly available, fault-tolerant systems.
- Cost Allocation (Preferred). Experience with cloud cost allocation and control.
- Capacity Planning (Preferred) . Experience with capacity planning and forecasting for infrastructure at scale.
- Multi-Cloud Familiarity (Preferred). Familiarity with AWS and/or Azure (GCP-primary).
RESIDENCY REQUIREMENT
Finalsite offers 100% fully remote employment opportunities, however, these opportunities are limited to permanent residents of the United States. Current residency, as well as continued residency, within the United States is required to obtain (and retain) employment with Finalsite.
DISCLOSURES
Finalsite is proud to be an equal opportunity workplace and is an affirmative action employer. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. We also consider qualified applicants regardless of criminal histories, consistent with legal requirements. EEO is the Law. If you have a disability or special need that requires accommodation, please contact Finalsite's People Operations Team. Finalsite is committed to the full inclusion of all qualified individuals. As part of this commitment, Finalsite will ensure that persons with disabilities or special needs are provided a reasonable accommodation. Ensure your Finalsite job offer is legitimate and don't fall victim to fraud. Ask your recruiter for a phone call or other type of verbal communication and ensure all email correspondence is from a finalsite.com email address. For added security, where possible, apply through our company website at finalsite.com/jobs.