Staff Site Reliability Engineer
Who We Are
KEV builds mission-critical financial software for K12 schools across North America. Our platform provides real-time visibility and control over how funds are collected and managed, replacing fragmented workflows with secure, modern systems that schools rely on every day.
Trusted by more than 27,000 K12 schools managing over $8B annually, KEV delivers mission-critical software for payments, accounting, and reporting, where reliability and security matter. Headquartered in Toronto with teams across North America, we are scaling quickly and investing deeply in cloud-native, data-driven technology.
KEV is a place for people who care about building durable software at scale and doing work that has real impact in public education. We focus on solving hard problems, raising the bar on engineering craft, and building products that earn long-term trust from the communities we serve.
This role is for an active vacancy.
About the Role
We are hiring a Staff Site Reliability Engineer to raise the bar on how reliably KEV's platforms run in production, from the SLOs and error budgets that define what “reliable” means for each service to the incident practice that keeps issues from repeating. This role requires deep hands-on SRE experience and the judgment to assess where reliability is at risk, shape a strategy for closing the gap, and break that strategy into work a team can execute.
You will partner closely with KEV's DevOps Engineering team and product engineering teams to embed reliability into how services are built and operated, and coach and mentor other engineers on reliability practices. This is a role for someone who can surface trade-offs clearly and communicate them well.
How You Will Contribute
- Reliability Strategy & SLOs Assess the current state of reliability across KEV's platforms, define and evolve the SLOs, SLIs, and error budgets that make “reliable” a measurable target rather than a feeling, and build a strategy for closing the gaps you find. Break that strategy into concrete workstreams and lead a team through execution.
- Incident Management & Postmortem Practice Own how KEV responds to production incidents, driving blameless postmortems that get to root cause rather than surface symptoms. Use recurring incidents and operational issues to identify durable fixes and longer-term reliability investments, not just one-off patches.
- Observability & Monitoring Platform Design and evolve the monitoring, alerting, logging, and tracing platform that gives engineering teams real visibility into system health. Use data to reduce noisy or ineffective alerts and make sure the signals teams rely on are ones they can trust.
- Automation & Toil Reduction Identify repetitive, manual operational work and replace it with automation and tooling that's maintainable, testable, and treated as production code. Build the tools yourself where that's the right answer, rather than only advocating for automation from the sidelines.
- Capacity & Performance Engineering Lead capacity planning and performance analysis so that KEV's platforms scale ahead of demand rather than in reaction to it. Use load testing and performance data to anticipate where systems will break before they do.
- AI-Assisted Reliability Operations Evaluate and apply AI tooling where it can meaningfully improve reliability and operational workflows — such as anomaly detection or incident triage — applying the same trade-off thinking to new tooling that you'd apply to any other operational decision.
- Coaching & Reliability Mentorship Coach and mentor other engineers on reliability practices — on-call hygiene, incident response, SLO-driven prioritization — raising the operational maturity of the broader engineering organization over time.
- Cross-Functional Partnership Partner closely with KEV's DevOps Engineering team and product engineering teams to embed reliability into how services are designed, built, and operated. Communicate trade-offs clearly so stakeholders understand what's changing, why, and what was weighed to get there.
Who You Are
- 10+ years of hands-on experience in Site Reliability Engineering, Platform Engineering, or a similar infrastructure/operations role, with a track record of leading reliability strategy and incident practice at scale.
- Familiarity working with Microsoft Azure and .NET, including legacy .NET Framework applications and IIS, so reliability practices hold up across KEV's full stack and not just its newest services.
- Demonstrated experience defining and operating SLOs, SLIs, and error budgets, and using them to guide engineering and operational priorities.
- Experience owning incident management processes and driving blameless postmortems that lead to durable fixes rather than one-off patches.
- Hands-on experience designing and building observability platforms — monitoring, alerting, logging, and tracing — that give engineering teams real visibility into system health.
- Strong automation skills, with experience building tools and scripts that eliminate repetitive manual operational work and are treated as production code.
- Experience with capacity planning and performance engineering, anticipating and preventing reliability issues rather than only reacting to them.
- Experience evaluating and applying AI tooling to improve reliability and operational workflows.
- Demonstrated ability to coach and mentor engineers on reliability practices, raising the operational maturity of a team over time.
- Strong ability to surface trade-offs and communicate them clearly to both engineering and cross-functional stakeholders.
What We Offer
- Competitive compensation – We believe in rewarding great work with fair, competitive pay.
- Meaningful benefits– Because your well-being matters; both at work and at home.
- Retirement Savings Support – We help you plan for your future with company-matched programs, including RRSP matching in Canada and 401(k) contributions in the U.S.
- Professional development – We invest in your growth with ongoing learning, stretch opportunities, and continuing education, including KEV Academy for onboarding and skill-building, plus KEV University, our online platform offering a wide range of courses.
- Hybrid model – 3 days in the office to collaborate and connect, with flexibility the rest of the week.
- Flexible PTO – Take the time you need to recharge with close to 4 weeks of vacation and a company-wide holiday closure
- Office perks – Enjoy a fully stocked snack bar and occasional catered lunches—because we know that great conversations (and ideas) often start around good food.
Salary Range - $150,000-180,000
Note: We set standard base pay ranges for all roles based on function, level, and country location, benchmarked against similar stage growth companies. Final offer amounts are determined by multiple factors, including skills, depth of work experience and relevant licenses/credentials, and may vary from the amounts listed above.
Why You’ll Love Working at KEV
- Make a real difference every day – At KEV, your work supports children, parents, and schools across North America. We don’t just build software, we create solutions that simplify lives and strengthen communities. Our mission is rooted in impact, and every team member plays a vital role in shaping the future of K-12 education.
- The KEV Way – At KEV, you’ll never feel stuck in the status quo. You’ll be part of a team that’s constantly questioning, improving, and innovating—always with one guiding focus: How does this help our schools and the students they serve? It’s a culture that challenges you to do your best work while reminding you why it matters.
- Grow with us – We’re scaling fast, and so are our people. At KEV, you’ll have real opportunities to learn, develop, and shape your career. Whether advancing in your role or exploring a new path, you’ll be supported every step of the way.
- Work alongside excellence – You’ll be part of a team with integrity who treat their colleagues with respect. who are accomplished but humble, collaborative but accountable. It’s a place where you can do your best work while feeling supported and inspired by the people around you.
- Celebrate Community and Culture – At KEV, we connect, recognize, and celebrate our people. Join Club KEV for team-building fun, hear directly from customers in our Voice of Customer Series, stay aligned with Monthly Townhalls, and be inspired by the KEVite Awards, where top contributors are recognized by their peers.
This job description indicates the general nature and level of work expected. It is not designed to cover or contain a comprehensive listing of activities, duties, or responsibilities required by an individual joining the KEV team in this or any other capacity.
KEV Group is pleased to accommodate individual needs in accordance with the Accessibility of Ontarians with Disabilities Act, 2005 (AODA), within our recruitment process. If you require accommodation at any time throughout the recruitment process, please speak with your recruiter.
KEV Group is an equal opportunity employer who agrees not to discriminate against any employee or job applicant because of race, color, religion, national origin, sex, physical or mental disability, or age.
KEV may use AI-enabled tools to support components of our recruitment process. These tools are used to support human decision making. All hiring decisions are made by our People and Culture team in partnership with Hiring Managers.
Visit our website for more information and details about working at KEV.
About the Company
KEV Group is a K-12 school finance software company that provides a platform for managing school fees, payments, and accounting. It offers real-time visibility over every fee, fund, and payment, helping districts and schools track and control their finances. The platform integrates with existing systems like ERP, SIS, nutrition, asset, and library systems, and is trusted by over 1,000 districts and 28,000 schools in North America. KEV emphasizes audit readiness, fraud prevention, and bank-grade security (SOC 2 Type II, PCI Level 1, FERPA-compliant).
More jobs at KEV Group
-
Staff DevOps
Toronto, Ontario, Canada · Hybrid · Sep 15, 2026
-
Staff Frontend Engineer
Toronto, Ontario, Canada · Hybrid · Sep 15, 2026
-
Sr. Test Automation Engineer
Toronto, Ontario, Canada · Hybrid · Aug 24, 2026
-
Senior Data Analyst (6 Month Contract)
Kitchener, Ontario, Canada; Toronto, Ontario, Canada · Contract · Aug 24, 2026
-
DevOps
Toronto, Ontario, Canada · Hybrid · Aug 19, 2026