Principal Software Engineer-SRE
PTCAbout the role
Our world is transforming, and PTC is leading the way. Our software brings the physical and digital worlds together, enabling companies to improve operations, create better products, and empower people in all aspects of their business.
Our people make all the difference in our success. Today, we are a global team of nearly 7,000 and our main objective is to create opportunities for our team members to explore, learn, and grow – all while seeing their ideas come to life and celebrating the differences that make us who we are and the work we do possible.
Principal Software Engineer (SRE)-Onshape-Remote US.
About the Role
Onshape’s Site Reliability Engineering team is looking for a Principal Software Engineer to play a critical role in ensuring the long‑term reliability, scalability, and operational excellence of our platform.
As a Principal Software Engineer, you will operate with a high degree of autonomy and influence. You will lead complex, cross‑organization reliability initiatives, shape reliability strategy, and serve as a technical authority and trusted advisor across engineering.
Your work will directly shape the experience of our customers by ensuring the platform is fast, resilient, and dependable. As a Principal Software Engineer, you will help protect customer trust by driving reliability across the entire system lifecycle.
This role is ideal for engineers who enjoy solving ambiguous, high‑impact problems at scale, influencing system design across teams, and raising the reliability bar for an entire organization.
What You’ll Do:
Own Reliability at Scale
Lead design, implementation, and evolution of reliability, availability, and resiliency strategies for large‑scale distributed systems written primarily in Java
Apply deep experience operating complex, distributed systems to guide architectural decisions, reliability strategies, and long‑term system evolution
Identify systemic risks in application architecture, data flows, and infrastructure, and drive architectural improvements that measurably improve availability, performance, and scalability
Set and evolve reliability standards, best practices, and operational principles across R&D
Drive Operational Excellence
Lead efforts to prevent, detect, and mitigate incidents through technical improvements and operational maturity
Serve as a senior coordination point during major incidents, helping manage response and guide long‑term remediation
Champion blameless post-incident reviews and ensure learnings translate into durable system improvements
Reduce Toil Through Engineering
Apply advanced software engineering practices to eliminate manual work, reduce operational load, and improve system observability
Design and build internal platforms, automation, and tooling that support Java‑based services and their operational needs
Raise the bar on monitoring, alerting, and SLO/SLI adoption across systems
Apply for this role
Generate a tailored application kit with a matched cover letter, interview prep, and CV highlights — in under 60 seconds.
Apply Now →Generate Application KitFree account required — sign up in 30s