Data Center Site Reliability Engineer
Data Center Site Reliability EngineerÂ
We Are
Synopsys is the leader in engineering solutions from silicon to systems, enabling customers to rapidly innovate AI-powered products. We deliver industry-leading silicon design, IP, simulation and analysis solutions, and design services. We partner closely with our customers across a wide range of industries to maximize their R&D capability and productivity, powering innovation today that ignites the ingenuity of tomorrow.
You Are
You are someone who takes ownership when things are complex, time-sensitive, and highly visible. You donât wait for perfect conditions or complete information. You create clarity by asking the right questions, confirming assumptions, and moving the work forward in a way that keeps others aligned and confident. You thrive in environments where priorities can shift quickly, and you stay steady under pressure by focusing on what matters most: protecting reliability, maintaining trust and delivering results that hold up over time.
You bring a practical, hands-on mindset to every day. You notice the small details that others overlook and you treat those details as signalsâearly warnings that help you prevent problems before they become incidents. You balance urgency with discipline, knowing when to act fast and when to slow down and verify. You take pride in leaving things better than you found them: cleaner, more organized, easier to understand, and easier to operate.
You work well as the person on the ground who makes progress visible. You communicate clearly and calmly, especially when coordinating across teams or guiding others through constraints. You document decisions so theyâre repeatable, and you follow through so commitments donât fade after the immediate issue is resolved. Above all, you care about operational excellenceânot as a slogan, but as a daily practice that keeps critical systems dependable and teams effective.
What You'll Be Doing
- Serve as the primary on-site technical resource supporting the Canonsburg data center and coordinating day-to-day operational needs with global engineering teams.
- Perform Linux systems administration tasks, including basic troubleshooting, log analysis, remote access support, and service management to keep critical systems running reliably.
- Install, rack, cable,relocate, and decommission physical server and networking equipment whilemaintainingclean, organized rack layouts and labeling standards.
- Maintainaccurateinfrastructure documentation and digital asset records in DCIM tools (e.g., Sunbird or similar), ensuring inventory, connectivity, and capacity datastayscurrent.
- Coordinate and oversee third-party vendors performing maintenance and infrastructure work, confirming scope, access requirements, safety practices, and completion criteria.
- Monitor data center health indicators such as power, cooling, rack capacity, and environmental conditions, escalatingrisksandinitiatingcorrective actions as needed.
- Respond to operational incidents as part of a shared on-call rotation, meeting establishedresponseSLAs and driving issues through resolution and follow-up.
The Impact You Will Have
- Keep mission-critical engineering systems dependable by reducing downtime and restoring service quickly when issues arise.
- Improve day-to-day operational confidence throughaccurateasset records and documentation that make capacity, ownership, and change planning clear.
- Accelerate infrastructure deployments by ensuring on-site execution istimely, consistent, and aligned with global engineering standards.
- Reduce operational risk byidentifyingearly warning signs in power, cooling, and environmental conditions and driving corrective actions before they become incidents.
- Strengthen vendor outcomes by ensuring work is properly scoped, safely executed, and fully completed with clear validation and follow-through.
- Increase cross-team effectiveness by serving as a reliable on-site partner who communicates clearly, escalates appropriately, and closes loops after changes and incidents.
- Support successful data center integration and modernization by helping standardize processes andstabilizingoperations during periods of change.
What You'll Need
- You have experience administering Linux systems (Red Hat, Rocky Linux, Ubuntu, or similar) and can navigate common operational tasks with confidence.
- You bring hands-on familiarity working in a physical enterprise data center environment, where safety, precision, and process matter.
- You have a working understanding of server hardware, rack infrastructure, structured cabling, power distribution, and cooling fundamentals.
- You bring strong troubleshooting and problem-solving habits, including the ability to stay calm, prioritize effectively, and drive issues to resolution.
- You have experience coordinating with third-party vendors and service providers and can ensure work is completed to scope and standard.
- Youare able towork independently as the primary on-site technical resource and communicate clearly with remote engineering partners.
- You bring differentiators such as exposure to DCIM platforms (e.g., Sunbird), network/storage/virtualization environments, and/or scripting and automation with Bash, Python, or PowerShell.
Who You Are
- You are the kind of person who takes ownership end-to-end, following through until the issue is fullyresolvedand the next steps are clear to everyone involved.
- You approach problems methodically, separating symptoms from root causes and validating changes before and after you act.
- You communicate with precision, tailoring updates to the audience and escalating early when risk, safety, or SLA impact is on the line.
- You stay organized in fast-moving environments, keeping documentation, labels, and recordsaccurateso others canoperateconfidently after you.
- You collaborate smoothly across teams and vendors, setting expectations upfront and ensuring work is completed safely, cleanly, and to standard.
The Team You'll Be Part Of
The Data Center Operations team supports reliable, secure day-to-day execution across Synopsysâ physical infrastructure footprint while partnering closely with global network, storage, security, and infrastructure engineering teams. This role serves as the primary on-site operator for the Canonsburg data center, ensuring operational excellence and enabling modernization and integration efforts.
Rewards and Benefits
We offer a comprehensive range of health, wellness, and financial benefits to cater to your needs. Our total rewards include both monetary and non-monetary offerings. Your recruiter will provide more details about the salary range and benefits during the hiring process.