About this role
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators.
At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there.
A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone.
As a Manager, Data Center Operations, you'll help us scale our Core Data Center and hardware infrastructure at a time of incredible growth for our business. At Roblox, you'll have boundless opportunities to shape the future of the Imagination Platform™ and demonstrate your passion for delivering thoughtful solutions in front of a global audience. If you know what it takes to build and operate hardware infrastructure that can sustain millions of concurrent players year-round and you take play as seriously as we do, you'll fit right into our highly experienced and ever-expanding engineering team. You will report to the Senior Manager of Data Center Operations. This will be a position based in Goodyear, AZ.
You will:
• Develop and maintain the Core Data Center and hardware infrastructure to meet the large-scale and real-time requirements of our Imagination Platform™ to ensure our community has an awesome experience anywhere in the world. This includes all aspects of the server, network infrastructure, power, and environmental monitoring.
• Lead a growing team of data center engineers focusing on rack deployments, hardware troubleshooting and break-fix, and decommissioning.
• Identify and solve critical problems and prevent them from re-occurring via root cause analysis and giving recommendations to improve automation. Guide, train and educate staff on best practices related to break/fix tasks related to the server hardware and network infrastructure.
• Create, influence, and improve the development platform, infrastructure, metrics, standards (Runbooks, SOPs, MOPs), and methods to ensure the goal of scalability and high availability can be achieved.
• Participate in the on-call rotation for our critical infrastructure.
• Build and implement Core sites around the world including low voltage cabling, creating BOMs, vendor management.
You have:
• At minimum 10+ years of experience working in large-scale Data Center Infrastructure environments and 3+ years experience leading a team of 3 or more data center engineers. While supporting the operations team, you are also expected to also perform as an individual contributor.
• Extensive experience installing, moni