Join us as we inspire
creativity and bring joy to
millions of users worldwide.
@2026 TikTok
Responsibilities
Team Insight: The DataCenter Service / Datacenter Cloud System (DCS) team sits within TikTok's global technology structure and supports the company's fast growth by building and operating hyper-scale datacenters, managing the life cycle of server fleet, providing cloud solutions, and developing various infrastructure services, making sure they are scalable and are reliable. Role Insight: We are seeking an experienced Data Center Operations Engineers to apply technical expertise in a dynamic, fast-paced environment. This role requires strong knowledge of server hardware and a foundational understanding of mechanical and electrical infrastructure in large-scale data centers. You will be responsible for diagnosing and resolving server and infrastructure issues, collaborating with remote teams, and supporting the full server rack lifecycle, including buildout of compute and storage environments. In addition, you will lead and guide DC Technicians, ensuring high-quality execution of operational tasks and adherence to best practices. Candidates should have hands-on experience in at least one of the following areas: Networking, Scripting, or Hardware Repair. Success in this role requires strong communication skills, the ability to work independently and within a team, and adaptability in a rapidly changing environment. Responsibilities - Own day-to-day IT infrastructure operations across assigned data centre sites, ensuring uptime and SLA compliance for all break-fix activity. - Lead daily operations by guiding Data Center Technicians and overseeing task allocation, hardware troubleshooting, maintenance, and diagnostics. - Manage the full incident and change lifecycle, resolving incidents to SLA, executing changes under formal change control, and leading root cause analysis and corrective actions to closure. - Track and report IT operational performance metrics, including availability, MTTR, ticket ageing, first-time fix, break-fix volumes, and rack capacity utilisation, using trend analysis to drive continuous improvement. - Manage data centre IT on-site stability, identifying risk events and driving improvement initiatives to enhance stability and reduce operational errors. - Support data centre projects and operational improvements, including capacity expansions, retrofits, infrastructure upgrades, and implementation of new tools and processes; support new site builds and operational handover readiness. - Maintain and improve site documentation (SOPs, MOPs, EOPs, escalation matrices), build strong cross-functional relationships, act as an escalation point, and identify recurring issues to drive vendor escalation and continuous improvement.
Qualifications
Minimum Qualifications - 5+ years of experience in data centre operations, IT infrastructure, hardware support, and structured cabling in a live data centre environment. - Strong hands-on knowledge of server, storage, and network hardware, including component-level diagnostics and replacement. - Capable of server disassembly, hardware replacement, rack deployment, and hardware fault diagnosis with component-level troubleshooting across server and rack systems. - Basic Linux OS administration and BMC/BIOS configuration skills. - Basic understanding of critical facility infrastructure (UPS, generators, PDUs, CRAC/CRAH, cooling topology) to work safely alongside facilities teams. - Competent with DCIM, ticketing, and monitoring platforms, with the ability to analyse ticket data to identify recurring faults and support IT infrastructure performance. - Clear written and verbal communication skills, with the ability to report to both technical and non-technical stakeholders. Preferred Qualifications - Relevant certification such as CDCP, CDCS, ITIL Foundation, CompTIA Server+, or CCNA. - Experience in large-scale data centre environments. - Strong working knowledge of data centre power infrastructure and cooling systems, with proven competence in identifying server overheating risks and responding to power emergencies. - Hands-on experience operating and maintaining GPU servers and integrated rack systems, including hardware troubleshooting on GB300, B200, or comparable GPU platforms (e.g., GPU error analysis, NVLink fault diagnosis). -Experience coordinating vendors, outsourced technicians, or smart-hand resources on site. - Experience supporting projects, incident management, or small team leadership. - Willingness to participate in on-call rotation and travel to other sites as required.
Job Information
About TikTok
TikTok is the leading destination for short-form mobile video. At TikTok, our mission is to inspire creativity and bring joy. TikTok's global headquarters are in Los Angeles and Singapore, and we also have offices in New York City, London, Dublin, Paris, Berlin, Dubai, Jakarta, Seoul, and Tokyo.
Why Join Us
Inspiring creativity is at the core of TikTok's mission. Our innovative product is built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and bring joy - a mission we work towards every day.
We strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. Every challenge is an opportunity to learn and innovate as one team. We're resilient and embrace challenges as they come. By constantly iterating and fostering an "Always Day 1" mindset, we achieve meaningful breakthroughs for ourselves, our company, and our users. When we create and grow together, the possibilities are limitless. Join us.
Diversity & Inclusion
TikTok is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At TikTok, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.