Incident Response Manager - San Jose
San Jose, California, United States
TikTok is the leading destination for short-form mobile video. At TikTok, our mission is to inspire creativity and bring joy. TikTok's global headquarters are in Los Angeles and Singapore, and its offices include New York, London, Dublin, Paris, Berlin, Dubai, Jakarta, Seoul, and Tokyo.
Why Join Us
Creation is the core of TikTok's purpose. Our platform is built to help imaginations thrive. This is doubly true of the teams that make TikTok possible.
Together, we inspire creativity and bring joy - a mission we all believe in and aim towards achieving every day.
To us, every challenge, no matter how difficult, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always.
At TikTok, we create together and grow together. That's how we drive impact - for ourselves, our company, and the communities we serve.
Join us.
About the team
The Data Systems Infrastructure (DSI) team sits within the ByteDance global technology structure and supports the company's fast growth by building and operating hyper-scale datacenters, managing the life cycle of server fleet, providing cloud solutions, and developing various infrastructure services, making sure they are scalable and are reliable.
The Incident Response Center (IRC) is the first layer of defense responsible for quick detection and incident response using various monitoring and automation tools, conducting thorough investigation of alerts, classification and triage. The Incident Response Manager is responsible for delivering operations within the IROC across all ByteDance datacenter sites in the respective regions. IRC team is expected to respond to all alarms/alerts set in Server Automation Operations System (SAOS), Data Center Infrastructure Management (DCIM) to quickly discover anomalies and engage Subject Matter Expert (SME) teams to start issue triage. The IRC team provides business intelligence through rigorous analysis of alerts and issues which reduce and prevent recurring incidents .
Responsibilities
- Delivering global operations within the IROC (Incident Response Operation Center) ByteDance datacenter.
- First responder and layer of defense responsible for quick detection and incident response using various monitoring and automation tools, conduct thorough investigation of alerts, classification and triage.
- Respond to all infrastructure, facilities, security, and safety events notified via various means, such as alarms/alerts set in Server Operations and Maintenance, Datacenter Infrastructure Management, Network & Grafana, and other functions.
- Respond to incidents and critical situations in a problem-solving manner, and conduct in-depth investigation of alerts.
- Provide insights into the effectiveness of the incident response and recovery process through regular reports
- Analyze trends and patterns in events to identify opportunities for improvement and optimization
- Monitor the performance of incident response against the agreed-upon SLAs by alerting and notifying stakeholders
- Escalation Management notifying or initiating discussions with higher-level support teams engaging in resolution processes
- Identify, assess and communicate potential risks arising through event monitoring that could affect customer's service
- Support program managers and facilitate project deliverables, improve overall operational security and engineering initiatives
- The Incident Response team is expected to work at ByteDance datacenter site. This is an on-site role.
Why Join Us
Creation is the core of TikTok's purpose. Our platform is built to help imaginations thrive. This is doubly true of the teams that make TikTok possible.
Together, we inspire creativity and bring joy - a mission we all believe in and aim towards achieving every day.
To us, every challenge, no matter how difficult, is an opportunity; to learn, to innovate, and to grow as one team. Status quo? Never. Courage? Always.
At TikTok, we create together and grow together. That's how we drive impact - for ourselves, our company, and the communities we serve.
Join us.
About the team
The Data Systems Infrastructure (DSI) team sits within the ByteDance global technology structure and supports the company's fast growth by building and operating hyper-scale datacenters, managing the life cycle of server fleet, providing cloud solutions, and developing various infrastructure services, making sure they are scalable and are reliable.
The Incident Response Center (IRC) is the first layer of defense responsible for quick detection and incident response using various monitoring and automation tools, conducting thorough investigation of alerts, classification and triage. The Incident Response Manager is responsible for delivering operations within the IROC across all ByteDance datacenter sites in the respective regions. IRC team is expected to respond to all alarms/alerts set in Server Automation Operations System (SAOS), Data Center Infrastructure Management (DCIM) to quickly discover anomalies and engage Subject Matter Expert (SME) teams to start issue triage. The IRC team provides business intelligence through rigorous analysis of alerts and issues which reduce and prevent recurring incidents .
Responsibilities
- Delivering global operations within the IROC (Incident Response Operation Center) ByteDance datacenter.
- First responder and layer of defense responsible for quick detection and incident response using various monitoring and automation tools, conduct thorough investigation of alerts, classification and triage.
- Respond to all infrastructure, facilities, security, and safety events notified via various means, such as alarms/alerts set in Server Operations and Maintenance, Datacenter Infrastructure Management, Network & Grafana, and other functions.
- Respond to incidents and critical situations in a problem-solving manner, and conduct in-depth investigation of alerts.
- Provide insights into the effectiveness of the incident response and recovery process through regular reports
- Analyze trends and patterns in events to identify opportunities for improvement and optimization
- Monitor the performance of incident response against the agreed-upon SLAs by alerting and notifying stakeholders
- Escalation Management notifying or initiating discussions with higher-level support teams engaging in resolution processes
- Identify, assess and communicate potential risks arising through event monitoring that could affect customer's service
- Support program managers and facilitate project deliverables, improve overall operational security and engineering initiatives
- The Incident Response team is expected to work at ByteDance datacenter site. This is an on-site role.
* Salary range is an estimate based on our InfoSec / Cybersecurity Salary Index 💰
Job stats:
0
0
0
Categories:
Incident Response Jobs
Leadership Jobs
Tags: Automation Business Intelligence Cloud Grafana Incident response Monitoring SLAs
Perks/benefits: Team events
Region:
North America
Country:
United States
More jobs like this
Explore more career opportunities
Find even more open roles below ordered by popularity of job title or skills/products/technologies used.
Senior Security Analyst jobsSenior Cloud Security Engineer jobsInformation System Security Officer jobsInformation Security Manager jobsSenior Cybersecurity Engineer jobsSenior Network Security Engineer jobsInformation Security Specialist jobsSecurity Consultant jobsSenior Penetration Tester jobsSecurity Specialist jobsSenior Information Security Analyst jobsSenior Cyber Security Engineer jobsIT Security Engineer jobsCyber Security Specialist jobsChief Information Security Officer jobsIT Security Analyst jobsPrincipal Security Engineer jobsStaff Security Engineer jobsCloud Security Architect jobsInformation System Security Officer (ISSO) jobsCyber Security Architect jobsSenior Product Security Engineer jobsSystems Engineer jobsSenior Information Security Engineer jobsSecurity Operations Analyst jobs
CI/CD jobsSaaS jobsForensics jobsMalware jobsEncryption jobsEDR jobsTop Secret jobsSplunk jobsSDLC jobsIDS jobsIPS jobsSQL jobsRMF jobsCompTIA jobsBash jobsIntrusion detection jobsDocker jobsFinance jobsThreat detection jobsDoDD 8570 jobsOWASP jobsITIL jobsActive Directory jobsTCP/IP jobsCRISC jobs
Terraform jobsVPN jobsGIAC jobsSANS jobsUNIX jobsBanking jobsHIPAA jobsIT infrastructure jobsClearance Required jobsJavaScript jobsSOX jobsAnsible jobsPolygraph jobsDNS jobsCCSP jobsJira jobsData Analytics jobsMITRE ATT&CK jobsSOC 2 jobsOSCP jobsGCIH jobsCISO jobsSOAR jobsMachine Learning jobsCyber defense jobs