Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
TikTok USDS logo

Senior Site Reliability Engineer, Product - USDS

TikTok USDS
  • ๐Ÿ‡บ๐Ÿ‡ธ United States
  • On-site
  • Senior
  • 5 months ago
  • Unix
  • Linux
  • TCP/IP
  • C++
  • Java
  • Python
  • Ruby
  • Rust
  • JavaScript
  • NGINX
  • Kubernetes
  • Docker
  • OpenStack
  • Hadoop
  • Flink
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV

About the team

The Product Engineering team monitors and maintains the availability of TikTok, including services such as video playback, content discovery/recommendations, live streaming, and customer service feedback.

In this role, you will

  • Gain a solid understanding of the various components and services that power the TikTok experience
  • Maintain services to meet service-level-agreements (SLAs) and service-level-objectives (SLOs) by measuring and monitoring availability, performance, and overall system health
  • Participate as part of a global team to support site-up issues to ensure that services are reliable, fault-tolerant, efficiently scalable and cost-effective
  • Scale systems sustainability through mechanisms such as automation; evolve systems reliability, efficiency, and velocity by pushing for changes
  • Provide user support, incident responses and postmortems

Minimum Qualifications

  • Bachelor's or above degree in Computer Science or a related technical discipline with 5+ years experience in the deployment and administration of large-scale distributed systems
  • Strong understanding of Unix/Linux operating systems internals and administration, networking (e.g. TCP/IP, routing, network topologies and hardware), storage systems, and database systems
  • Experience in one or more programming languages, such as C, C++, Java, Python, Go, Ruby, Rust, JavaScript
  • Experience in debugging and optimizing code and automate routine tasks
  • Experience in development, testing, deployment and administration of one or more of the following types of systems: Nginx, Kubernetes, Docker, OpenStack, Hadoop, Spark, Flink, Kafka
  • Experience in designing and analyzing large-scale distributed systems is preferred
  • Strong skills in problem solving and communication

Senior Site Reliability Engineer, Product - USDS ยท TikTok USDS

Auto apply with Likeremote