TikTok

Site Reliability Engineer

London, UK
Git Docker Kubernetes AWS Shell Go Python
Description
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed infrastructures. Our SREs are tasked to ensure the traffic services are reliable, fault-tolerant, efficiently scalable and cost-effective. You will have the opportunity to manage a variety of complex systems at scale, including traffic systems that serve hyperscale datacenters and public cloud, global load balancer that handles Tbps of traffic etc..

Responsibilities
- Build, expand and operate Bytedance’s global traffic platform, including large-scale systems in public and private clouds, edge data centers.
- Build tools, automations, visualizations and monitors to facilitate the operation and optimization of the global traffic platform.
- Work in a fast-paced environment. Participate in technical operations and rotations in response to performance and reliability issues.
- Help improve the whole lifecycle of infrastructure services from inception and design throughout development, to deployment, user support and refinementMinimum Qualifications
- Master’s degree (or Bachelor's degree with 3+) years of experience in Computer Engineering, Electrical Engineering, Computer Science or related major
- 3+ years experience working with Linux systems from kernel to shell and beyond with experience working with system libraries, file systems, and client-server protocols.
- 3+ years experience in one or more programming languages such as Go, Python and Shell script.
- Familiar with Cloud and CI/CD framework/Tools, such as GIT, Docker, Kubernetes, etc.
- Self-driven and capable of coping with ambiguity and moving projects from concept to delivery.
- Strong in analytical skills and the ability to solve real world problems in a fast moving environment.

Preferred Qualifications
- Experience in designing, analyzing and building automation and tools for large scale systems
- Experience in building solutions with AWS, Google, Azures and other cloud services.
- Experience in networking technologies such TCP/IP, HTTP, DNS, etc. in a carrier-grade environment.
- Experience in developing and operating one or more of following systems: Kubernetes, Nginx, ipvs, ELK stack, etc.
TikTok
TikTok

0 applies

0 views

There are more than 50,000 engineering jobs:

Subscribe to membership and unlock all jobs

Engineering Jobs

60,000+ jobs from 4,500+ well-funded companies

Updated Daily

New jobs are added every day as companies post them

Refined Search

Use filters like skill, location, etc to narrow results

Become a member

🥳🥳🥳 452 happy customers and counting...

Overall, over 80% of customers chose to renew their subscriptions after the initial sign-up.

To try it out

For active job seekers

For those who are passive looking

Cancel anytime

Frequently Asked Questions

  • We prioritize job seekers as our customers, unlike bigger job sites, by charging a small fee to provide them with curated access to the best companies and up-to-date jobs. This focus allows us to deliver a more personalized and effective job search experience.
  • We've got over 200,000 jobs from 15,000+ vetted companies. No fake or sleazy jobs here!
  • We aggregate jobs from 15,000+ companies' career pages, so you can be sure that you're getting the most up-to-date and relevant jobs.
  • We're the only job board *for* software engineers, *by* software engineers… in case you needed a reminder! We add thousands of new jobs daily and offer powerful search filters just for you. 🛠️
  • Every single hour! We add 2,000-3,000 new jobs daily, so you'll always have fresh opportunities. 🚀
  • Typically, job searches take 3-6 months. EchoJobs helps you spend more time applying and less time hunting. 🎯
  • Check daily! We're always updating with new jobs. Set up job alerts for even quicker access. 📅

What Fellow Engineers Say