Talent.com
TikTok
Cloud Site Reliability Engineer - Cloud and SystemTikTok • City Of London, England, GB
Cloud Site Reliability Engineer - Cloud and System

Cloud Site Reliability Engineer - Cloud and System

TikTok • City Of London, England, GB
30+ days ago
Job type
  • Full-time
Job description

Overview

Cloud Site Reliability Engineer - Cloud and System. Our Infrastructure Engineering team supports the company\'s fast growth by building and operating hyper-scale datacenters, managing the life cycle of the server fleet, providing cloud solutions, and developing various infrastructure services to ensure they are scalable and reliable.

Responsibilities

  • Build, expand, and operate global infrastructures, including large-scale systems in public and private clouds, data centers, and content delivery networks.
  • Build tools, automation, visualizations, and monitors to facilitate the operation and optimization of the global infrastructure.
  • Work in a fast-paced environment. Participate in technical operations and rotations in response to performance and reliability issues.
  • Help improve the whole lifecycle of infrastructure services from inception and design throughout development to deployment, user support, and refinement.

Qualifications

Minimum Qualification(s)

  • Master’s degree (or Bachelor\'s degree) with 3+ years of experience in Computer Engineering, Electrical Engineering, Computer Science, or related major.
  • 3+ years of experience working with Unix/Linux systems from kernel to shell and beyond, with experience working with system libraries, file systems, and client-server protocols.
  • 2+ years of experience working on Public Cloud Platforms, familiar with basic components of cloud products. Experience in building solutions with AWS, Google, OCI, or other cloud services.
  • 2+ years experience in one or more programming languages such as Java, C++, Go, or scripting experience in Shell and Python.
  • 2+ years experience with essential system-level apps, like DNS, APT, LDAP, Nginx, CI/CD, Ansible, Packer, etc.

Preferred Qualification(s)

  • Experience in system and data security.
  • Self-driven and capable of coping with ambiguity and moving projects from concept to delivery.
  • Strong analytical skills and the ability to solve real-world problems in a fast-moving environment.
  • Experience in designing, analyzing, and building automation and tools for large-scale systems.
  • Strong communication and collaboration skills.
  • The passion for solving problems.
  • Patience for supporting cases.

About TikTok

TikTok is the leading destination for short-form mobile video. At TikTok, our mission is to inspire creativity and bring joy. TikTok\'s global headquarters are in Los Angeles and Singapore, and we also have offices in New York City, London, Dublin, Paris, Berlin, Dubai, Jakarta, Seoul, and Tokyo.

Why Join Us

Inspiring creativity is at the core of TikTok\'s mission. Our innovative product is built to help people authentically express themselves, discover and connect – and our global, diverse teams make that possible. Together, we create value for our communities, inspire creativity and bring joy - a mission we work towards every day. We strive to do great things with great people. We lead with curiosity, humility, and a desire to make impact in a rapidly growing tech company. Every challenge is an opportunity to learn and innovate as one team. We\'re resilient and embrace challenges as they come. By constantly iterating and fostering an "Always Day 1" mindset, we achieve meaningful breakthroughs for ourselves, our company, and our users. When we create and grow together, the possibilities are limitless. Join us.

Diversity & Inclusion

TikTok is committed to creating an inclusive space where employees are valued for their skills, experiences, and unique perspectives. Our platform connects people from across the globe and so does our workplace. At TikTok, our mission is to inspire creativity and bring joy. To achieve that goal, we are committed to celebrating our diverse voices and to creating an environment that reflects the many communities we reach. We are passionate about this and hope you are too.

#J-18808-Ljbffr

Create a job alert for this search

Cloud Site Reliability Engineer - Cloud and System • City Of London, England, GB

Similar jobs

Senior Cloud Reliability & Platform Engineer

Carta HealthcareGreater London, England, GB
Full-time

Carta Healthcare seeks a Senior Site Reliability Engineer to build and scale internal platform services, ensuring reliability and performance for applications.You will design monitoring and inciden... Show more

 • Promoted

Site Reliability Engineer, Mistral Cloud

MistralGreater London, England, GB
Full-time

Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute.We partner with enterprises tackling the hardest problems—across high-stakes industries ... Show more

 • Promoted

Azure Site Reliability Engineer

BOSS ERP ConsultingGreater London, England, GB
Full-time

We are hiring for a Senior Azure Support Engineer to join a growing business and be responsible for the maintenance and support of the Azure Cloud Environment that hosts SaaS web-based applications... Show more

 • Promoted

Senior Site Reliability Engineer (Application / API Focused)

Xpertise RecruitmentGreater London, England, United Kingdom
Full-time

Senior Site Reliability Engineer (Application / API Focused).We are hiring a Senior SRE to support a large-scale digital organisation undergoing a major commercial re-platforming across web and mob... Show more

 • Promoted

Senior Site Reliability Engineer - Cloud Observability & Automation

OmiliaGreater London, England, GB
Full-time

Omilia is seeking a Senior Site Reliability Engineer with Cloud platform experience to join our production reliability team.You will operate and maintain production clusters, develop observability ... Show more

 • Promoted

Principal SRE: Cloud Reliability Lead (Azure/AWS)

FourthGreater London, England, GB
Full-time

Fourth is seeking an experienced Principal Site Reliability Engineer to accelerate cloud adoption and build automated, highly reliable infrastructure pipelines across Azure and AWS.You will collabo... Show more

 • Promoted

Remote Site Reliability Engineer - Healthcare Platform

Altera Digital Health APACGreater London, England, GB
Remote
Full-time

Altera Digital Health APAC is seeking a Site Reliability Engineer (SRE) to ensure the reliability and performance of healthcare platforms.This remote role focuses on monitoring, troubleshooting, an... Show more

 • Promoted

Site Reliability Engineer

Selby JenningsGreater London, England, GB
Full-time

Our client, a world leading systematic multi strat hedgefund is looking for a Site Reliability Engineer to work within their ETF Trading systems.This role combines building and owning the observabi... Show more

 • Promoted

Senior Site Reliability Engineer

P2PGreater London, England, GB
Full-time

Senior Site Reliability Engineer (SRE) - GCP/Kubernetes.We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small, agile engineering team.This role... Show more

 • Promoted

Cloud Infrastructure Engineer — Scalable, Reliable Systems

OpenAIGreater London, England, GB
Full-time

OpenAI in Greater London is seeking an experienced infrastructure engineer to design and build cloud platforms that ensure reliability and security at scale.The position entails maintaining systems... Show more

 • Promoted

Lead Site Reliability Engineer Global IT

Fintech Farm LtdGreater London, England, GB
Full-time

We are a UK fintech creating successful neobanks in emerging markets in partnerships with local traditional banks.The mission is to make banking services accessible, simple and fun to use worldwide... Show more

 • Promoted

Site Reliability Engineer — Global Travel Platform

DuffelCity Of London, England, GB
Full-time

A dynamic technology company in travel is seeking a Site Reliability Engineer to ensure the reliability and performance of infrastructure and applications.You will collaborate closely with engineer... Show more

 • Promoted

Site Reliability Engineer

CapitalontapGreater London, England, GB
Full-time

At Capital On Tap, we run a hybrid embedded SRE model.We aim to work closely with the teams within Capital On Tap to provide them the best support.Our main objective currently is to gain as much vi... Show more

 • Promoted

Lead Site Reliability Engineer

JPMorgan Chase & Co.Greater London, England, GB
Full-time

Our trading technology stack is undergoing a multi‑year convergence and modernization journey.You will play a pivotal role in shaping our next‑generation SRE patterns, reliability frameworks, obser... Show more

 • Promoted

Site Reliability Engineer

Sanderson Government & DefenceGreater London, England, GB
Full-time

UK-based (hybrid working; London / client site as required).SC clearance required (or eligibility).You’ll work closely with platform and delivery teams, designing resilient cloud infrastructure and... Show more

 • Promoted

Remote Senior Systems Engineer, Cloud Emulation

LocalStackCity Of London, England, GB
Remote
Full-time

A progressive cloud development platform is seeking a Systems Engineer in London to enhance their cutting-edge technology.The role requires over 5 years of experience in systems engineering and a s... Show more

 • Promoted

Senior Site Reliability Engineer - AWS Kubernetes

Source TechnologyLondon, England, GB
Full-time

Senior Site Reliability Engineer - AWS Kubernetes.Senior Site Reliability Engineer - AWS Kubernetes.Get AI-powered advice on this job and more exclusive features.A truly unique opportunity to help ... Show more

 • Promoted

Site Reliability Engineer — Remote/Hybrid, DevSecOps & Cloud

Made TechGreater London, England, GB
Remote
Full-time

Made Tech is hiring a Site Reliability Engineer to facilitate service onboarding and ongoing maintenance.The role is based in the UK with flexible hybrid working options.The successful candidate wi... Show more