Cloud Site Reliability Engineer (SRE) - Data Center Solutions (DCS)
Join ByteDance as a Cloud Site Reliability Engineer (SRE) in our Data Center Solutions (DCS) team and play a pivotal role in powering the infrastructure behind one of the worldβs fastest-growing technology companies. Our DCS team is at the heart of ByteDance's operations, ensuring that our global platform runs smoothly and efficiently. In this role, you will be responsible for maintaining and optimizing our cloud infrastructure, ensuring high availability, and driving continuous improvement in our data center solutions.
ByteDance is a global leader in technology, with a strong presence in Indonesia. We are committed to fostering a diverse and inclusive workplace, where everyone can thrive and contribute to our mission of connecting the world. As a Cloud SRE in Bali, you will have the opportunity to work in a dynamic and innovative environment, with access to the latest tools and technologies.
In this role, you will work closely with our engineering teams to design, implement, and maintain our cloud infrastructure. You will be responsible for monitoring and troubleshooting issues, ensuring that our systems are running optimally. You will also be involved in capacity planning and performance tuning, working to improve the efficiency and reliability of our data center solutions.
π Tanggung Jawab Pekerjaan
- Design, implement, and maintain cloud infrastructure using AWS, GCP, and Azure.
- Monitor and troubleshoot issues to ensure high availability and performance of our cloud services.
- Develop and implement automation scripts using Python, Bash, and other relevant tools.
- Collaborate with cross-functional teams to drive continuous improvement in our data center solutions.
- Conduct capacity planning and performance tuning to optimize the efficiency of our cloud infrastructure.
- Implement and manage security best practices to protect our cloud environment.
- Stay up-to-date with the latest industry trends and best practices in cloud computing.
- Provide technical leadership and mentorship to junior engineers and interns.
π Kualifikasi & Syarat
- Bachelor's degree in Computer Science, Engineering, or a related field.
- 5+ years of experience in cloud infrastructure and site reliability engineering.
- Strong knowledge of cloud platforms such as AWS, GCP, and Azure.
- Proficiency in scripting languages such as Python, Bash, and PowerShell.
- Experience with containerization technologies such as Docker and Kubernetes.
- Familiarity with monitoring and logging tools such as Prometheus, Grafana, and ELK Stack.
- Certifications in relevant cloud platforms (e.g., AWS Certified Solutions Architect, Google Professional Cloud Architect) are a plus.
- Excellent problem-solving skills and a strong attention to detail.
π οΈ Keahlian
Kirim lamaran sekarang sebelum batas waktu.π Lamar Sekarang