
ByteDance
Platform Engineer -Traffic Infrastructure Operation & Maintenance
직무인프라
경력중급
위치San Jose, Canada, United States
근무오피스 출근
고용Regular
게시오늘
포지션 소개
About the Team:
The Global Traffic Infrastructure (GTI) team leverages unified platform capabilities to manage edge infrastructure outside China (both self-built and third-party) providing standardized, compliant, scalable, and cost-effective traffic infrastructure capabilities for edge services. Our vision is to build a global edge traffic infrastructure platform and become the long-term cornerstone of Byte Dance’s global edge business in terms of scale, performance, and cost.
Responsibilities:
- Responsible for the architecture design and engineering of the "network-traffic infrastructure" operation and maintenance & efficiency platform.
- Responsible for the engineering of the CMDB, operation and maintenance automation, observability, stability, and change management systems for the "network-traffic infrastructure".
- Responsible for the interactive design and system development of the efficiency tools for the "network-traffic infrastructure" business to improve the operational management efficiency.
- Explore the application and implementation of intelligent operation and maintenance scenarios, and promote the intelligent evolution of system operation and maintenance.
Requirements:
- Minimum Qualifications
- Bachelor's degree or above in computer science or a related field, with at least 3 years of relevant experience in R&D, system operation and maintenance, or SRE.
- Solid foundation in computer theory, with proficiency in at least one programming language such as Go, C, Python, etc.
- Strong analytical and communication skills, strong sense of responsibility and team spirit.
- Passionate about programming, with a strong thirst for knowledge, curiosity and ambition.
Preferred Qualifications:
- Experience in system engineering of large-scale distributed systems, management platforms or operation and maintenance platforms.
- Familiarity with infrastructure architecture, and have a solid understanding of Kubernetes, edge computing, cloud networking, Load Balance, micro-services architecture and other related technologies.
- Solid understanding of distributed systems, micro-services architecture, high availability, stability assurance, and emergency response systems.
- Exploratory experience in LLM large model and Agent development.
필수 스킬
Facilities maintenance
Troubleshooting
Safety procedures
ByteDance 소개
San Jose
본사 위치