Skip to content
← Back to job listings

Platform Engineer - Traffic Infrastructure Operation & Maintenance

ByteDance · San Jose, California, United States of America

External listingfull-timeRecently

About The Role

About the Team

The Global Traffic Infrastructure (GTI) team leverages unified platform capabilities to manage edge infrastructure outside China (both self-built and third-party) providing standardized, compliant, scalable, and cost-effective traffic infrastructure capabilities for edge services. Our vision is to build a global edge traffic infrastructure platform and become the long-term cornerstone of ByteDance’s global edge business in terms of scale, performance, and cost.

Responsibilities

  • Responsible for the architecture design and engineering of the "network-traffic infrastructure" operation and maintenance & efficiency platform.
  • Responsible for the engineering of the CMDB, operation and maintenance automation, observability, stability, and change management systems for the "network-traffic infrastructure".
  • Responsible for the interactive design and system development of the efficiency tools for the "network-traffic infrastructure" business to improve the operational management efficiency.
  • Explore the application and implementation of intelligent operation and maintenance scenarios, and promote the intelligent evolution of system operation and maintenance.

Minimum Qualifications

  • Bachelor's degree or above in computer science or a related field, with at least 3 years of relevant experience in R&D, system operation and maintenance, or SRE.
  • Solid foundation in computer theory, with proficiency in at least one programming language such as Go, C, Python, etc.
  • Strong analytical and communication skills, strong sense of responsibility and team spirit.
  • Passionate about programming, with a strong thirst for knowledge, curiosity and ambition.

Preferred Qualifications

  • Experience in system engineering of large-scale distributed systems, management platforms or operation and maintenance platforms.
  • Familiarity with infrastructure architecture, and have a solid understanding of Kubernetes, edge computing, cloud networking, Load Balance, micro-services architecture and other related technologies.
  • Solid understanding of distributed systems, micro-services architecture, high availability, stability assurance, and emergency response systems.
  • Exploratory experience in LLM large model and Agent development.

This is an external listing. JobSpring does not represent or verify the employer. Report this listing