中级农民
- 积分
- 107
- 大米
- 颗
- 鳄梨
- 个
- 水井
- 尺
- 蓝莓
- 颗
- 萝卜
- 根
- 小米
- 粒
- 学分
- 个
- 注册时间
- 2019-3-10
- 最后登录
- 1970-1-1
|
今日精选职位
Software Engineer, TikTok SRE
山景城·社招·正式
职位描述
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed and fault-tolerant systems. Infrastructure SRE ensures that TikTok's infrastructure services reliability and uptime appropriate to the needs of users and fast iterations of improvement. Our software development pays great attention to optimizing existing systems, building infrastructure and eliminating work through automation.
Responsibilities:
1. Help improve the whole lifecycle of infrastructure services from inception and design, throughout development, capacity planning and launch reviews, to deployment, operation and refinement;
2. Design and implement software platforms and monitor frameworks for efficient, automated and intelligent service-oriented architecture (SOA) governance;
3. Scale systems sustainability through mechanisms such as automation; evolve systems reliability, efficiency, and velocity by pushing for changes;
4. Maintain services to meet service-level-agreements (SLAs) or service-level-objectives (SLOs) by measuring and monitoring availability, performance, and overall system health;
5. Provide user support, incident responses and postmortems.
职位要求
1. Bachelor or above degree in Computer Science or a related technical discipline with 3-5+ years experience in the deployment and administration of large-scale distributed systems;
2. Familiar with Unix/Linux operating systems internals and administration, networking (e.g. TCP/IP, routing, network topologies and hardware), storage systems, and database
3. Experience in one of the following programmings: C, C++, Java, Python, Go, Perl, Ruby or shell scripting;
4. Experience in debugging and optimizing code and automate routine tasks;
5. Experience in the development, test, deployment and administration of one of the following types of systems: Ngnix, Kubernetes, Docker, OpenStack, Hadoop, Spark, Flink, etc. is preferred;
6. Experience in designing and analyzing large-scale distributed systems is preferred;
7. Strong skills in problem solving and communication. |
|