活跃农民
- 积分
- 956
- 大米
- 颗
- 鳄梨
- 个
- 水井
- 尺
- 蓝莓
- 颗
- 萝卜
- 根
- 小米
- 粒
- 学分
- 个
- 注册时间
- 2012-11-23
- 最后登录
- 1970-1-1
|
最近AMD DCGPU硬件组和MLSE软件组均有很多职位放出,但一般标注MTS、SMTS、PMTS的不接受应届毕业生,请留意。. From 1point 3acres bbs
Machine Learning Performance Engineer.- 100521
Location: Austin, Texas, US.google и
The Role:
. 1point3acres.com This is a highly visible position with the opportunity to work closely with teams across AMD and impact MLPerf optimizations on AMD hardware and software architectures.
The Person:
We are seeking a machine learning specialist with a background in frameworks such as PyTorch/TensorFlow, and machine learning workloads. The successful candidate will have a passion for machine learning algorithms and understand the mapping of the workloads (such as CNNs, Transformer, BERT, recommendation models) on CPU/GPU architecture.
Preferred Experience:
Expertise with PyTorch/TensorFlow and performance tools such as nvprof and rocprof.
Experience with standard machine learning models such as ResNet50, Transformer, SSD, BERT, DLRM and Reinforcement Learning
Familiarity with large scale execution of machine learning models on clusters connected with Ethernet/InfiniBand
Accelerated Computing Network Engineer- 99256
Location: Austin, Texas, US
THE ROLE:
Do you want to be a key contributor to highly-scaled HPC and ML training clusters from AMD? The Network Validation and Performance team is looking for computer networking engineers to ensure GPUs perform optimally across high performance networks. You will be validating different networking technologies and identifying performance optimizations in areas such as: platform configurations, BIOS settings, network configurations, OS settings, GPU architecture, etc. You will be interfacing with networking partners as well as senior hardware and software engineers and architects in a very flat AMD organizational structure. You will be exposed to cutting edge technology and will have the opportunity to build out the test environment and contribute to test plans and methodology.
PREFERRED EXPERIENCE:
Comfortable setting up servers, changing hardware, updating firmware and installing server software
RDMA network and PCIe configuration and troubleshooting. .и
Industry standard performance benchmarks
System and software performance debug tools
ML and/or HPC system design |
|