Site Reliability Engineer (SRE) Intern — AI Infrastructure
TencentAdded 2d ago
Create a free account to save this job and browse the whole board.
Create free accountAI analysis
ProJev · full postingQuick summaryby newgrad.ai
What you'll do
- Set up, tune, and keep AI compute servers and network infrastructure running smoothly
- Run diagnostics on hardware, validate system performance, and apply firmware patches
- Work with engineering teams to build customized environments including bare-metal and container platforms
- Handle initial troubleshooting for on-site problems and record solutions for future reference
What they're looking for
- Enrolled in or recently finished a Bachelor's or Master's degree in computer science or engineering
- Knowledge of server components, firmware management, and working with Linux systems
- Some background with Bash or Python for automation tasks
- Interest in or experience with tools like Kubernetes, Slurm, Prometheus, or similar infrastructure software
Pay and perks
Apply on Tencent's site.
Apply nowListing wrong or expired, or you're the employer and want it removed? Let us know.
