New York, New York, United StatesInternshipEntry-levelPosted Today
Our Server Fleet Operations team manages the full lifecycle of servers powering our self-built data centers across the United States and Europe: from new hardware rollout to decommissioning. We work at the center of a large network of teams, including hardware design, data center operations, field maintenance, supply chain, and automation engineering, to keep our global infrastructure running reliably and efficiently at scale.
As a Project Intern, you will contribute to impactful short-term projects and gain hands-on experience in a fast-paced, professional environment. This internship offers the opportunity to develop practical skills, apply your knowledge to real-world challenges, and explore your career interests.
Applications are reviewed on a rolling basis, so we encourage you to apply early.
What You'll Do
- Get hands-on with server deployment, monitoring, and maintenance across CPU and GPU fleets
- Write scripts and small tools (Python, Bash, or similar) to automate repetitive operational tasks
- Learn to troubleshoot real Linux issuesL OS, hardware, storage, networking, and performance
- Explore GPU and AI infrastructure, and contribute to the tools that keep it reliable
- Dig into infrastructure data and metrics to spot trends and improvement opportunities
- Experiment with applying AI/LLMs to infrastructure troubleshooting and operations
- Support incident investigations and root-cause analysis alongside senior engineers
- Help write and improve documentation, runbooks, and internal knowledge bases
- Collaborate with engineers across hardware, ops, platform, and supply chain teams on real projects
annually.
As a Project Intern, you will contribute to impactful short-term projects and gain hands-on experience in a fast-paced, professional environment. This internship offers the opportunity to develop practical skills, apply your knowledge to real-world challenges, and explore your career interests.
Applications are reviewed on a rolling basis, so we encourage you to apply early.
What You'll Do
- Get hands-on with server deployment, monitoring, and maintenance across CPU and GPU fleets
- Write scripts and small tools (Python, Bash, or similar) to automate repetitive operational tasks
- Learn to troubleshoot real Linux issuesL OS, hardware, storage, networking, and performance
- Explore GPU and AI infrastructure, and contribute to the tools that keep it reliable
- Dig into infrastructure data and metrics to spot trends and improvement opportunities
- Experiment with applying AI/LLMs to infrastructure troubleshooting and operations
- Support incident investigations and root-cause analysis alongside senior engineers
- Help write and improve documentation, runbooks, and internal knowledge bases
- Collaborate with engineers across hardware, ops, platform, and supply chain teams on real projects
annually.
