SpaceXAI’s mission is to create AI systems that can accurately understand the universe and aid humanity in its pursuit of knowledge. The Site Operations Supervisor at SpaceXAI Memphis will oversee the health and performance of server and network infrastructure across data centers and global points of presence, leading a team of technicians to ensure optimal operations and continuous improvement in key metrics such as mean time to detect (MTTD) and mean time to repair (MTTR).
Supervisor, Site Operations at xAI
On-site - Memphis, TN
More jobs at xAIRequirements
Skills
- High school diploma or equivalency certificate
- 3+ years of experience working with server, storage, compute, and network hardware
- 3+ years of experience troubleshooting and repairing servers and networking infrastructure
- Proven leadership experience in a data center or technical operations environment
- Strong Linux skills, including navigating system directories, manipulating files in the Linux shell, user permission configuration, and package installation
- Experience with Python, Bash or other scripting languages
- Experience leading data center infrastructure projects
- Familiarity with structured cabling (copper/fiber) and power and cooling concepts inside the data center
- Excellent prioritization and time management skills
- Ability to work in a fast‑paced environment and maintain attention to detail
- Position is subject to pre‑employment and annual post‑employment background checks
- Ability to lift up to 35 lbs. unassisted
- Comfortable working at elevated heights (up to 50 feet) with appropriate safety gear
- Comfortable working in an environment requiring exposure to noise
- Available to work evenings and weekends, with flexibility required
Responsibilities
- Lead and mentor a team of data center technicians, fostering a culture of excellence and continuous improvement
- Oversee the installation, maintenance, and troubleshooting of server and network infrastructure
- Manage and optimize data center operations, including power supply cabling, fiber/optics labeling, and hardware decommissioning
- Develop and enforce standard operating procedures (SOPs) and ensure adherence to safety protocols
- Coordinate with engineering and provisioning teams to ensure seamless hardware intake and repair processes
- Utilize internal applications for inventory and asset management, ensuring accurate tracking and reporting
- Manage data center operations tickets via Jira, ensuring timely resolution and documentation
- Collaborate with cross‑functional teams to design and implement network layouts and solutions
- Lead initiatives to improve operational efficiency and reduce downtime
- Provide on‑call support and respond to critical events as needed
Technologies
LinuxPythonBashJirastructured cabling
See if your resume is ready for this job
See how our AI can optimize your resume and improve your chances for this role.