Lambda→
Data Center Operations System Engineer (Kansas… at Lambda · Kansas…
Entry LevelOn-siteFull-timeKansas City, MO$89k–$119k/yr
Skills
data center infrastructuretroubleshooting hardwarenetwork topologydcim softwarelinux administrationaction-orientedwillingness to learnjirazendesksupermicro hardwarenvidia hardware
Job Description
Summary: Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. The role involves ensuring the proper setup and management of server, storage, and network infrastructure in the data center while troubleshooting hardware and software issues.
Responsibilities:
- Ensure new server, storage and network infrastructure is properly racked, labeled, cabled, and configured
- Troubleshoot hardware and software issues in some of the world’s most advanced systems
- Document data center layout and network topology in DCIM software
- Work with supply chain & manufacturing teams to ensure timely deployment of systems and project plans for large-scale deployments
- Manage a parts depot inventory and track equipment through the delivery-store-stage-deploy-handoff process in each of our data centers
- Work closely with HW Support team to ensure data center infrastructure-related support tickets are resolved
- Work with RMA team to ensure faulty parts are returned and replacements are ordered
- Follow installation standards and documentation for placement, labeling, and cabling to drive consistency and discoverability across all data centers
Required Qualifications:
- Ensure new server, storage and network infrastructure is properly racked, labeled, cabled, and configured
- Troubleshoot hardware and software issues in some of the world's most advanced systems
- Document data center layout and network topology in DCIM software
- Work with supply chain & manufacturing teams to ensure timely deployment of systems and project plans for large-scale deployments
- Manage a parts depot inventory and track equipment through the delivery-store-stage-deploy-handoff process in each of our data centers
- Work closely with HW Support team to ensure data center infrastructure-related support tickets are resolved
- Work with RMA team to ensure faulty parts are returned and replacements are ordered
- Follow installation standards and documentation for placement, labeling, and cabling to drive consistency and discoverability across all data centers
- Are familiar with critical infrastructure systems supporting data centers, such as power distribution, air flow management, environmental monitoring, capacity planning, DCIM software, structured cabling, and cable management
- Are someone who pays attention to detail and has the ability to follow instructions
- Are action-oriented and have a strong willingness to learn
- Are willing to travel for bring up of new data center locations
Preferred Qualifications:
- Experience with troubleshooting server hardware
- Experience with/or knowledge of network topology
- Familiarity with ticketing systems like JIRA and Zendesk
- Experience with Linux administration
- Experience with working in large-scale distributed data center environments
- Experience with Supermicro & Nvidia hardware
Required Skills: Data center infrastructure, Troubleshooting hardware, Network topology
Important Skills: DCIM software, Linux administration
Nice-to-Have Skills: Action-oriented, Willingness to learn, JIRA, Zendesk, Supermicro hardware, Nvidia hardware
Benefits: Generous cash & equity compensation, Health, dental, and vision coverage for you and your dependents, Wellness and commuter stipends for select roles, 401k Plan with 2% company match (USA employees), Flexible paid time off plan that we all actually use
Benefits
Generous cash & equity compensation
Health, dental, and vision coverage for you and your dependents
Wellness and commuter stipends for select roles
401k Plan with 2% company match (USA employees)
Flexible paid time off plan that we all actually use