Data Center Operations Systems Engineer (Atlanta) at Lambda

Lambda is a leader in AI cloud infrastructure serving thousands of customers. This role focuses on ensuring data center operations run smoothly by managing server, storage, and network infrastructure deployment, configuration, and maintenance across Lambda's advanced data center facilities. What You'll Do: Rack, label, cable, and configure new server, storage, and network infrastructure; Troubleshoot hardware and software issues in advanced systems; Document data center layout and network topology using DCIM software; Coordinate with supply chain and manufacturing teams on system deployment and large-scale project planning; Manage parts depot inventory and track equipment through delivery, storage, staging, and deployment phases; Collaborate with HW Support and RMA teams to resolve infrastructure-related tickets and manage faulty equipment returns; Maintain installation standards and documentation for placement, labeling, and cabling consistency across all data centers. What You Need: Familiarity with critical data center infrastructure systems including power distribution, air flow management, environmental monitoring, capacity planning, DCIM software, structured cabling, and cable management; Strong attention to detail and ability to follow instructions; Action-oriented mindset with willingness to learn; Willingness to travel for new data center bring-ups. Nice to Have: Experience troubleshooting server hardware; Knowledge of network topology; Familiarity with ticketing systems such as JIRA and Zendesk; Linux administration experience; Experience working in large-scale distributed data center environments; Experience with Supermicro and NVIDIA hardware. Compensation: Salaried non-exempt role, eligible for overtime. Health, dental, and vision coverage for employee and dependents; wellness and commuter stipends for select roles; 401k plan with 2% company match (USA employees); generous cash and equity compensation; flexible paid time off.