Data Center Operations Engineer - Chicago ORD

Elk Grove Village , United States
full-time On-site

AI overview

Ensure the performance and reliability of AI infrastructure through oversight of data center operations and collaboration across teams to meet operational goals.

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you'd like to build the world's best AI cloud, join us.


*Note: This position requires presence in our Chicago/Elk Grove Village Data Center location 5 days per week.


The Operations team is at the heart of keeping our AI-IaaS infrastructure running smoothly from start to finish. They handle everything from sourcing the right hardware and components to keeping our data centers performing at their best day in and day out. The team also works closely across the company, making sure our operational capabilities stay in sync with product goals and overall strategy. By managing the entire lifecycle — from procurement through deployment and ongoing efficiency — the Operations team ensures our AI infrastructure stays reliable, scalable, and ready to support the business as it grows.

What You’ll Do

  • Make sure new servers, storage, and networking gear are racked, labeled, cabled, and configured the right way.

  • Keep data center layouts and network topologies up to date in our DCIM software.

  • Coordinate with supply chain and manufacturing teams so systems are deployed on time, especially for large-scale projects.

  • Evaluate current and future data center needs based on growth and technology trends.

  • Manage parts depot inventory and track equipment as it moves from delivery → storage → staging → deployment → handoff.

  • Work closely with hardware support teams to get tickets resolved quickly.

  • Create and manage RMA tickets when needed, making sure faulty parts are replaced and reinstalled without delay.

  • Develop and maintain installation standards (placement, labeling, cabling) to ensure consistency across all data centers.

  • Act as a subject matter expert on data center deployments, supporting sales engagements for major deployments in our facilities or at customer sites.

You

Sourcing & Procurement

  • Researching, evaluating, and securing the right hardware and infrastructure components.

  • Building relationships with peers and supply chain to ensure cost-effective and timely supply.

Data Center Operations

  • Monitoring day-to-day performance of data centers to maintain uptime and efficiency.

  • Troubleshooting and resolving hardware or infrastructure issues quickly.

  • Performing regular maintenance and upgrades to keep systems running at peak performance.

Deployment & Lifecycle Management

  • Overseeing the full lifecycle of infrastructure, from initial setup to ongoing optimization.

  • Coordinating deployments of new hardware and ensuring seamless integration with existing systems.

  • Managing capacity planning to make sure infrastructure can scale with business growth.

Cross-Team Collaboration

  • Working with product management, support, and other teams to align operational capabilities with company goals.

  • Translating business priorities into technical and operational requirements.

  • Supporting cross-functional projects where infrastructure plays a critical role.

Reliability & Scalability

  • Ensuring infrastructure remains stable, secure, and scalable as demand increases.

  • Continuously improving processes to boost efficiency and reduce downtime risks.

Nice to Have

  • Certifications: Any Linux or project management.

  • Military background.

  • Experience in the machine learning or computer hardware industry

Salary Range Information

The annual salary range for this position has been set based on market data and other factors. However, a salary higher or lower than this range may be appropriate for a candidate whose qualifications differ meaningfully from those listed in the job description.

About Lambda

  • Founded in 2012, with 500+ employees, and growing fast

  • Our investors notably include TWG Global, US Innovative Technology Fund (USIT), Andra Capital, SGW, Andrej Karpathy, ARK Invest, Fincadia Advisors, G Squared, In-Q-Tel (IQT), KHK & Partners, NVIDIA, Pegatron, Supermicro, Wistron, Wiwynn, Gradient Ventures, Mercato Partners, SVB, 1517, and Crescent Cove

  • We have research papers accepted at top machine learning and graphics conferences, including NeurIPS, ICCV, SIGGRAPH, and TOG

  • Our values are publicly available: https://lambda.ai/careers

  • We offer generous cash & equity compensation

  • Health, dental, and vision coverage for you and your dependents

  • Wellness and commuter stipends for select roles

  • 401k Plan with 2% company match (USA employees)

  • Flexible paid time off plan that we all actually use

A Final Note:

You do not need to match all of the listed expectations to apply for this position. We are committed to building a team with a variety of backgrounds, experiences, and skills.

Equal Opportunity Employer

Lambda is an Equal Opportunity employer. Applicants are considered without regard to race, color, religion, creed, national origin, age, sex, gender, marital status, sexual orientation and identity, genetic information, veteran status, citizenship, or any other factors prohibited by local, state, or federal law.

Perks & Benefits Extracted with AI

  • Health Insurance: Health, dental, and vision coverage for you and your dependents
  • 401k match: 401k Plan with 2% company match (USA employees)
  • Paid Time Off: Flexible paid time off plan that we all actually use
  • Wellness Stipend: Wellness and commuter stipends for select roles
Salary
$89,000 – $134,000 per year
Get hired quicker

Be the first to apply. Receive an email whenever similar jobs are posted.

Ace your job interview

Understand the required skills and qualifications, anticipate the questions you may be asked, and study well-prepared answers using our sample responses.

Operations Engineer Q&A's
Report this job
Apply for this job