NOC Engineer
About the Company
Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company's Most Innovative Companies and to the CNBC Disruptor 50, Armada’s solutions are deployed in over 60 countries globally for organizations ranging from energy to defense.
With nearly $500 million in funding to date, Armada is backed by leading investors including Founders Fund, Lux, BlackRock and Microsoft (M12), alongside strategic partnerships with Microsoft, Dell, Palantir, NVIDIA, SpaceX, and Skydio. We are building the infrastructure layer for sovereign and edge AI - rugged, deployable compute for customers that cannot rely on centralized cloud.
Working at Armada means taking ownership, driving autonomy, and delivering impact. You’ll tackle challenges that haven’t been solved before and help build something transformative from the ground up. What you do here will not only define your career but help further Armada’s mission to bridge the digital divide for customers around the world.
About the Role
Armada is seeking an experienced NOC / Data Center Operations Engineer to support the monitoring, operation, and reliability of our critical physical infrastructure and distributed edge data center environments.
In this role, you will monitor and troubleshoot power, cooling, networking, environmental, and facility systems across Armada’s infrastructure. You will serve as a Tier 2 escalation point, lead incidents through resolution, coordinate with engineering and facilities teams, and help continuously improve our monitoring, alerting, and operational processes.
The ideal candidate brings hands-on experience with data center operations, PLC/BMS/DCIM platforms, mechanical and electrical infrastructure, incident management, and remote or edge infrastructure.
What You'll Do
Infrastructure Monitoring & Incident Response
- Monitor critical physical infrastructure using PLC, BMS, and DCIM platforms such as Distech, Radix IoT/Mango, Schneider Electric, Siemens, or similar systems.
- Monitor and respond to alerts involving power, cooling, network connectivity, environmental conditions, and facility infrastructure.
- Provide Tier 2 support for escalations from L1 technicians and coordinate escalation to engineering, facilities, vendors, or other teams as required.
- Own incidents through their lifecycle, from initial triage and troubleshooting through resolution, stakeholder communication, and documentation.
- Support incident response and post-incident reviews, identifying opportunities to improve operational reliability and response procedures.
- Develop and refine monitoring dashboards, alerting thresholds, and operational health indicators.
Data Center Operations
- Perform and coordinate routine health checks of critical infrastructure, including UPS systems, PDUs, CRAC/CRAH units, backup generators, and environmental monitoring systems.
- Troubleshoot and coordinate resolution of mechanical and electrical infrastructure issues using a working knowledge of MEP systems.
- Read and interpret technical documentation, including electrical one-line diagrams, network diagrams, schematics, and equipment documentation.
- Coordinate scheduled and emergency maintenance activities with internal teams, vendors, and remote hands personnel.
- Participate in change management processes and assess the operational and infrastructure impact of proposed changes.
- Ensure physical and logical security policies and procedures are followed within data center and edge environments.
Modular & Edge Infrastructure
- Operate and support modular, containerized, micro, and distributed edge data center environments.
- Help maintain availability, continuity, and resiliency across geographically distributed and remotely operated infrastructure.
- Coordinate remote troubleshooting and hands-on support for edge deployments.
- Implement and continuously improve operational practices for remote infrastructure monitoring, fault detection, escalation, and recovery.
- Support environmental sensors and IoT-based monitoring integrations across remote infrastructure.
Tools, Automation & Documentation
- Use operational platforms such as Grafana, Zoho Desk, Zenduty, ServiceNow, Jira, SolarWinds, or similar tools for monitoring, analysis, ticketing, and incident management.
- Maintain accurate documentation of configurations, incidents, changes, maintenance activities, and recurring operational tasks.
- Support automation and reporting initiatives focused on infrastructure performance, availability, capacity, and uptime.
- Assist with gathering operational evidence and infrastructure data for compliance and audit requirements.
- Identify opportunities to automate repetitive operational activities using scripting and monitoring integrations.
Collaboration & Operational Excellence
- Partner closely with Facilities, Infrastructure Engineering, Networking, IT, Security, and other operational teams to maintain reliable infrastructure.
- Communicate clearly during incidents, maintenance activities, escalations, and operational handoffs.
- Participate in incident reviews and root cause analysis and help drive corrective and preventive actions.
- Contribute to the development of runbooks, standard operating procedures, escalation processes, and operational best practices.
- Participate in an on-call or shift-based operating model as required to support 24x7 infrastructure.
Required Qualifications
- 5+ years of experience in NOC, data center operations, critical infrastructure operations, or a related technical role.
- Strong working knowledge of PLC, BMS, and/or DCIM platforms.
- Hands-on experience monitoring or supporting power and cooling infrastructure in data center or other mission-critical environments.
- Working knowledge of mechanical and electrical systems, including UPS, PDU, CRAC/CRAH, generators, and environmental monitoring systems.
- Solid understanding of networking fundamentals, infrastructure monitoring, alerting, and monitoring protocols.
- Experience with critical infrastructure monitoring platforms such as Distech, Radix IoT/Mango, Schneider Electric, Siemens, or equivalent systems.
- Experience with incident management, troubleshooting, escalation, and operational change management.
- Ability to read and interpret electrical one-lines, network schematics, technical diagrams, and equipment documentation.
- Strong analytical, troubleshooting, communication, and documentation skills.
- Ability to work effectively in a shift-based and/or on-call environment, including nights and weekends when required.
Preferred Qualifications
- Experience operating or supporting modular, containerized, micro, or edge data centers.
- Experience supporting geographically distributed infrastructure and coordinating remote hands activities.
- Familiarity with environmental sensors, telemetry, IoT integrations, and remote monitoring systems.
- Working knowledge of Linux and Windows server environments.
- Basic scripting and automation experience using Python, PowerShell, Bash, or similar technologies.
- Experience with monitoring and operational platforms such as Grafana, ServiceNow, Jira, SolarWinds, Zenduty, or Zoho Desk.
- Familiarity with compliance, audit evidence collection, and operational security requirements.
- Certifications such as CDCP, CDCTP, CompTIA Server+, CCNA, or equivalent industry certifications.
- Highly organized, detail-oriented, and comfortable operating in fast-paced, mission-critical environments.
#LI-Onsite
You're a Great Fit if You're
- A go-getter with a growth mindset. You're intellectually curious, have strong business acumen, and actively seek opportunities to build relevant skills and knowledge
- A detail-oriented problem-solver. You can independently gather information, solve problems efficiently, and deliver results with a "get-it-done" attitude
- Thrive in a fast-paced environment. You're energized by an entrepreneurial spirit, capable of working quickly, and excited to contribute to a growing company
- A collaborative team player. You focus on business success and are motivated by team accomplishment vs personal agenda
- Highly organized and results-driven. Strong prioritization skills and a dedicated work ethic are essential for you
Equal Opportunity Statement
At Armada, we are committed to fostering a work environment where everyone is given equal opportunities to thrive. As an equal opportunity employer, we strictly prohibit discrimination or harassment based on race, color, gender, religion, sexual orientation, national origin, disability, genetic information, pregnancy, or any other characteristic protected by law. This policy applies to all employment decisions, including hiring, promotions, and compensation. Our hiring is guided by qualifications, merit, and the business needs at the time.
Unsolicited Resumes and Candidates
Armada does not accept unsolicited resumes or candidate submissions from external agencies or recruiters. All candidates must apply directly through our careers page. Any resumes submitted by agencies without a prior signed agreement will be considered unsolicited and Armada will not be obligated to pay any fees.
Apply for this job
*
indicates a required field

