HPC Data Center Developer
Chicago, IL ยท On-site
Develop and maintain operational tooling that supports day-to-day data center workflows such as hardware lifecycle tracking, data center inventory/spares, change management, and diagnostics.
Chicago, IL ยท On-site
Develop and maintain operational tooling that supports day-to-day data center workflows such as hardware lifecycle tracking, data center inventory/spares, change management, and diagnostics.
Chicago, IL ยท On-site
Develop and maintain operational tooling that supports day-to-day data center workflows such as hardware lifecycle tracking, data center inventory/spares, change management, and diagnostics.
Chicago, IL ยท On-site
Develop and maintain operational tooling that supports day-to-day data center workflows such as hardware lifecycle tracking, data center inventory/spares, change management, and diagnostics.
Chicago, IL ยท On-site
Develop and maintain operational tooling that supports day-to-day data center workflows such as hardware lifecycle tracking, data center inventory/spares, change management, and diagnostics.
Elk Grove Village, IL ยท On-site
$109K - $145K/yr
Partner with HW Support teams to ensure data center hardware incidents with higher level troubleshooting challenges are resolved, reported on and solutions are disseminated to the large operations ...
Elk Grove Village, IL ยท On-site
$109K - $145K/yr
Partner with HW Support teams to ensure data center hardware incidents with higher level troubleshooting challenges are resolved, reported on and solutions are disseminated to the large operations ...
Elk Grove Village, IL ยท On-site
$89K - $119K/yr
Document data center layout and network topology in DCIM software * Work with supply chain & manufacturing teams to ensure timely deployment of systems and project plans for large-scale deployments
Elk Grove Village, IL ยท On-site
$89K - $119K/yr
Document data center layout and network topology in DCIM software * Work with supply chain & manufacturing teams to ensure timely deployment of systems and project plans for large-scale deployments
AWS Data Center Operations is seeking an innovative, customer-obsessed Technical Training Manager to own the development and implementation of a new Full-Time Training Team (FTT) for Data Center ...
AWS Data Center Operations is seeking an innovative, customer-obsessed Technical Training Manager to own the development and implementation of a new Full-Time Training Team (FTT) for Data Center ...
Elk Grove Village, IL ยท On-site
$35 - $37/hr
The Data Center Shift Technician is responsible for the daily operation, inspection, and maintenance of critical facility systems, including power, cooling, backup generation, and physical security ...
Quick apply
Elk Grove Village, IL ยท On-site
$35 - $37/hr
The Data Center Shift Technician is responsible for the daily operation, inspection, and maintenance of critical facility systems, including power, cooling, backup generation, and physical security ...
$160K - $200K/yr
Take our data center vertical from 1 to 10. You will drive operational excellence-whether that means streamlining complex site-access and safety protocols, standardizing onboarding for massive ...
$160K - $200K/yr
Take our data center vertical from 1 to 10. You will drive operational excellence-whether that means streamlining complex site-access and safety protocols, standardizing onboarding for massive ...
Chicago, IL ยท On-site
$109K - $148K/yr
Incident Management and Operation Improvement: - Manages incidents that impact data center infrastructure and the proactive and timely resolution of such incidents. -Collaborates with design ...
Chicago, IL ยท On-site
$109K - $148K/yr
Incident Management and Operation Improvement: - Manages incidents that impact data center infrastructure and the proactive and timely resolution of such incidents. -Collaborates with design ...
Aurora, IL ยท On-site
Opportunity The Data Center Site Operations Manager will be a vital member of the leadership team of the North American division of Edged Data Centers. We are looking for someone who has experience ...
Aurora, IL ยท On-site
Opportunity The Data Center Site Operations Manager will be a vital member of the leadership team of the North American division of Edged Data Centers. We are looking for someone who has experience ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Quick apply
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Deep expertise in data center infrastructure systems * Experience managing large-scale operations and budgets * Excellent written and verbal communication skills. * Ability to handle a multitude of ...
Deep expertise in data center infrastructure systems * Experience managing large-scale operations and budgets * Excellent written and verbal communication skills. * Ability to handle a multitude of ...
We are seeking a Project Manager - Data Center Operations (Cx) who thrives in fast-paced environments, leads from the front, and takes ownership of complex project outcomes. Why Top Talent Chooses ...
Quick apply
We are seeking a Project Manager - Data Center Operations (Cx) who thrives in fast-paced environments, leads from the front, and takes ownership of complex project outcomes. Why Top Talent Chooses ...
Deep expertise in data center infrastructure systems * Experience managing large-scale operations and budgets * Excellent written and verbal communication skills. * Ability to handle a multitude of ...
Deep expertise in data center infrastructure systems * Experience managing large-scale operations and budgets * Excellent written and verbal communication skills. * Ability to handle a multitude of ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Chicago, IL ยท On-site
The world' biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Chicago, IL ยท On-site
The world' biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Northlake, IL ยท On-site
The world' biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Quick apply
Northlake, IL ยท On-site
The world' biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
Quick apply
The world's biggest companies trust T5 with their data center operations. At T5, our success is fueled by our team. With over 400 engineers, technicians and professional staff, we're proud to foster ...
$53.6K - $67.4K
6% of jobs
$67.4K - $81.3K
5% of jobs
$81.3K - $95.1K
13% of jobs
$95.8K is the 25th percentile. Wages below this are outliers.
$95.1K - $109K
17% of jobs
The median wage is $122.1K / yr.
$109K - $122.9K
9% of jobs
$122.9K - $136.7K
8% of jobs
$136.7K - $150.6K
8% of jobs
$161.8K is the 75th percentile. Wages above this are outliers.
$150.6K - $164.4K
9% of jobs
$164.4K - $178.3K
9% of jobs
$178.3K - $192.2K
8% of jobs
$192.2K - $206K
5% of jobs
$53.6K
$132.4K
$206K
| Aspect | Data Center Operations | Data Center Technician |
|---|---|---|
| Primary Focus | Overseeing overall data center functions, including infrastructure management, monitoring, and maintenance coordination. | Performing hands-on hardware installation, troubleshooting, and repairs within the data center. |
| Certifications | Often requires certifications like CompTIA Data Center or Cisco CCNA. | Typically requires certifications such as CompTIA A+ or Network+. |
| Work Environment | Office-based with some on-site presence; involves coordination and supervision. | Primarily on-site, working directly with hardware and equipment. |
| Employer Usage | Used by data center managers, operations teams, and facilities management. | Commonly searched by technicians and entry-level hardware staff. |
In summary, Data Center Operations involves managing and coordinating data center functions, while Data Center Technicians focus on hands-on hardware tasks. Both roles are essential but differ in scope and responsibilities.

HPC Data Center Production Engineer
Location: Chicago or New York (On-site 5 days/week)
Jump Trading Group is committed to world class research. We empower exceptional talents in Mathematics, Physics, and Computer Science to seek scientific boundaries, push through them, and apply cutting edge research to global financial markets. Our culture is unique. Constant innovation requires fearlessness, creativity, intellectual honesty, and a relentless competitive streak. We believe in winning together and unlocking unique individual talent by incenting collaboration and mutual respect. At Jump, research outcomes drive more than superior risk adjusted returns. We design, develop, and deploy technologies that change our world, fund start-ups across industries, and partner with leading global research organizations and universities to solve problems.
Trading Infrastructure is a global organization of Engineers who architect, build and maintain our world-class infrastructure. From colo design/implementation, to optimizing our exchange connectivity, to building world class low latent Wide Area Networks, we leverage research and automation to consistently adapt and innovate our infrastructure to scale and drive our trading and evolving business.
We are looking for an HPC Data Center Production Engineer to build and own the automation and tooling that powers Jump's HPC data center operations. This is a development-heavy role focused on automating the onboarding and lifecycle management of data center hardware-servers, switches, rack PDUs, CDUs, and environmental sensors-and building tools for capacity planning, outage simulation, monitoring, and metrics integration. You will work hand-in-hand with the HPC Planning, Engineering, and Operations leads to turn tooling and monitoring vision into production-ready systems. Heavy, daily use of AI tools is expected to accelerate development and raise the quality bar on everything you build.
What You'll Do:
Hardware Onboarding Automation
Design, develop, and maintain automation to onboard new hardware devices into Jump's HPC data centers, including servers, network switches, rack PDUs, CDUs, and environmental sensors.
Build end-to-end provisioning workflows that take hardware from racked-and-cabled through discovery, configuration, validation, and production-ready state with minimal manual intervention.
Extend and adapt onboarding automation as new hardware platforms and device types are introduced.
Data Center Tooling Development
Develop tools for power and cooling capacity planning-enabling the operations and planning teams to model current utilization, forecast growth, and identify constraints before they become problems.
Build outage simulation tooling to model the impact of power, cooling, or network failures across HPC facilities and validate redundancy/failover configurations.
Develop and maintain operational tooling that supports day-to-day data center workflows such as hardware lifecycle tracking, data center inventory/spares, change management, and diagnostics.
Monitoring & Metrics Integration
Build and maintain monitoring integrations for HPC data center infrastructure-pulling telemetry from servers, switches, PDUs, CDUs, environmental sensors, and facility systems into centralized observability platforms.
Integrate metrics feeds from colocation and data center providers into Jump's monitoring stack, normalizing data for alerting and capacity reporting.
Work with the Operations Lead to implement the monitoring and alerting strategy, translating requirements into deployed, production-grade instrumentation.
Cross-Team Collaboration
Work very closely with the HPC Planning, Engineering, and Operations leads to understand tooling and monitoring needs and bring their vision to fruition.
Partner with HPC Engineering on integration points between data center automation and compute/storage/network provisioning systems.
Translate operational pain points and manual processes into automated, maintainable solutions.
Systems Maintenance & Reliability
Own the reliability and lifecycle of all systems and tools you develop-monitor for failures, respond to issues, and iterate based on operational feedback.
Maintain comprehensive documentation for all tooling, automation workflows, and integrations.
Participate in large, coordinated maintenance operations, including during evenings and weekends.
AI-Driven Development
Use AI tools daily across all aspects of the role: writing and reviewing code, analyzing data, debugging, generating documentation, and accelerating development velocity.
Identify opportunities to apply AI to data center operations problems-anomaly detection, predictive capacity planning, intelligent alerting, and beyond.
Additional duties as assigned or needed.
Skills You'll Need:
5+ years of professional experience in production engineering, infrastructure automation, or site reliability engineering, preferably in HPC or large-scale data center environments.
Proven track record of building and shipping production automation and tooling-not just scripts, but maintained, reliable systems.
Experience automating hardware provisioning and lifecycle management (servers, network devices, power/cooling infrastructure).
Strong understanding of data center infrastructure: power distribution, cooling systems (air and liquid), environmental monitoring, and structured cabling.
Experience integrating with hardware management interfaces (IPMI/BMC/Redfish, SNMP, vendor APIs) for discovery, configuration, and telemetry collection.
Demonstrates a high level of energy, results driven, and able to work under pressure with tight deadlines.
Technical Skills:
High proficiency in Golang and at least one additional language (e.g., Python). You will write a lot of code in this role.
Strong Linux systems knowledge-you should live in Linux. Proficient with system administration, networking, storage, process management, log analysis, and troubleshooting at the OS level.
Experience with Grafana for building dashboards, alerting, and visualization of infrastructure metrics. Experience with Prometheus, InfluxDB, or similar observability platforms and building custom integrations/exporters.
Experience with configuration management and infrastructure-as-code tools (SaltStack, Ansible, Terraform, or similar).
Solid understanding of networking concepts: L2/L3 protocols, VLANs, BGP, SNMP, and switch/router configuration (Arista, Cisco).
Experience with APIs and data integration-consuming vendor APIs, normalizing heterogeneous data sources, building data pipelines for metrics and reporting.
Experience with ClickHouse and MySQL-writing queries, designing schemas, and building tooling that reads from and writes to these databases.
Experience with GitHub for version control, code review, CI/CD workflows, and collaborative development.
Demonstrated heavy use of AI tools (e.g., LLM-based coding assistants, AI-driven analytics) in a professional setting. You should already be using AI daily and be eager to push its application further.
A compulsion to perform root cause analysis.
Excellent written and verbal communication skills with the ability to work across a global engineering team.
Extremely high personal standards for work quality.
Reliable and predictable availability, including ability to work evenings and weekends as required.
Bachelor's degree preferred.
Sourced by ZipRecruiter
Finance and insurance
501 - 1,000 Employees
Chicago, IL, US
1999