1

Distributed Systems Manager Jobs (NOW HIRING)

Work with system-level concerns including scheduling, memory management, I/O optimization, storage ... Debug concurrency issues and distributed coordination challenges. * Design and maintain APIs and ...

Moment is the AI operating system for investment management, built for the world's largest wealth ... Responsibilities : • Architect and ship systems that solve genuinely hard problems - distributed ...

Distributed Systems Engineer

New York, NY · On-site

$200K - $400K/yr

Distributed Systems Engineer Build the future of investment management with us The infrastructure managing $300 trillion in assets was built in the 90s. Now, all of it is up for grabs. The winner of ...

You will define both how AI interacts with the world, and how humans build and manage those tools ... Experience with high-scale distributed systems * Experience leading teams and mentoring more junior ...

Make the system's view of itself always match reality: integrate fleet state as a machine-readable ... Unified machine management, actual state inspection, distributed command execution -- and the ...

Make the system's view of itself always match reality: integrate fleet state as a machine-readable ... Unified machine management, actual state inspection, distributed command execution - and the ...

Water Systems Manager

Milwaukee, WI · On-site

$79K - $96K/yr

The Water Systems Manager is responsible for overseeing the installation, operation, and maintenance of the City's water distribution system to ensure the safe and reliable delivery of potable water.

next page

Showing results 1-20

Distributed Systems Manager information

See salary details

$46K

$102.1K

$153K

How much do distributed systems manager jobs pay per year?

As of Sep 11, 2026, the average yearly pay for distributed systems manager in the United States is $102,067.00, according to ZipRecruiter salary data. Most workers in this role earn between $77,000.00 and $125,000.00 per year, depending on experience, location, and employer.

What is a distributed systems manager?

A Distributed Systems Manager is an IT professional responsible for overseeing the design, implementation, and maintenance of distributed computing systems—networks of computers that work together to achieve a common goal. They coordinate teams, manage system architecture, ensure system reliability and scalability, and address any issues that may arise in distributed environments. Their role is crucial in organizations that rely on cloud computing, large-scale web services, or enterprise software requiring high availability and fault tolerance.

What are the key skills and qualifications needed to thrive as a distributed systems manager?

To thrive as a Distributed Systems Manager, you need deep expertise in distributed computing, network architecture, and systems engineering, often supported by a degree in computer science or a related field. Familiarity with cloud platforms (such as AWS, Azure, or Google Cloud), containerization tools (like Docker and Kubernetes), and monitoring systems is essential, along with certifications such as AWS Certified Solutions Architect. Strong leadership, problem-solving abilities, and communication skills help you guide teams and coordinate complex projects effectively. These skills ensure the reliable operation, scalability, and security of distributed systems in dynamic technical environments.

How does a distributed systems manager typically collaborate with cross-functional teams to ensure system reliability and scalability?

As a Distributed Systems Manager, you will frequently collaborate with software engineers, DevOps specialists, and product managers to design, implement, and maintain scalable architectures. This role often involves leading incident response efforts, coordinating system upgrades, and facilitating regular meetings to align on technical priorities and business goals. Clear communication and proactive problem-solving are key, as you'll need to bridge the gap between technical challenges and business requirements to ensure high availability and reliability across distributed platforms.

What cities are hiring for Distributed Systems Manager jobs?

Cities with the most Distributed Systems Manager job openings:

What are popular job titles related to Distributed Systems Manager jobs?

For Distributed Systems Manager jobs, the most frequently searched job titles are:

Infographic showing various Distributed Systems Manager job openings in the United States as of August 2026, with employment types broken down into 1% As Needed, 85% Full Time, 9% Part Time, 4% Contract, and 1% Nights. Highlights an 89% Physical, 3% Hybrid, and 8% Remote job distribution, with an average salary of $102,067 per year, or $49.1 per hour.

Distributed Systems Engineer

Remote

ThisWay
Software Development • 11 - 50 employees

Full-time

Re-posted yesterday


Job description

ThisWay Global is looking for a Distributed Systems Engineer in a remote role within the United States.
ThisWay Global, Inc. is an AI-first technology company headquartered in Texas, operating at the intersection of artificial intelligence, data center infrastructure, and workforce solutions.
The company operates across three primary business areas:
  • ADCAP - AI & Data Center Acceleration Platform: A data center development and operations platform supporting accelerated deployment timelines and NVIDIA NVL72/GB300 GPU clusters.
  • Amalgamy.ai - AI Orchestration Software: An enterprise AI orchestration platform focused on GPU and compute utilization across AI environments.
  • Staffing & Workforce Solutions: AI-powered talent matching solutions connecting employers with candidates at scale.

This role focuses on building foundational distributed systems and operational infrastructure that support AI and exascale computing environments. The work emphasizes systems programming, distributed architecture, fault tolerance, and HPC-grade reliability.
Location: Remote - United States
Department: Engineering
Employment Type: Full-Time, Exempt
Responsibilities
  • Design and build distributed systems that tolerate latency, bandwidth constraints, and intermittent connectivity.
  • Implement fault-tolerant communication strategies, retry logic, backpressure, caching, and eventual consistency patterns.
  • Write maintainable, resilient, and tested code following development standards and methodologies.
  • Debug and improve system behavior, including networking and distributed coordination issues.
  • Contribute to systems written primarily in Rust.
  • Work with system-level concerns including scheduling, memory management, I/O optimization, storage hierarchy management, and system reliability.
  • Optimize performance and memory usage in resource-constrained environments.
  • Debug concurrency issues and distributed coordination challenges.
  • Design and maintain APIs and communication layers between distributed components.
  • Identify and reduce tight coupling across services and systems.
  • Diagnose and resolve cross-system failures in production environments.
  • Design and implement secure, reliable solutions aligned with engineering standards.
  • Collaborate with engineers and computer scientists on operating systems internals, compiler internals, fault tolerance, file system architecture, and trusted systems.
  • Contribute to resolving architectural and systemic issues.
  • Continue developing expertise in distributed systems, HPC infrastructure, and related tooling.

Requirements
  • Experience building distributed systems in environments with low bandwidth, high latency, or unreliable communication links.
  • Production experience developing systems in Rust or Go.
  • Understanding of distributed systems failure modes and mitigation strategies.
  • Knowledge of consistency models, coordination strategies, and state replication.
  • Experience designing APIs and communication layers between distributed components.
  • Experience working within established architectures and delivering production-quality components.
  • Understanding of systems-level concepts including durability, reliability, and operational behavior.
  • Ability to work independently while collaborating with technical leadership.
Preferred Qualifications
  • Experience with HPC environments, exascale computing, or AI/ML infrastructure.
  • Exposure to operating systems internals, compiler design, or language runtimes.
  • Experience with edge computing or constrained network environments.
  • Familiarity with message queues, event-driven systems, or streaming architectures.
  • Exposure to consensus algorithms or distributed coordination primitives.
  • Experience with concurrency, memory management, or performance optimization in production systems.
  • Experience contributing to developer tooling, internal platforms, or infrastructure-layer components.

Benefits
  • Remote work within the United States.
  • Opportunity to work on distributed systems supporting AI and exascale workloads.
  • Collaboration with engineers experienced in operating systems internals, compiler internals, fault tolerance, file system architecture, and trusted systems.
  • Exposure to AI infrastructure, HPC, and large-scale distributed computing environments.