2

Remote Gpu Engineer Jobs in Oregon (NOW HIRING)

It transforms legacy data silos into data pipelines that dramatically increase GPU utilization and ... Proactively monitor customer environments (Ceph and WEKA) using observability and remote monitoring ...

Solutions Architect, HPC Systems Engineer

OR · On-site +1

$63 - $83/hr

Working with NVIDIA Consumer Internet and IT Services customers on data center GPU server and ... open to remote work location and look forward to have you join our team. NVIDIA is widely ...

Showing results 21-28

Remote Gpu Engineer information

What is a remote GPU engineer?

Remote GPU Engineers are specialized software or hardware engineers who work primarily with Graphics Processing Units (GPUs) from a remote location. They focus on designing, optimizing, and maintaining GPU-based systems for applications such as machine learning, high-performance computing, and graphics rendering. These professionals often collaborate with teams virtually, leveraging cloud-based GPU resources and remote access tools. Their work enables companies to efficiently utilize GPU technology without requiring engineers to be on-site.

What are the key skills and qualifications needed to thrive as a remote GPU engineer?

To thrive as a Remote GPU Engineer, you need strong expertise in GPU architectures, parallel programming (CUDA/OpenCL), and a solid background in computer science or engineering. Familiarity with tools like CUDA Toolkit, performance profilers, and version control systems, as well as experience with relevant certifications, is typically required. Excellent problem-solving abilities, communication skills, and the capacity to collaborate effectively in remote, distributed teams are standout soft skills. These competencies ensure efficient GPU solution development, effective troubleshooting, and seamless teamwork in a remote engineering environment.

What are some common challenges faced by remote GPU engineers when collaborating with distributed teams?

Remote GPU Engineers often work with global teams, which can present challenges such as coordinating across different time zones, ensuring consistent communication, and managing access to high-performance hardware remotely. To overcome these hurdles, it's important to leverage collaboration tools, maintain clear documentation, and establish regular check-ins. Additionally, using remote desktop solutions and cloud-based GPU environments can help facilitate smoother development and debugging processes.

What are the most commonly searched types of Gpu Engineer jobs in Oregon?

The most popular types of Gpu Engineer jobs in Oregon are:

What job categories do people searching Remote Gpu Engineer jobs in Oregon look for?

The top searched job categories for Remote Gpu Engineer jobs in Oregon are:

What cities in Oregon are hiring for Remote Gpu Engineer jobs?

Cities in Oregon with the most Remote Gpu Engineer job openings:

Designated Service Engineer - Ceph Expert

OR • On-site, Remote

WEKA
Software Development • 201 - 500 employees

Full-time

Re-posted 25 days ago


Key responsibilities

  • Architect, deploy, and operate large-scale production Ceph clusters supporting S3 with an emphasis on availability, performance, and operational simplicity.

  • Own cluster lifecycle activities including upgrades, patching, configuration management, routine health checks, and proactive risk remediation.

  • Troubleshoot complex issues across the Ceph stack, lead incident response and root-cause analysis.


Job description

About the job

WEKA is architecting a new approach to the enterprise data stack built for the age of reasoning. NeuralMesh by WEKA sets the standard for agentic AI data infrastructure with a cloud and AI-native software solution that can be deployed anywhere. It transforms legacy data silos into data pipelines that dramatically increase GPU utilization and make AI model training and inference, machine learning, and other compute-intensive workloads run faster, work more efficiently, and consume less energy.
WEKA is a pre-IPO, growth-stage company on a hyper-growth trajectory. We've raised $375M in capital with dozens of world-class venture capital and strategic investors. We help the world's largest and most innovative enterprises and research organizations, including 12 of the Fortune 50, achieve discoveries, insights, and business outcomes faster and more sustainably. We're passionate about solving our customers' most complex data challenges to accelerate intelligent innovation and business value. If you share our passion, we invite you to join us on this exciting journey.

What You'll Be Doing

This is a customer-facing Premium Services role that combines deep, hands-on Ceph architecture and administration with the high-touch, outcome-driven approach of a Senior Designated Services Engineer (DSE). You will be the primary Ceph subject matter expert for assigned strategic customers and internal initiatives, owning the design, deployment, lifecycle operations, and performance of Ceph-based object storage environments. In parallel, you will play a key role in ensuring WEKA's Customer Success, contributing to our five-star Gartner reviews. You will work with cutting-edge technologies and top-tier customers, providing technical expertise and strengthening customer relationships.

Collaborating closely with Account Teams, you will gain deep insight into customers' business requirements, technical needs, and system environments. Your role involves resolving technical issues, bridging gaps between customers and Engineering, and ensuring the highest level of service.

Ceph Architecture & Operations
  • Architect, deploy, and operate large-scale production Ceph clusters supporting S3 with an emphasis on availability, performance, and operational simplicity.
  • Own cluster lifecycle activities: upgrades, patching, configuration management, routine health checks, and proactive risk remediation.
  • Troubleshoot complex issues across the Ceph stack, lead incident response and root-cause analysis.
  • Establish and maintain runbooks, operational best practices, and customer-facing documentation; drive continuous improvement in reliability, observability, and automation.
  • Partner with customer teams on security and compliance requirements
  • Advise on hardware and topology choices to meet workload requirements.
Designated Services Engineering 
  • Serve as the primary technical liaison between customers and WEKA Engineering/Product to address feature gaps, reliability concerns, and documentation improvements.
  • Own, track, and document customer issues via the ticketing system; drive issues to resolution with clear, timely communication and executive-ready updates when needed.
  • Proactively monitor customer environments (Ceph and WEKA) using observability and remote monitoring tools to identify and remediate risks before they impact production.
  • Support account teams (Customer Success, Sales Engineering, Partners/Resellers) with deep technical expertise and credibility in front of senior customer stakeholders.
  • Contribute to knowledge sharing through internal and customer-facing documentation (FAQs, KB articles, runbooks) and repeatable troubleshooting playbooks.
  • Manage multiple engagements and cases concurrently, balancing urgency, impact, and long-term customer outcomes.
  • Participate in on-call and follow-the-sun support rotations as required; work occasional alternative hours (nights/weekends/holidays) and travel as needed.
Learning & Growth at WEKA
  • Ramp on WEKA's architecture, tooling, and support model, and progressively take ownership of designated services engagements beyond Ceph.
  • Develop deeper expertise in S3-compatible object storage concepts and ecosystems (clients, load balancing, performance testing, multi-tenancy), with mentorship from WEKA SMEs.
  • Partner with internal teams to improve product supportability and operational excellence for object-storage use cases.
Requirements

We're looking for a senior, customer-facing engineer who can lead Ceph architecture and operations today, and who is excited to grow into a broader object-storage and WEKA role.

  • 10+ years in customer-facing technical roles solving complex enterprise infrastructure issues.
  • 5+ years of hands-on Ceph experience in production: cluster design, deployment, upgrades, and day-2 operations.
  • Strong understanding of Ceph internals and operational mechanics: MON quorum, MGR active/standby, OSD behavior, CRUSH and CRUSH maps, pools and placement groups (PGs), recovery/backfill and rebalancing.
  • Experience operating large-scale (multi-PB) Ceph environments and navigating the operational challenges of fleet size, PG scaling, and long-running recovery events.
  • Practical experience with Ceph RGW and S3 concepts (buckets, users/tenants, load balancing, scaling patterns, performance troubleshooting).
  • Expertise in Linux/Unix administration in multi-platform, distributed environments.
  • Strong troubleshooting skills across hardware, OS, networking, and distributed storage layers (including diagnosing performance bottlenecks and failure scenarios).
  • Deep understanding of networking (Infiniband, Ethernet, DPDK, UCX), cloud computing, and distributed storage.
  • Experience with observability and monitoring stacks (Prometheus/Grafana and common log/metrics tooling).
  • Proficiency in Python and/or Bash; comfort building automation for monitoring, diagnostics, and repeatable operational tasks.
  • Excellent written and verbal communication skills, with the ability to explain complex technical topics to both technical and non-technical stakeholders.
It's Nice If You Have
  • Experience with Kubernetes/Containers and/or cloud platforms (AWS, Azure, OCI, GCP) in storage-heavy environments.
  • Familiarity with Jira, Confluence, Slack, and collaborating across Support, Engineering, and Product teams.
  • Experience supporting HPC or AI/ML infrastructure (GPU clusters, high-throughput networking, performance benchmarking).
  • Experience with infrastructure-as-code or config management (Ansible, Terraform, etc.).
  • Strong technical writing skills and a habit of creating reusable runbooks and playbooks.
What Success Looks Like
  • Customers view you as a trusted technical leader for Ceph and object storage, and they proactively engage you for architectural guidance and operational reviews.
  • You can independently assess Ceph cluster health, identify risk, and lead remediation plans that improve availability and performance.
  • You deliver consistent, high-quality customer communications and drive issues to resolution while partnering effectively with internal teams (CS, Product and Engineering).
  • You build durable artifacts (runbooks, dashboards, automation, postmortems) that raise the operational maturity of customer environments and WEKA's supportability.
  • You ramp quickly on WEKA's products and processes and expand your ownership beyond Ceph into broader designated services engagements.
The WEKA Way
  • We are Accountable: We take full ownership, always-even when things don't go as planned. We lead with integrity, show up with responsibility & ownership, and hold ourselves and each other to the highest standards.
  • We are Brave: We question the status quo, push boundaries, and take smart risks when needed. We welcome challenges and embrace debates as opportunities for growth, turning courage into fuel for innovation.
  • We are Collaborative: True collaboration isn't only about working together. It's about lifting one another up to succeed collectively. We are team-oriented and communicate with empathy and respect. We challenge each other and conduct positive conflict resolution. We are being transparent about our goals and results. And together, we're unstoppable.
  • We are Customer Centric: Our customers are at the heart of everything we do. We actively listen and prioritize the success of our customers, and every decision we make is driven by how we can better serve, support, and empower them to succeed. When our customers win, we win.