2

Remote Contract Reliability Engineer Jobs in Toronto, ON

Senior Infrastructure Engineer

Toronto, ON · Remote

CA$170K - CA$220K/yr

Remote - Remote - Based In ET+2 / -3, NY Preferred Remote | Full-time Compensation: $170K - $220K ... reliability, and speed. This individual will bring a software engineering mindset to infrastructure ...

Contract Location: Remote Job Summary: Our customer's team is seeking a Game Tools & Modding Engineer with strong rendering and engine-level experience to build high-performance, real-time ...

As a remote-first organization, we are building a modern, high-performing team focused on ... Contribute to improving application performance, scalability, reliability, and maintainability ...

As a remote-first organization, we are building a modern, high-performing team focused on ... Contribute to improving application performance, scalability, reliability, and maintainability ...

As a remote-first organization, we are building a modern, high-performing team focused on ... Contribute to improving application performance, scalability, reliability, and maintainability ...

... scalability, and reliability. * Assist customers to resolve their performance problems ... This position can be fully remote for the right candidate.

Backend Engineer

Toronto, ON · Remote

CA$140K - CA$240K/yr

This is a fully remote position that offers a competitive salary range of $140,000 to $240,000 USD ... Continuously improve service performance, reliability, scalability, security, and developer ...

Remote within Canada Build the Future of Global Payments Veem is transforming how businesses move ... Establish engineering standards for architecture, reliability, observability, security, testing ...

Showing results 41-60

Remote Contract Reliability Engineer information

What is a remote contract reliability engineer?

Remote Contract Reliability Engineers are professionals who work remotely, usually on a contract basis, to ensure that systems, equipment, or software operate reliably and efficiently. Their main focus is on analyzing data, troubleshooting issues, and implementing improvements to enhance the dependability and performance of products or processes. They collaborate with teams virtually to identify potential failures, recommend solutions, and help organizations minimize downtime and maintenance costs. Their work spans various industries such as manufacturing, technology, and energy, and typically involves using specialized tools and methodologies to predict and prevent problems before they occur.

What are the key skills and qualifications needed to thrive as a remote contract reliability engineer, and why are they important?

To thrive as a Remote Contract Reliability Engineer, you need a solid background in reliability engineering, failure analysis, and maintenance planning, often supported by a degree in engineering and relevant industry experience. Familiarity with reliability analysis software (like ReliaSoft), asset management systems, and certifications such as Certified Reliability Engineer (CRE) are typically required. Strong problem-solving, communication, and self-motivation skills are essential for effectively collaborating and delivering results in a remote, contract-based environment. These skills ensure the engineer can optimize system reliability, reduce downtime, and meet client expectations efficiently from a remote location.

How does a remote contract reliability engineer typically collaborate with on-site teams and stakeholders?

As a Remote Contract Reliability Engineer, you will frequently collaborate with on-site teams through virtual meetings, project management platforms, and real-time data sharing tools. Strong communication skills are essential, as you'll provide recommendations, troubleshoot issues, and review maintenance or performance data remotely. You may also participate in regular status updates and coordinate with cross-functional teams—such as operations, maintenance, and safety—to implement reliability improvements and ensure asset uptime. Successful remote collaboration often relies on proactive communication and clear documentation of your analyses and recommendations.

What is the difference between Remote Contract Reliability Engineer vs Remote Contract Maintenance Technician?

AspectRemote Contract Reliability EngineerRemote Contract Maintenance Technician
CredentialsEngineering degree, certifications like Six Sigma or Reliability EngineeringTechnical diploma or certifications in maintenance or HVAC
Work EnvironmentDesigning reliability strategies, analyzing data remotely, consultingPerforming repairs, inspections, and preventive maintenance remotely or on-site
Employer & Industry UsageManufacturing, energy, aerospace industriesManufacturing plants, facilities management, industrial sectors

The Remote Contract Reliability Engineer focuses on analyzing systems, improving reliability, and providing remote consulting, while the Remote Contract Maintenance Technician handles hands-on repairs and maintenance tasks. Both roles may work remotely or on-site, but their core responsibilities and required skills differ significantly.

What are popular job titles related to Remote Contract Reliability Engineer jobs in Toronto, ON? For Remote Contract Reliability Engineer jobs in Toronto, ON, the most frequently searched job titles are:
What job categories do people searching Remote Contract Reliability Engineer jobs in Toronto, ON look for? The top searched job categories for Remote Contract Reliability Engineer jobs in Toronto, ON are:
Infographic showing various Remote Contract Reliability Engineer job openings in Toronto, ON as of August 2026, with employment types broken down into 92% Full Time, 2% Part Time, and 6% Contract. Highlights an 87% Physical, 4% Hybrid, and 9% Remote job distribution.

Senior Software Engineer (Java)

Fusemachines

Toronto, ON • Remote

Contractor

Posted 16 days ago


Job description

About Fusemachines
Founded in 2013, Fusemachines is a global provider of enterprise AI products and services, on a mission to democratize AI. Leveraging proprietary AI Studio and AI Engines, the company helps drive the clients’ AI Enterprise Transformation, regardless of where they are in their Digital AI journeys. With offices in North America, Asia, and Latin America, Fusemachines provides a suite of enterprise AI offerings and specialty services that allow organizations of any size to implement and scale AI. Fusemachines serves companies in industries such as retail,  manufacturing, and government.Fusemachines continues to actively pursue the mission of democratizing AI for the masses by providing high-quality AI education in underserved communities and helping organizations achieve their full potential with AI.Type: Full-Time, RemoteAbout the Role

We are seeking a Senior Software Engineer to build and operate high-performance backend services using Java. This role is ideal for an experienced backend engineer who understands distributed systems, testing, observability, and production reliability, with enough machine learning knowledge to integrate applications with models and feature stores.

You will work on systems with demanding availability, latency, throughput, and service-level requirements. You will collaborate with software, machine learning, data, infrastructure, and product teams to deliver reliable production solutions.

This is primarily a software engineering role. Deep machine learning research experience is not required.

Responsibilities
  • Design, build, test, and maintain scalable backend services using Java.

  • Develop APIs, microservices, and event-driven components for high-volume production systems.

  • Apply established Java architecture patterns, coding standards, and engineering best practices.

  • Integrate backend applications with machine learning models, inference endpoints, and feature stores.

  • Design reliable model-calling workflows with appropriate timeouts, retries, validation, and fallback behavior.

  • Build highly available, low-latency services that meet defined SLAs and service-level objectives.

  • Monitor and optimize latency, throughput, availability, resource utilization, and error rates.

  • Implement resilience patterns such as caching, circuit breakers, rate limiting, and graceful degradation.

  • Develop unit, integration, contract, performance, and end-to-end tests.

  • Implement logging, metrics, dashboards, distributed tracing, and actionable alerts.

  • Use platforms such as Datadog or similar tools to monitor systems and investigate production issues.

  • Participate in incident response, root-cause analysis, and reliability improvements.

  • Use AI-assisted coding tools responsibly to support development, testing, documentation, and debugging.

  • Participate in architecture discussions, technical design reviews, and code reviews.

  • Communicate technical decisions, risks, dependencies, and tradeoffs clearly to technical and nontechnical stakeholders.

  • Promote strong software engineering practices across the team.

Required Qualifications
  • Strong professional experience developing production applications with Java.

  • Experience designing backend services, APIs, microservices, or distributed systems.

  • Strong understanding of Java design patterns, object-oriented programming, and software architecture.

  • Experience building systems with demanding availability, scalability, throughput, or latency requirements.

  • Experience with automated testing and continuous integration and delivery practices.

  • Experience with relational or non-relational databases.

  • Experience implementing production logging, monitoring, metrics, tracing, and alerting.

  • Familiarity with Datadog, OpenTelemetry, Grafana, Prometheus, New Relic, or comparable platforms.

  • Understanding of SLAs, service-level indicators, service-level objectives, and production reliability.

  • Familiarity with cloud platforms, containers, and modern deployment environments.

  • Basic understanding of machine learning concepts and how applications interact with deployed models.

  • Strong troubleshooting, problem-solving, and production support skills.

  • Excellent written and verbal communication skills.

Preferred Qualifications
  • Experience integrating applications with machine learning inference services or feature stores.

  • Experience with low-latency or real-time decisioning systems.

  • Experience with Kafka, message queues, streaming platforms, or event-driven architectures.

  • Experience with Kubernetes, Redis, distributed caching, or performance testing.

  • Familiarity with model versioning, feature freshness, prediction logging, and controlled model rollouts.

  • Experience in advertising technology, digital marketplaces, auction systems, recommendation systems, or personalization.

  • Experience using AI-assisted development tools such as GitHub Copilot, Claude Code, or Cursor.

Fusemachines is an Equal Opportunities Employer, committed to diversity and inclusion. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other characteristic protected by applicable federal, state, or local laws.
 

Powered by JazzHR

fzTgSRZ2jw