Job Summary:
RIT Solutions, Inc. is seeking an experienced DevOps Engineer with strong hands-on expertise in AWS and knowledge of Google Cloud Platform. The role involves designing and maintaining scalable infrastructure, implementing Infrastructure as Code, and collaborating with teams to improve platform reliability and automation.
Responsibilities:
• Design, build, and maintain scalable and highly available infrastructure in AWS
• Deploy, manage, and troubleshoot Kubernetes clusters (EKS, GKE, or self‐managed)
• Implement and manage Infrastructure as Code using Terraform and AWS CloudFormation
• Build and support CI/CD pipelines to automate application build, test, and deployment processes
• Containerize applications using Docker and manage deployment strategies (rolling, blue/green, canary)
• Monitor system performance, availability, and reliability using cloud‐native and third‐party monitoring tools
• Collaborate with development teams to improve DevOps practices and cloud architecture
• Implement security best practices, including IAM, secrets management, and policy enforcement
• Troubleshoot production issues and participate in on‐call or incident response rotations as needed
• Optimize cloud environments for performance, cost, and scalability
Qualifications:
Required:
• 7+ years of experience in a DevOps, SRE, or Cloud Engineering role
• Strong hands‐on experience with AWS and/or GCP
• Proven experience managing Kubernetes in production environments
• Extensive experience with Terraform and CloudFormation (writing, maintaining, and refactoring IaC)
• Solid experience with CI/CD tools (Jenkins, GitHub Actions, GitLab CI, Azure DevOps, etc.)
• Strong knowledge of Linux/Unix systems and networking fundamentals
• Proficiency in scripting (Bash, Python, or similar)
• Experience with Docker and container orchestration best practices
• Familiarity with logging, monitoring, and observability tools
Preferred:
• Experience with Helm and Kubernetes package management
• Multi‐account / multi‐project cloud architecture experience
• Experience with security automation, vulnerability scanning, or compliance tooling
• Exposure to Site Reliability Engineering (SRE) principles
• Cloud certifications (AWS, GCP, Kubernetes) are a plus
Company:
Jobdiva Job Portal: https://www1.jobdiva.com/candidates/myjobs/searchjobsdone.jsp?a=xbjdnwgjodtga1y1im2g881fkkeiwd0775lbvq8yqgps8vb2q36w2vj1ga6xxork&compid=-1 Recruitment (contingency search and campus selection). Founded in 2019, the company is headquartered in Arlington, USA, with a team of 201-500 employees. The company is currently Growth Stage.