Job Title: AEM Cloud Platform Reliability Engineer
Location: Remote
Duration: Long-term contract
Must Have Skills/Attributes: AEM, Cloud, Enterprise Architecture, Governance
Experience Desired: Adobe Experience Manager (AEM) as a Cloud Service (5+ yrs); AEM Platform Operations & Incident Management (5+ yrs); CI/CD, Cloud Deployments & Platform Automation (5+ yrs); Enterprise Integrations & Modern Web Platforms (5+ yrs); Platform Governance, Security & Performance Optimization (5+ yrs)
Job Description
- New platform - Vercel - is entering the AdobePlatform Services portfolio. Provide operational support for AdobeExperience Manager (AEM) as a Cloud Service platforms, including Author,Publish, Dispatcher, CDN, and supporting integrations.
- Monitor platform health, system performance,application availability, and service reliability using enterprisemonitoring and alerting tools.
- Troubleshoot and resolve production incidents,platform issues, deployment failures, and integration problems inpartnership with development, infrastructure, Adobe, and vendor supportteams.
- Manage and support CI/CD deployment pipelines,release activities, environment promotions, and application configurationchanges.
- Coordinate incident, problem, and root causeanalysis (RCA) activities to identify recurring issues and implementpreventive solutions.
- Support platform security, access management,certificate renewals, and compliance activities in accordance withenterprise standards.
- Maintain and optimize integrations between AEMand connected enterprise systems, including DAM, Workfront, Azureservices, Akamai/CDN, publishing partners, and other digital marketingplatforms.
- Perform environment administration activitiesincluding user management, permissions, workflow support, replicationmonitoring, and configuration management.
- Collaborate with application development teams toensure platform stability, scalability, performance optimization, andadherence to architecture standards.
- Participate in on-call support, major incidentresponse, and disaster recovery planning and testing activities.
- Create and maintain operational procedures,support documentation, knowledge articles, and platform runbooks.
- Evaluate platform enhancements, Adobe releases,and feature updates; recommend adoption strategies and implementationplans.
- Support capacity planning, performance tuning,and platform roadmap initiatives to ensure reliable service delivery.
- Partner with product owners, businessstakeholders, and technical teams to prioritize platform improvements andsupport requests.
- Lead continuous improvement initiatives focusedon operational excellence, automation, service reliability, and reductionof technical debt.
Position's Contributions to Work Group:
- This is net-new to the team. Neither has adefined service owner, documented processes, control mappings, supportmodel, or trained staff today.
- The Vercel Platform Support Engineer isresponsible for the operational support, administration, reliability, andcontinuous improvement of the Vercel hosting platform used for enterpriseweb applications.
- This role partners with development teams,product owners, architects, security teams, and external vendors to ensurehigh availability, performance, security, and scalability of applicationshosted on Vercel.
- The engineer will provide platform governance,incident management, deployment support, monitoring, user accessadministration, and technical consultation for teams building modern webapplications using Next.js, React, APIs, and cloud-native architectures.
- This is arriving on top of an existing portfolio(Adobe Platform, AEM Sites, AEM Assets, Target, Workfront, Akamai) thatthe team already runs at capacity. To bring this platform online securely,compliantly, and as repeatable service offerings - not as ad-hoc tooldependent on individual heroics.
- Their mandate is to stand each platform up toLevel 3 (Defined) maturity against our published Service Maturity Model,complete all 21 required service-offering artifacts, evidence ClientGeneral Control (ITGC) and Security Directive alignment, train theexisting team, and establish proactive operational processes before theseplatforms carry production, customer-facing traffic.
Interaction with team:
Working with PO, team of 6-7 of platform engineers ( various levels), also working with other support teams ( level 1 or level 2 )
Required Technical Skills:
- 5+ years maintaining AM platform
- Demonstrates expertise supporting AEM as a CloudService environment and associated digital experience platforms.
- Leads incident resolution, root cause analysis,and service restoration efforts for critical business applications.
- Applies platform monitoring, observability, andautomation techniques to improve service reliability.
- Possesses strong knowledge of AEM architecture,Dispatcher/CDN configurations, cloud deployments, and enterpriseintegrations.
- Leads incident resolution, root cause analysis,and service restoration efforts for critical business applications.
- Applies platform monitoring, observability, andautomation techniques to improve service reliability.
- Possesses strong knowledge of AEM architecture,Dispatcher/CDN configurations, cloud deployments, and enterpriseintegrations.
- Experience in providing platform governance,incident management, deployment support, monitoring, user accessadministration, and technical consultation for teams building modern webapplications using Next.js, React, APIs, and cloud-native architectures.