They will develop GCP cloud logging and monitoring reports to support visibility across the platform. Activities are comprised of: 1. SLA & Reliability Reporting 1. Establish the initial framework ...
Quick apply
They will develop GCP cloud logging and monitoring reports to support visibility across the platform. Activities are comprised of: 1. SLA & Reliability Reporting 1. Establish the initial framework ...
Quick apply
They will develop GCP cloud logging and monitoring reports to support visibility across the platform. Activities are comprised of: 1. SLA & Reliability Reporting 1. Establish the initial framework ...
New York, NY ยท On-site
$150K - $200K/yr
The platform processes millions of clinical documents monthly across multi-tenant deployments in ... Take ownership of at least one customer deployment end-to-end - monitoring, alerting, incident ...
New York, NY ยท On-site
$150K - $200K/yr
The platform processes millions of clinical documents monthly across multi-tenant deployments in ... Take ownership of at least one customer deployment end-to-end - monitoring, alerting, incident ...
Manhattan, NY ยท On-site
Own API versioning, rate limiting, SLA monitoring, and client onboarding * Security & Access ... Ensure the platform scales from a small number of pilot clients to 2,000+ users without ...
Manhattan, NY ยท On-site
Own API versioning, rate limiting, SLA monitoring, and client onboarding * Security & Access ... Ensure the platform scales from a small number of pilot clients to 2,000+ users without ...
Manhattan, NY ยท On-site
The team is also responsible for the production environment which requires knowledge with monitoring, alerting and on-call practices. The Platform team performs many functions whether that is ...
Manhattan, NY ยท On-site
The team is also responsible for the production environment which requires knowledge with monitoring, alerting and on-call practices. The Platform team performs many functions whether that is ...
New York, NY ยท On-site
As a QIS Platform Analyst, you will operate at the intersection of Python engineering, data, AI and front office processes, with direct ownership of platform analytics, intelligent monitoring and ...
New York, NY ยท On-site
As a QIS Platform Analyst, you will operate at the intersection of Python engineering, data, AI and front office processes, with direct ownership of platform analytics, intelligent monitoring and ...
Farmingdale, NY ยท On-site
$23 - $25/hr
Assemble and prepare delivery routes for next-day operations using the FarEye platform. * Monitor route execution throughout the day and track driver progress to ensure timely deliveries and ...
Quick apply
Farmingdale, NY ยท On-site
$23 - $25/hr
Assemble and prepare delivery routes for next-day operations using the FarEye platform. * Monitor route execution throughout the day and track driver progress to ensure timely deliveries and ...
Farmingdale, NY ยท On-site
$23 - $25/hr
Assemble and prepare delivery routes for next-day operations using the FarEye platform. * Monitor route execution throughout the day and track driver progress to ensure timely deliveries and ...
Farmingdale, NY ยท On-site
$23 - $25/hr
Assemble and prepare delivery routes for next-day operations using the FarEye platform. * Monitor route execution throughout the day and track driver progress to ensure timely deliveries and ...
Farmingdale, NY ยท On-site
$23 - $25/hr
Assemble and prepare delivery routes for next-day operations using the FarEye platform. * Monitor route execution throughout the day and track driver progress to ensure timely deliveries and ...
Farmingdale, NY ยท On-site
$23 - $25/hr
Assemble and prepare delivery routes for next-day operations using the FarEye platform. * Monitor route execution throughout the day and track driver progress to ensure timely deliveries and ...
Manhattan, NY ยท On-site
$147K - $170K/yr
... monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable. - Drive platform reliability and operational excellence through incident response, root ...
Manhattan, NY ยท On-site
$147K - $170K/yr
... monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable. - Drive platform reliability and operational excellence through incident response, root ...
Manhattan, NY ยท On-site
$147K - $170K/yr
... monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable. - Drive platform reliability and operational excellence through incident response, root ...
Manhattan, NY ยท On-site
$147K - $170K/yr
... monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable. - Drive platform reliability and operational excellence through incident response, root ...
Manhattan, NY ยท On-site
$147K - $170K/yr
... monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable. - Drive platform reliability and operational excellence through incident response, root ...
Manhattan, NY ยท On-site
$147K - $170K/yr
... monitoring, logging, tracing, alerting, and SLO practices to keep the developer platform stable and measurable. - Drive platform reliability and operational excellence through incident response, root ...
Manhattan, NY ยท On-site
Monitor platform support Slack channels and address customer questions and issues* Stay up-to-date with the latest Google Workspace features and best practices, implementing updates and improvements ...
Manhattan, NY ยท On-site
Monitor platform support Slack channels and address customer questions and issues* Stay up-to-date with the latest Google Workspace features and best practices, implementing updates and improvements ...
Manhattan, NY ยท On-site +1
Dario is seeking a hands-on Platform Engineer with at least three years of experience to join our ... Manage, monitor, and optimize AWS cloud infrastructure. * Build and maintain CI/CD pipelines and ...
Manhattan, NY ยท On-site +1
Dario is seeking a hands-on Platform Engineer with at least three years of experience to join our ... Manage, monitor, and optimize AWS cloud infrastructure. * Build and maintain CI/CD pipelines and ...
Integrate monitoring, observability, alerting, and security scanning into platform and application delivery processes. * Serve as a senior escalation point for production incidents and drive root ...
Integrate monitoring, observability, alerting, and security scanning into platform and application delivery processes. * Serve as a senior escalation point for production incidents and drive root ...
Manhattan, NY ยท On-site
Implement data quality monitoring with traceability back to source systems * Define schemas ... Collaborate on platform architecture decisions and help establish engineering best practices
Manhattan, NY ยท On-site
Implement data quality monitoring with traceability back to source systems * Define schemas ... Collaborate on platform architecture decisions and help establish engineering best practices
Manhattan, NY ยท On-site
Implement data quality monitoring with traceability back to source systems * Define schemas ... Collaborate on platform architecture decisions and help establish engineering best practices
Manhattan, NY ยท On-site
Implement data quality monitoring with traceability back to source systems * Define schemas ... Collaborate on platform architecture decisions and help establish engineering best practices
New York, NY ยท On-site +1
$150K - $250K/yr
... monitoring standards that translate into actionable insights for product teams downstream. You've ... founding Platform Engineer at a fintech company where your work directly determines what we can ...
New York, NY ยท On-site +1
$150K - $250K/yr
... monitoring standards that translate into actionable insights for product teams downstream. You've ... founding Platform Engineer at a fintech company where your work directly determines what we can ...
New York, NY ยท On-site
We're hiring a Platform Engineer to help world-class financial institutions automate away their ... Ensure systems are resilient through monitoring, alerting, and automated recovery About You * 5+ ...
New York, NY ยท On-site
We're hiring a Platform Engineer to help world-class financial institutions automate away their ... Ensure systems are resilient through monitoring, alerting, and automated recovery About You * 5+ ...
We're hiring a Platform Engineer to help world-class financial institutions automate away their ... Ensure systems are resilient through monitoring, alerting, and automated recovery About You * 5+ ...
We're hiring a Platform Engineer to help world-class financial institutions automate away their ... Ensure systems are resilient through monitoring, alerting, and automated recovery About You * 5+ ...
Manhattan, NY ยท On-site
Automate the deployment, scaling, and monitoring of Kubernetes clusters * Assist with developing ... Keep pace with emerging tools, techniques, and cloud platforms * Experience building solutions in ...
Manhattan, NY ยท On-site
Automate the deployment, scaling, and monitoring of Kubernetes clusters * Assist with developing ... Keep pace with emerging tools, techniques, and cloud platforms * Experience building solutions in ...
$33.98 - $39.73
3% of jobs
$39.73 - $45.48
9% of jobs
$45.48 - $51.23
12% of jobs
$51.95 is the 25th percentile. Wages below this are outliers.
$51.23 - $56.98
13% of jobs
$56.98 - $62.74
12% of jobs
The median wage is $63.50 / hr.
$62.74 - $68.49
16% of jobs
$73.98 is the 75th percentile. Wages above this are outliers.
$68.49 - $74.24
12% of jobs
$74.24 - $79.99
10% of jobs
$79.99 - $85.74
5% of jobs
$85.74 - $91.50
5% of jobs
$91.50 - $97.25
4% of jobs
$33
$65
$97
New York, NY โข On-site
Contractor
Re-posted 3 days ago
Role : GCP Agentic Platform Support Lead
Location : New York, NY 10019 (Need local candidates/Hybrid)
Client: Persistent
Detailed JD:
The platform support lead will set the foundation and requirements for support on the GCP Data & AI platform. They will define standards for platform health, managing incident resolution, and executing routine maintenance to support the platform. They will develop GCP cloud logging and monitoring reports to support visibility across the platform.
Activities are comprised of:
1. SLA & Reliability Reporting
1. Establish the initial framework for tracking Mean Time to Repair (MTTR) and Mean Time Between Failures (MTBF)
2. Configure self-service billing and uptime dashboards for Con Edison stakeholders
2. Foundation, Maintenance & Optimization
1. Develop and deploy the initial suite of Cloud Logging and Monitoring reports to establish platform visibility
2. Monitor GCP billing for anomalies (e.g., BigQuery slot spikes) and implement tactical fixes to ensure budget adherence
3. Build and maintain the "Golden Path" runbooks to ensure operational procedures are documented as they are established
3. Platform Monitoring & Incident Management
1. Conduct solo reviews of overnight batch processing logs (e.g., Cloud Composer/Dataflow) to verify completion and identify failures before business hours progress
2. Receive and prioritize platform-related tickets; determine if issues stem from infrastructure, pipelines, or upstream sources
3. Execute root cause analysis (RCA) and apply fixes for code-based failures, IAM errors, or configuration drifts
4. Act as the primary technical point of contact for Google Cloud Support or Con Edison Source System teams (SAP, GIS) when issues are external to the platform
4. Minor Enhancements (Capacity-Based
1. Maintain a prioritized backlog of minor requests to be addressed only after platform stability and incidents are managed
2. Within available bandwidth, execute minor schema updates, ingestion schedule tweaks, or IAM modifications
Workstream Deliverables:
1. Operations Runbook: The definitive MS Word resource reflecting current operational procedures and recovery steps (MS Word)
2. Integrated Health & Cost Reporting: Automated tracking of service uptime and GCP spend via Cloud Monitoring (Cloud Monitoring Reports)
3. Unified Incident & RCA Logs: A centralized record of Critical/High severity incidents and their resolutions, stored in the agreed management tool (ServiceNow/Jira or similar)
4. Recovery & Maintenance Code: Validated code merged into the repository for bug fixes and configuration updates, including detailed release notes (GCP Code)