About the role

Owning incidents from detection to resolution, the full-time remote Platform Operations Engineer will manage communications and diagnostics for platform issues, ensuring effective troubleshooting and stakeholder updates.
Key responsibilities: Receive and process requests through Telegram, Slack, and email while clarifying the nature of issues Perform issue localization and basic infrastructure troubleshooting, including logs and service status Take end-to-end ownership of incidents, escalating to L2/L3 support with prepared context and creating tickets as necessaryRequired qualificationsExperience with monitoring and logging systems such as Grafana, Kibana, or Loki Basic understanding of Kubernetes and application configuration management Background in QA, technical support, or a similar role Ability to read and analyze logs for issue localization Willingness to work night shifts covering European hours

Matching similar jobs

JOB OVERVIEW

Experience level

Lead

Location

New York, NY

Occupation

Computer Systems Engineers/Architects

Industry

Computer Systems Design Services

Posted

9 days ago

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE