About the role

FULLY REMOTE: ScolerTec is seeking multiple Data Site Reliability Engineers (SREs) to support a large-scale cloud data modernization and governance program. The Data SRE will provide technical leadership for the reliability, availability, performance, observability, security, and continuous improvement of cloud-based data platforms, pipelines, applications, and analytics services. This role will work across AWS infrastructure, DevSecOps, CI/CD, monitoring, incident management, disaster recovery, automation, and cloud cost optimization.
Key Responsibilities: Monitor and maintain the reliability, availability, health, and performance of cloud data platforms and services. Implement real-time monitoring, dashboards, alerting, event management, and operational reporting. Lead incident detection, diagnosis, escalation, root-cause analysis, and resolution. Improve system reliability, observability, automation, scalability, and operational resilience. Implement automated health checks across operating systems, applications, databases, and PaaS services. Design and support high availability, disaster recovery, automated failover, and continuity-of-operations capabilities. Integrate SSO, MFA, IAM, RBAC, least-privilege access, and secrets management. Integrate security tooling such as SAST, DAST, SCA, CSPM, SIEM, and SBOM processes into DevSecOps pipelines. Build and maintain CI/CD and automated deployment pipelines. Automate infrastructure and operational activities using Terraform, CloudFormation, Python, Bash, PowerShell, or similar tools. Support containerized application and platform environments. Perform cloud infrastructure, database, and application performance tuning. Monitor CPU, memory, disk, network, application, database, and platform utilization. Develop operational thresholds, alerting rules, escalation procedures, and runbooks. Support canary releases and controlled deployment strategies. Perform FinOps, cloud cost analysis, forecasting, and optimization. Support security audits, compliance requirements, and operational documentation.
Qualifications: 5+ years of experience in systems engineering, cloud engineering, SRE, DevOps, platform engineering, or related technical roles. Bachelor's degree in Computer Science, Software Engineering, or related field. Strong hands-on experience with AWS infrastructure and services, including compute, networking, storage, monitoring, and logging. Hands-on experience with Terraform, AWS CloudFormation, or comparable Infrastructure as Code technologies. Experience with monitoring/APM platforms such as CloudWatch, Datadog, Prometheus, Grafana, New Relic, or similar tools. Strong experience with incident response, problem management, root-cause analysis, and production support. Experience with CI/CD and GitOps platforms such as Jenkins, GitHub Actions, GitLab, or ArgoCD.Strong scripting experience using Python, Bash, or PowerShell. Experience implementing DevSecOps practices. Experience supporting containerized applications and cloud platforms. Experience with enterprise CI/CD build and release processes. Experience with cloud cost optimization and FinOps practices. Familiarity with federal security/compliance frameworks such as FedRAMP, NIST, or similar frameworks preferred. Preferred CertificationsAWS SysOps Administrator, AWS Solutions Architect, AWS DevOps Engineer, ITIL, Google Professional Cloud DevOps Engineer, or related cloud/SRE certifications.

Matching similar jobs

JOB OVERVIEW

Experience level

Senior

Location

Jacksonville, FL

Occupation

Computer Systems Engineers/Architects

Industry

Computer Systems Design Services

Posted

today

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE