About the role

Responsibilities Include:
  • Responsible for all incidents, service requests and all alerts.
  • Responsible for driving all P1, P2, incidents, change requests and OEM/ vendor management for the tower
  • Driving initiatives to improve SLAs by analysis and responses.
  • Provide technical expertise for system transitions, migrations, and consolidations.
  • Be able to work independently and lead/support multiple projects/clients.
  • Responsible for Technology refresh in respective domain.
  • Monitor and tune the system to achieve optimum performance levels in standalone and multi-tiered environments.
  • Monitor online performance of all servers and take appropriate action to address performance issues.
  • Conduct system analysis, configuration management and develops improvements for system software performance, availability, and reliability.
  • Provide, maintain, and execute Online & Operating Systems availability health checks.
  • Perform server installation, software installation, server decommission, server update, server migration, transition new account, scripting, changes to the systems and other engineer tasks/projects within the Service Level Agreements.
  • Design, develops, recommends, and implements new or revised system software, utilities, and automated processes as necessary.
  • Propose recommendations to improve the server management process.
  • Provide in-depth diagnosis for operating systems software/hardware failures and develops solutions.
  • Work with appropriate hardware vendors / principle to keep all system up to date on firmware levels at the latest.
  • Recommend system performance enhancements (i.e. configuration, H/W, S/W, consolidation of workloads, etc.)Implement recommendations to improve the server management process.
  • Implement root cause analysis recommendations as requested/assigned for respective areas of service responsibility.
  • Ensure server data integrity by evaluating, implementing, and managing appropriate software and hardware solutions.
  • Prescribe system backup / disaster recovery procedures and directs recovery operations in the event of destruction of all or part of the operating system or other system components.
  • Implement appropriate levels of system security and standard best practices.
  • Work with internal and external auditors to identify and mitigate any risks
  • Focus area in this domain would be to increase the compliance on house-keeping tasks for the systems in this domain through automation.
Posting Date: Department: Engineering-Platform Data - 1486250021

Matching similar jobs

JOB OVERVIEW

Experience level

Lead

Location

Austin, TX

Occupation

Network and Computer Systems Administrators

Industry

Computer Systems Design Services

Posted

18 days ago

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE