Systems Engineer position is a senior technical leader for the design, development, and deployment of key servers and cloud service platforms that run Driscoll’s infrastructure.
-
Design, build, and deploy the critical infrastructure servers, platforms, and cloud services needed to operate the applications and systems Driscoll’s requires.
-
Provide technical leadership and guidance to implement new systems, re-define services, management platforms, and other infrastructure components as needed.
-
Mentor team members to define or evaluate and follow best practices among the other systems engineering, administration, and technical staff.
-
Evaluate processes and automate common practices and procedures to streamline infrastructure operations, including basic system builds, deployments, and initial configuration of new accounts and services.
-
Document standard operating procedures for routing and complex tasks to be carried out by systems administrators and technicians.
-
Assist with incident resolution and major incident recovery of escalated issues.
-
Prepare the monitoring, feedback mechanisms, and customer engagement necessary to ensure that our infrastructure meets the needs of the business.
-
Ensuring high performance, high availability, and disaster recovery are designed into every deployment/project.
-
Design and implement monitoring, alerting, and incident response plans to ensure services and applications are available at the highest levels.
-
Bachelor's degree, or equivalent experience, in computer science and related technical field.
-
Job experience in infrastructure systems for 8+ years.
-
Strong teamwork, leadership, and technical knowledge skills.
-
Expert level configuration, management, and troubleshooting of Microsoft product platforms: O365, Exchange, Azure AD, Windows Active Directory, Okta, PowerShell, and SQL Server or Sharepoint.
-
Strong Infrastructure as Code experience with automation of standard processes and a rigorous procedural mindset including experience with tools such as Terraform, Packer, Ansible, Chef, Puppet, Octopus, Git.
-
Extensive virtualization experience, both on-prem and cloud (VMware, HyperV and AWS preferred) as well as hyperconverged technologies such Nutanix and Rubrik for backups.
-
Strong knowledge of data prioritization, storage, and management technologies to ensure performance meets the requirements of workload/workflow in cloud and on-premises.
-
Knowledge of application, system, and service monitoring tools (New Relic), alerting and notification, automated/scripted response, and triage/escalation process.
-
Expertise in patch management and vulnerability remediation with industry-leading tools and best practices.
-
Extensive experience with AWS services including IAM, EC2, S3, CloudWatch, CloudTrail, VPC, and Systems Manager are desired.
-
Strong analytical abilities: provide actionable insights from the data to identify and solve problems that may arise.
-
Extensive performance and reliability planning experience required. Continuous improvement mindset a strong predictor of success for this position.
-
Disaster recovery design, planning, and implementation.
-
ITIL process management experience.