ECS is seeking an experienced Sr. Infrastructure / DevSecOps Systems Engineer to work remotely providing scientific computing support for the work performed under this contract for NIH NIAID Enabling and Advancing Technologies (NEAT). This role will lead complex incident resolution, infrastructure automation, and advanced troubleshooting for scientific computing environments and deployment pipelines.
Responsibilities:
- Lead complex incident resolution, infrastructure automation, performance tuning, operational engineering standards, and advanced troubleshooting across enterprise and scientific computing environments
- Oversee CI/CD pipelines, container orchestration, infrastructure automation, security integration, configuration management, and reproducible scientific computing deployment pipelines
- Supervise, coordinate and/or perform additions and changes to network hardware and operating systems, and attached devices; including investigation, analysis, recommendation, configuration, installation, and testing of new network hardware and software
- Provide direct support in the day-to-day operations on network hardware and operating systems including the evaluation of system utilization, monitoring response time and primary support for detection and correction of operational problems
- Participate in planning design, technical review and implementation for new network infrastructure hardware and network operating systems for voice and data communication networks
- Maintain network infrastructure standards including network communication protocols such as TCP/IP
- Provide technical consultation, training and support to IT staff as designated by the government
Requirements:
- Experience in leading complex incident resolution and infrastructure automation
- Proficiency in performance tuning and operational engineering standards
- Advanced troubleshooting skills across enterprise and scientific computing environments
- Experience overseeing CI/CD pipelines and container orchestration
- Knowledge of infrastructure automation and security integration
- Experience with configuration management and reproducible scientific computing deployment pipelines
- Ability to supervise, coordinate and/or perform additions and changes to network hardware and operating systems
- Experience in investigating, analyzing, recommending, configuring, installing, and testing new network hardware and software
- Direct support experience in day-to-day operations on network hardware and operating systems
- Experience in evaluating system utilization and monitoring response time
- Primary support experience for detection and correction of operational problems
- Participation in planning design, technical review, and implementation for new network infrastructure hardware and network operating systems
- Knowledge of maintaining network infrastructure standards including network communication protocols such as TCP/IP
- Ability to provide technical consultation, training, and support to IT staff