ME00616-Cloud System Administrator 2
Momentum Engineering, Inc.
- Location
- Onsite (Annapolis Junction, Maryland)
- Compensation
- $150k - $205k/yr
- Employment
- Full-time
- Level
- Mid Level
Posted 4 weeks ago
About the Role
Momentum Engineering seeks a Cloud Systems Administrator to support a mission-driven Cyber platform team. The role involves sustaining and optimizing enterprise cloud platforms for large-scale analytics and mission operations.
Skills
Linux Administration
Kubernetes
Docker
Hadoop
Python
Bash
Red Hat Enterprise Linux
Network Diagnostics
Prometheus
Grafana
Ansible
Salt
DNS/DHCP Management
Cloud Infrastructure
Troubleshooting
Jira
Benefits
- Medical
- Dental
- Vision
- Life Insurance
- Short Term Disability
- Long Term Disability
- PTO
- Paid Holidays
Full job details
Job Summary
- Seeking a Cloud Systems Administrator to support a mission-driven Cyber platform team responsible for sustaining and optimizing enterprise cloud platforms supporting large-scale analytics and mission operations
- The successful candidate will support day-to-day operations of cloud-based environments built on Java and Free and Open-Source Software (FOSS) technologies, including Kubernetes, Hadoop, and Accumulo
- This role is responsible for maintaining platform stability, supporting customer operations, troubleshooting complex Linux-based systems, and ensuring the availability, performance, and security of mission-critical infrastructure
- The Cloud Systems Administrator will provide Tier 1 through Tier 3 operational support and participate in an on-call rotation supporting enterprise cloud platforms in a dynamic, high-tempo operational environment
Primary Responsibilities
- Support the daily operations, administration, and sustainment of cloud-based enterprise analytics platforms
- Provide Tier 1, Tier 2, and Tier 3 operational support for cloud infrastructure, operating systems, applications, and distributed platforms
- Monitor system health, performance, and availability while proactively identifying and resolving operational issues
- Troubleshoot Linux operating system, network, application, and infrastructure issues to maintain continuous mission operations
- Install, configure, maintain, and provision Red Hat Enterprise Linux (RHEL), CentOS, and similar Linux environments using Kickstart and automated deployment tools
- Perform Linux and UNIX network diagnostics using tools such as nmap, tcpdump, and other troubleshooting utilities
- Administer and support DNS, DHCP, hosts file management, and related network services
- Support containerized applications and distributed computing environments utilizing Kubernetes, Docker, Hadoop, and HDFS
- Collaborate with software engineers, cloud engineers, and infrastructure teams to diagnose and resolve system performance, scalability, and reliability issues
- Develop and maintain automation scripts using Python and Bash to streamline operational tasks and improve system efficiency
- Utilize monitoring and observability tools such as Prometheus and Grafana to identify, analyze, and resolve system issues
- Support configuration management and infrastructure automation using tools such as Salt, Ansible, or similar technologies
- Document operational procedures, troubleshooting activities, and system configurations while maintaining accurate technical documentation
- Track operational issues, incidents, and change requests using Jira
- Participate in a rotating on-call schedule to provide after-hours operational support and respond to mission-critical incidents
Required Qualifications
- Must have active Top Secret/SCI clearance with NSA Full Scope Polygraph
- Bachelor's degree in Computer Science, Information Systems, Engineering, Mathematics, or a related technical discipline, or equivalent combination of education and experience
- Experience administering Linux operating systems, including Red Hat Enterprise Linux (RHEL), CentOS, or similar distributions
- Experience supporting enterprise cloud environments and distributed computing platforms
- Experience troubleshooting Linux operating systems, infrastructure, networking, and cloud platform issues
- Experience with containerization technologies such as Docker and Kubernetes
- Experience administering DNS, DHCP, and Linux networking services
- Experience developing automation scripts using Python and/or Bash
- Strong analytical, troubleshooting, and problem-solving skills
- Excellent written and verbal communication skills
- Ability to work independently within a fast-paced operational environment while effectively collaborating with multidisciplinary technical teams
- Ability and willingness to participate in an on-call operational support rotation
Desired Qualifications
- Experience supporting large-scale data analytics platforms utilizing Hadoop, HDFS, or Accumulo
- Experience with cloud platforms such as AWS or OpenStack
- Experience using Prometheus, Grafana, or similar enterprise monitoring and observability platforms
- Experience with configuration management and automation tools such as Salt or Ansible
- Familiarity with Java application environments and Free and Open-Source Software (FOSS) technologies
- Experience supporting enterprise cybersecurity or Intelligence Community mission environments
- Experience working within Agile or DevSecOps environments
- Familiarity with Infrastructure as Code (IaC), CI/CD pipelines, and automated deployment processes
- Experience using Jira for issue tracking, incident management, and operational workflow management
Exempt hourly position. 11 paid holidays, minimum of 3 weeks PTO, company sponsored group medical plan, company paid dental, vision, life insurance, and STD/LTD plans. Salary is dependent upon the candidate’s experience and qualifications.