Cloud System Administrator 2
Job description
At Wyetech, you’ll be at the center of an award-winning corporate culture, breaking technological barriers and solving real-world problems for our federal government customers. We are committed to hiring the best of the best, and in return, we offer a world-class, truly unique employee experience that is rare within our industry. We are seeking a highly skilled Senior Cloud & Distributed Systems Engineer to support mission-critical, cloud-based data repositories serving thousands of Intelligence Community users. This role operates in a dynamic, high-tempo operational environment where requirements evolve rapidly in response to global events. The selected candidate will administer and engineer large-scale Hadoop and Accumulo clusters, ensure system reliability and security, and collaborate across infrastructure, networking, and security domains to maintain continuous mission availability. This is not a traditional system administration role — it is a reliability-focused distributed systems engineering position operating at scale. Due to federal contract requirements, United States Citizenship and position appropriate security clearance is required. (e.g. Active TS/SCI security clearance with agency appropriate polygraph).
Capabilities
Required Qualifications
Required Technical Skills
Linux Systems Administration (7+ Years)
- Deep understanding of Linux operating systems and internals
- User and group account management (LDAP)
- Configuration and administration of DHCP, DNS, and TFTP
- System patching, upgrades, and security hardening
- Performance tuning and resource optimization
Distributed Systems & Cluster Administration (3+ Years)
Experience supporting large distributed systems consisting of:
• Multiple clusters
• Clusters spanning at least three racks
• Minimum of 60 nodes per site
Experience with:
• Hadoop (HDFS, YARN tuning)
• Accumulo (tablet balancing, performance optimization)
• Cassandra, Scality, Swift, Gluster, Lustre, GPFS, Amazon S3, or comparable technologies
Cloud & Container Technologies
- Kubernetes orchestration services (CKA-level knowledge preferred)
- Docker containerization and image management
- Helm charts and cluster configuration
- StatefulSets and persistent volume management
- Cloud-based storage architectures
Automation & Infrastructure as Code
• 5+ years scripting in Bash, Python, or Perl
• Experience with configuration management tools:
o Puppet
o Ansible
o Salt
• Infrastructure as Code:
o Terraform or CloudFormation
• CI/CD pipeline integration and Git-based workflows
Observability & Reliability Engineering
• Experience implementing and managing monitoring solutions such as:
o Prometheus / Grafana
o ELK / OpenSearch
o Splunk
o Cloud-native monitoring platforms
• Design and tuning of alerting frameworks
• Experience defining and supporting SLAs/SLOs
• Incident response participation and documentation
• Capacity planning and performance analysis
Networking & Infrastructure
- Understanding of VLANs, port channel bonding, and Layer 2/Layer 3 interactions
- TCP/IP troubleshooting
- Load balancing (F5, HAProxy, NGINX)
- Firewall rule management
- Network performance analysis
Storage & High Availability
- RAID and storage architecture knowledge
- Object storage optimization
- Data replication and backup strategies
- Multi-site failover and disaster recovery (DR) planning
- RPO/RTO considerations
- Active/Active or Active/Passive cluster design
Security & Compliance
- System hardening (STIG implementation preferred)
- Vulnerability scanning tools (e.g., ACAS/Nessus)
- RMF familiarity
- Security logging and audit compliance
- Experience operating in TS/SCI environments
One of the following certifications is required:
- AWS Certified SysOps Administrator – Associate
- AWS DevOps Engineer – Professional
- Certified Kubernetes Administrator (CKA)
Operational Environment
- Mission-critical cloud repositories supporting thousands of users
- High-tempo, operationally responsive environment
- Daily interaction with infrastructure, hardware, and security teams
- Requirements shift in response to world events and mission needs
- Emphasis on automation-first operations and continuous improvement
Ideal Candidate Profile
The successful candidate:
• Thinks like a reliability engineer, not just a system administrator
• Automates repetitive processes and improves operational maturity
• Remains calm and analytical during high-impact incidents
• Understands distributed systems behavior at scale
• Communicates effectively across technical domains
• Thrives in mission-driven, dynamic environments
The Benefits Package
Additional benefits include:
Full-time employees have the option to participate in a variety of voluntary benefit plans including:
Company Environment & Perks
Wyetech, LLC is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or status as a protected veteran. Affirmative Action Statement: Wyetech, LLC is committed to the principles of affirmative action in all hiring and employment for minorities, women, individuals with disabilities, and protected veterans. Accommodations: Wyetech, LLC is committed to providing an inclusive and accessible hiring process. If you need any accommodations during the application or interview process, please contact Brittney Wood. at 844-WYETECH x727 or staffing@wyetech.com. We are happy to provide reasonable accommodations to ensure equal access to all candidates.