Capabilities
- Monitor system health, performance, and reliability across large, distributed clusters
- Troubleshoot and resolve complex hardware, software, network, and cloud platform issues
- Perform root cause analysis (RCA) and contribute to post-incident reviews
- Maintain and optimize large-scale Hadoop and Accumulo environments
- Administer and maintain distributed storage systems
- Engineer and improve monitoring, observability, and alerting frameworks
- Participate in architecture and engineering design discussions
- Create and maintain automation scripts and Infrastructure-as-Code deployments
- Patch, upgrade, and harden systems in accordance with security compliance standards
- Administer LDAP-based user and group accounts
- Maintain hardware inventory and asset tracking
- Provide after-hours on-call support in a mission-driven environment
- Interface with hardware, network, infrastructure, and security teams
Required Qualifications
- TS/SCI with agency appropriate poly
- Minimum 3 years experience administering large distributed systems
- 7 years Linux systems administration experience
- 5 years scripting experience
- Bachelor’s degree in Engineering, Systems Engineering, Computer Science, Mathematics, or related field highly desired
- May substitute for two (2) years of experience
Required Technical Skills
Linux Systems Administration (7+ Years)
• Deep understanding of Linux operating systems and internals
• User and group account management (LDAP)
• Configuration and administration of DHCP, DNS, and TFTP
• System patching, upgrades, and security hardening
• Performance tuning and resource optimization
Distributed Systems & Cluster Administration (3+ Years)
Experience supporting large distributed systems consisting of:
• Multiple clusters
• Clusters spanning at least three racks
• Minimum of 60 nodes per site
Experience with:
• Hadoop (HDFS, YARN tuning)
• Accumulo (tablet balancing, performance optimization)
• Cassandra, Scality, Swift, Gluster, Lustre, GPFS, Amazon S3, or comparable technologies
Cloud & Container Technologies
• Kubernetes orchestration services (CKA-level knowledge preferred)
• Docker containerization and image management
• Helm charts and cluster configuration
• StatefulSets and persistent volume management
• Cloud-based storage architectures
Automation & Infrastructure as Code
• 5+ years scripting in Bash, Python, or Perl
• Experience with configuration management tools:
o Puppet
o Ansible
o Salt
• Infrastructure as Code:
o Terraform or CloudFormation
• CI/CD pipeline integration and Git-based workflows
Observability & Reliability Engineering
• Experience implementing and managing monitoring solutions such as:
o Prometheus / Grafana
o ELK / OpenSearch
o Splunk
o Cloud-native monitoring platforms
• Design and tuning of alerting frameworks
• Experience defining and supporting SLAs/SLOs
• Incident response participation and documentation
• Capacity planning and performance analysis
Networking & Infrastructure
• Understanding of VLANs, port channel bonding, and Layer 2/Layer 3 interactions
• TCP/IP troubleshooting
• Load balancing (F5, HAProxy, NGINX)
• Firewall rule management
• Network performance analysis
Storage & High Availability
• RAID and storage architecture knowledge
• Object storage optimization
• Data replication and backup strategies
• Multi-site failover and disaster recovery (DR) planning
• RPO/RTO considerations
• Active/Active or Active/Passive cluster design
Security & Compliance
• System hardening (STIG implementation preferred)
• Vulnerability scanning tools (e.g., ACAS/Nessus)
• RMF familiarity
• Security logging and audit compliance
• Experience operating in TS/SCI environments
One of the following certifications is required:
• AWS Certified SysOps Administrator – Associate
• AWS DevOps Engineer – Professional
• Certified Kubernetes Administrator (CKA)
Operational Environment
• Mission-critical cloud repositories supporting thousands of users
• High-tempo, operationally responsive environment
• Daily interaction with infrastructure, hardware, and security teams
• Requirements shift in response to world events and mission needs
• Emphasis on automation-first operations and continuous improvement
Ideal Candidate Profile
The successful candidate:
• Thinks like a reliability engineer, not just a system administrator
• Automates repetitive processes and improves operational maturity
• Remains calm and analytical during high-impact incidents
• Understands distributed systems behavior at scale
• Communicates effectively across technical domains
• Thrives in mission-driven, dynamic environments
The Benefits Package
- Wyetech believes in generously supporting employees as they prepare for retirement. The company automatically contributes 20% of each employee's gross compensation to a Simplified Employee Pension (SEP) IRA, with no requirement for employee matching. All contributions are fully vested from day one, ensuring immediate ownership of retirement funds.
- Wyetech provides a generous PTO plan of up to 200 hours annually, aligned with applicable state leave regulations. Employees have the flexibility to adjust their PTO allocation at the start of each calendar year, ensuring it meets their evolving needs.
- A Choice of Medical Plan Options, some with Health Savings Account (HSA)
- Vision and Dental
- Life and AD&D Benefits
- Short and Long-Term Disability
- Hospital Indemnity, Accident, and Critical Illness Insurances
- Optional Identity Theft and Legal Protection Services
Company Environment & Perks
- Employee Referral Bonus Eligibility up to $10,000
- Mobility Among Wyetech-supported Contracts
- Various contract and work locations throughout Maryland, Virginia, Colorado, Texas, Utah, Alaska, Hawaii and OCONUS
- Various team-building events throughout the year such as: monthly lunches, summer company picnic, and an annual holiday party.
- Employees receive two complementary branded clothing orders annually.
Skills Required
- Active TS/SCI security clearance with agency-appropriate polygraph
- United States citizenship
- At least 3 years administering large distributed systems
- At least 7 years of Linux systems administration experience
- At least 5 years of scripting experience
- One of AWS Certified SysOps Administrator Associate, AWS DevOps Engineer Professional, or Certified Kubernetes Administrator certification
- Bachelor's degree in Engineering, Systems Engineering, Computer Science, Mathematics, or a related field
- Experience with Hadoop and Accumulo cluster administration
- Experience with Kubernetes, Docker, Helm, StatefulSets, and persistent volumes
- Experience with configuration management and infrastructure-as-code tools
- Experience with monitoring, observability, alerting, incident response, and SLA/SLO support
- Knowledge of networking, storage, high availability, disaster recovery, and security compliance
What We Do
Wyetech offers quality engineering services in the fields of Software Engineering, Systems Engineering, Cloud Engineering, Data Analysis, and Cyber Security to federal and commercial customers. Wyetech has qualified employees in a broad spectrum of engineering disciplines; however, it is our quality that sets us apart from the rest. Candidates are internal referrals and are thoroughly scrutinized. The result of this has been tremendous, our customers recognize the level of quality and professionalism that Wyetech staff offer. Wyetech is always seeking new quality engineering talent. Please refer to our website for inquries and applications.






