Full-Time

Senior Site Reliability Engineer

Barbaricum

Barbaricum

51-200 employees

Government contracting firm delivering mission support

No salary listed

Washington, DC, USA

In Person

Category
DevOps & Infrastructure (1)
Required Skills
PowerShell
Chef
Bash
Microsoft Azure
Python
Puppet
Computer Networking
AWS
Ansible
Linux/Unix
Google Cloud Platform

Get referred to Barbaricum

Find people who can refer or advise you

Requirements
  • Expert knowledge of site reliability engineering practices, system monitoring, incident management, automation, performance tuning, and operational resilience.
  • Strong understanding of Windows and Linux administration, infrastructure operations, system configuration, service management, and troubleshooting practices.
  • Experience with automation platforms and configuration management tools such as Ansible, Puppet, Chef, or similar technologies.
  • Proficiency with scripting languages such as Python, Shell, PowerShell, or similar tools used to automate operational and infrastructure tasks.
  • Knowledge of cloud services and infrastructure across AWS, Microsoft Azure, Google Cloud, or comparable cloud environments.
  • Strong understanding of network troubleshooting, configuration, connectivity analysis, system dependencies, and performance bottleneck identification.
  • Ability to design, interpret, and maintain dashboards, alerts, metrics, logs, and operational reporting that support service health and decision-making.
  • Ability to conduct root cause analysis, post-incident reviews, and corrective action planning in complex technical environments.
  • Strong problem-solving skills and the ability to work under pressure during outages, impairments, and time-sensitive operational issues.
  • Excellent written and verbal communication skills, with the ability to explain technical findings, incident impacts, and reliability recommendations to technical and non-technical stakeholders.
  • Bachelor's degree in Computer Science, Information Technology, Systems Engineering, Cybersecurity, or a related field; Master's degree preferred.
  • Certifications related to cloud computing, system administration, site reliability engineering, DevSecOps, or automation are beneficial.
  • 10+ years of experience in site reliability engineering, systems administration, infrastructure operations, cloud operations, DevSecOps, or a similar technical role, particularly in a government, federal, defense, or secure IT setting.
  • Demonstrated experience maintaining reliable, scalable, and efficiently managed IT systems across on-premises, cloud, or hybrid environments.
  • Experience developing automated infrastructure, operational scripts, monitoring solutions, dashboards, runbooks, and configuration standards.
  • Experience supporting incident response, system outage resolution, post-incident reviews, root cause analysis, and operational improvement initiatives.
  • Experience collaborating with development, infrastructure, cloud, cybersecurity, and program teams to improve reliability, security, and service performance.
  • DoD Secret Security Clearance.
Responsibilities
  • Monitor and maintain system reliability, availability, and performance across on-premises, cloud, and hybrid IT environments supporting MC&FP mission requirements.
  • Implement proactive performance monitoring, automated alerting, incident response workflows, and resilience engineering practices to reduce downtime and improve operational visibility.
  • Develop, maintain, and improve scalable automated infrastructure solutions that support reliable system operations and repeatable service delivery.
  • Implement rollback strategies, recovery approaches, and chaos engineering practices to validate resilience, reduce operational risk, and improve system stability.
  • Analyze usage patterns, capacity trends, and performance indicators to support dynamic scaling, resource optimization, and system improvement decisions.
  • Develop and maintain real-time operational dashboards, reports, and metrics that enable rapid decision-making, leadership awareness, and system optimization.
  • Respond to and resolve system outages, impairments, and service disruptions while coordinating with technical teams to minimize mission impact.
  • Conduct post-incident reviews to identify root causes, document lessons learned, and implement preventative measures that reduce recurrence.
  • Collaborate with software developers, cloud engineers, cybersecurity personnel, and operations teams to improve services, reliability patterns, deployment practices, and operational standards.
  • Create and maintain system documentation, configuration standards, operational runbooks, monitoring procedures, and service reliability guidance.
  • Automate common operations tasks to reduce manual workloads, improve consistency, and increase system efficiency.
  • Implement security best practices across operational activities, infrastructure automation, monitoring, incident response, and system administration functions.

Barbaricum is a service-disabled veteran-owned contracting firm that supports U.S. government clients in Integrated Communications, Mission Support, Research and Analysis, Cyber Security/Intelligence, and Technology-Enabled Services. It combines strategy development with deploying emerging technologies through a hands-on, all-inclusive approach and builds long-term partnerships driven by repeat business. The firm differentiates itself with SDVOB status, ISO 9001:2015 certification, CMMI Level 3 appraisal, and a global team spanning five continents, along with industry recognition for growth and client work. Its goal is to transform how the U.S. Government tackles complex problems by delivering effective, technology-enabled solutions that advance national security.

Company Size

51-200

Company Stage

N/A

Total Funding

N/A

Headquarters

Washington DC, District of Columbia

Founded

2008

Get referred to Barbaricum

Find people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • Pete Brady's August 2023 appointment as Chief Growth Officer accelerates global market expansion.
  • Multiple offices in DC, Tampa, Dayton, and Crane enhance DoD regional mission support.
  • DevSecOps and CBM+ services boost US Army, Navy, and USSOCOM readiness and cost savings.

What critics are saying

  • ODL integration fails, clashing airborne intel with cyber services, losing contracts in 6-12 months.
  • Booz Allen Hamilton outcompetes Pete Brady's initiatives, blocking DoD wins in 12-24 months.
  • 2026 recertification disqualifies SDVOSB status, barring $500M+ vehicles due to revenue growth.

What makes Barbaricum unique

  • Barbaricum delivers ISO 9001:2015-certified and CMMI Level 3-appraised national security solutions.
  • SDVOSB status enables superior access to restricted federal contract vehicles for DoD clients.
  • ODL acquisition adds airborne ISR and AI/ML expertise to core cyber/intelligence services.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Flexible Work Hours

Company News

Technical.ly
Aug 5th, 2024
Reallist Engineers 2024: Meet 15 Technologists Helping The Dmv Ecosystem Grow

Who’s been helping make the DMV tech scene a hot commodity? We aren’t talking about the blistering heat and humidity this summer, but experts in the region with technological acumen who simultaneously make a difference where they live. Technical.ly’s 2024 RealLIST Engineers features savvy leaders who not only have technological prowess, but also make a difference in the DMV scene. Maybe they host a meetup group, dedicate time to educating the next generation, or even help beautify local hiking trails.The list, which we’ve published in the DC region since 2019, is put together with help of reader nominations (you can submit names all year round), and our own research and past reporting. Keep reading to meet the 15 impressive people who made the list this year. Leslie WelchLeslie Welch. ExcellaWhen she’s not working as a principal fellow in AI and analytics at the Arlington software company Excella, Welch is very involved in data community meetups throughout the region, including Data Viz DC and Data Science DC. She’s a co-organizer of Power BI DC, a group focused on bringing together users of Microsoft’s business intelligence data visualization software product. She previously served a team lead for the inclusion, diversity, and equity ambassadors at Excella. David CurryDavid Curry. BarbaricumA senior software engineer and solutions architect at the government relations firm Barbaricum, Curry has more than 25 years of experience in the field. Outside of work, he volunteers at the Potomac Appalachian Trail Club, where he helps maintain trails and cabins in the region

ExecutiveBiz
Apr 25th, 2024
Barbaricum Acquires ODL to Boost Capabilities

Barbaricum, a service-disabled veteran-owned government contracting firm, has acquired ODL Services to enhance its defense and technology services. The acquisition adds expertise in intelligence, surveillance, reconnaissance, and mission support, aiming to advance modernization, predictive analytics, AI/ML, and battlespace fusion. This move is expected to benefit Barbaricum's defense and intelligence clients by providing a broader range of advanced technological and operational services.

Washington Technology
Apr 24th, 2024
Barbaricum acquires airborne intelligence provider

Barbaricum is also purchasing ODL Solutions to grow its access with agencies in the defense and intelligence communities, including those in special operations.

Barbaricum LLC
Apr 15th, 2024
Barbaricum Expands Capabilities with Strategic Acquisition

Barbaricum, a leader in innovative defense and technology solutions, announced the strategic acquisition of ODL Services.

Barbaricum LLC
Aug 11th, 2023
Barbaricum Adds Chief Growth Officer

WASHINGTON - August 11, 2023 - Barbaricum proudly welcomes Pete Brady as its new Chief Growth Officer, tasked with spearheading the company's global growth and rapidly expanding market footprint.