Full-Time

Site Reliability Engineer

Axle

Axle

201-500 employees

Translational research informatics and data science

Compensation Overview

$140k - $155k/yr

Frederick, MD, USA

In Person

On-site in Frederick, Maryland; no remote option stated.

Category
DevOps & Infrastructure (1)
Required Skills
PowerShell
Chef
Bash
Kubernetes
Microsoft Azure
Python
Grafana
Puppet
GitHub Actions
NoSQL
Node.js
SQL
Java
Docker
.NET
AWS
Prometheus
Jenkins
Terraform
Ansible
Splunk
Google Cloud Platform

Get referred to Axle

See people who can refer or advise you

Requirements
  • Must have total of 6+ experience DevOps / SRE roles with monitoring and observability tools (Prometheus, Grafana, ELK, or cloud-native equivalents) for on-prem and cloud hosted workloads.
  • Must have 4+ years of Hands-on Linux experience that includes Ubuntu/CentOS/Red Hat operating systems, containers, dependency management and administration support
  • Must have 4+ years of experience automating Infrastructure-as-Code (IaC) deployments to one of the following cloud platforms Amazon AWS, Google GCP and Microsoft Azure
  • Must have 4+ years of experience with CI/CD and automation tools such as Terraform, Ansible, Chef, Puppet, Jenkins, GitHub Actions
  • Strong scripting skills (Python, Bash, PowerShell or similar)
  • Must be proficient using vibe coding and coding assistants to develop scripts, tools and applications for the DevOps and SRE use cases
  • Must have proficiency to debug or troubleshoot and/or deploying SQL and/or NoSQL databases, object storage, web servers, open-source programming stack for Node.JS, R, Python, .NET Core, Java is desired but not mandatory
  • Must be willing to learn new technologies, adopt and adapt to emerging technologies or needs from a project to a project
  • Cloud certifications is preferred
  • Certifications in Grafana, Splunk, Docker, Kubernetes is preferred but optional
Responsibilities
  • Design and implement enterprise-grade monitoring and observability frameworks (metrics, logs, traces) across distributed systems using enterprise Splunk, Grafana and Open-telemetry tools
  • Establish and manage SLIs, SLOs, and error budgets to drive reliability improvements
  • Develop and maintain real-time asset inventory systems across cloud, on-prem, and hybrid environments
  • Automate workload onboarding and offboarding processes, ensuring standardization and governance
  • Track system ownership, dependencies, and lifecycle states for operational transparency
  • Build proactive detection mechanisms using AIOps and intelligent alerting to minimize incident impact
  • Design and operate scalable, resilient, and secure infrastructure platforms across cloud and hybrid environments
  • Implement automated compliance tracking and enforcement aligned with organizational and regulatory standards (e.g., NIST, FISMA, FedRAMP)
  • Embed ITIL processes (incident, change, problem, configuration management) into SRE workflows
  • Build and maintain automated deployment environments and pipelines that enforce security, compliance, and operational standards
  • Develop “golden paths” and standardized platform templates for consistent workload deployment
  • Automate provisioning, patching, configuration management, and environment lifecycle
  • Leverage AI/ML coding assistants and vibe coding practices to rapidly develop automation scripts, tools, and internal platforms
  • Integrate AI-driven tooling into DevOps pipelines for code quality, security scanning, and operational insights
  • Lead adoption of AI-enhanced SRE practices, including intelligent remediation and predictive operations
  • Champion DevOps and SRE practices including Infrastructure as Code, CI/CD, observability, and reliability engineering
  • Build developer-friendly platforms (“golden paths”) that simplify deployments, reduce friction, and improve velocity
  • Enable and optimize infrastructure for AI/ML workloads, including data pipelines, storage systems, and inference environments, GPU-enabled and high-performance compute workloads
  • Build and manage containerized and orchestrated platforms (Docker, Kubernetes)
  • Support cloud migration, modernization, and platform standardization initiatives
  • Ensure systems meet security, compliance, backup, and disaster recovery requirements
  • Evangelize and promote best practices in DevOps, SRE, and platform engineering to developer communities
  • Stay abreast of new technologies in your areas but not limited to AIOps, MLOps, cloud computing & deployment, site reliability engineering, infrastructure automation, security best practices, data engineering etc.
Desired Qualifications
  • Cloud certifications is preferred
  • Certifications in Grafana, Splunk, Docker, Kubernetes is preferred but optional

Axle Informatics provides specialized informatics solutions for translational research, health informatics, and data science to biomedical research centers and healthcare organizations. Its offerings are customized software and data management platforms that help researchers collect, integrate, analyze, and visualize large research datasets, with end-to-end tools that automate data aggregation and deliver analytics, dashboards, and decision-support features. The company differentiates itself with an integrated, scientifically informed approach that bridges data science with application development, specifically focused on translational research and clinical data work rather than generic software. Axle’s goal is to advance public health by moving biomedical discoveries from the lab to bedside, improving healthcare outcomes.

Company Size

201-500

Company Stage

N/A

Total Funding

N/A

Headquarters

Rockville, Maryland

Founded

2002

Get referred to Axle

See people who can refer or advise you

Simplify Jobs

Simplify's Take

What believers are saying

  • NIH awarded Axle the $11.2 million N3C Dynamic Workspaces task order on September 30, 2025.
  • NCATS extended Axle through June 27, 2026, signaling trust during transition risk.
  • SK pharmteco partnered with Axle and NIH on rare-disease gene therapy programs on May 7, 2026.

What critics are saying

  • Axle depends heavily on NIH funding; one recompete failure hits revenues by 2027.
  • NCATS limited-source notices in 2025-2026 show mission-critical work gets rebid fast.
  • A larger integrator like Dovel or Leidos can underprice Axle and compress margins.

What makes Axle unique

  • Axle embeds inside NIH, winning NCATS and NIAID support work in 2025-2026.
  • Its translational-research niche blends biomedical science, software engineering, and federal program management.
  • Repeated sole-source extensions show NIH treats Axle as operational infrastructure, not a commodity vendor.

Help us improve and share your feedback! Did you find this helpful?

Benefits

Health Insurance

Dental Insurance

Vision Insurance

Paid Vacation

Paid Holidays

401(k) Company Match

Educational Benefits for Career Growth

Employee Referral Bonus

Flexible Spending Accounts

Company News

TheGWW.com
Jul 25th, 2023
“Octo awarded $64.7M IT Infrastructure Call Order to support NCI’s Cancer research”

Octo, in partnership with Unissant, Axle Informatics, and TRex, has been awarded an IT Infrastructure and Operations Call Order to support the NCI’s OCIO.

GlobeNewswire
Jan 12th, 2023
Digital Pathology Market Worth $1.86 Billion by 2030 -

For instance, in May 2020, Indica Labs (U.S.), a provider of digital pathology solutions, collaborated with information technology consulting companies, Octo (U.S.) and Axle Informatics (U.S.) and the National Institutes of Health (NIH) (U.S.), to develop an online collection of high-resolution histopathology images of tissues from COVID-19 patients using Indica’s HALO Link platform.