Full-Time

Senior / Staff Engineer

Platform

Sei Network

Sei Network

No salary listed

San Francisco, CA, USA + 1 more

More locations: New York, NY, USA

Hybrid

Hybrid work is indicated for New York City, San Francisco, and New York.

Category
DevOps & Infrastructure (1)
Required Skills
Kubernetes
Incident Response
Software Testing
Threat modeling
OpenTelemetry
Docker
Blockchain
Prometheus
Observability
DevOps
Requirements
  • At least 5 years of experience building production systems at scale.
  • Experience designing and shipping infrastructure that other engineers rely on, with end-to-end understanding beyond the surface API.
  • Knowledge of security fundamentals, including threat modeling, common vulnerability classes, and integrating security work into a development pipeline.
  • Ability to reason about distributed consensus protocols at the algorithmic level, including safety, liveness, fault assumptions, and practical failure modes.
  • Experience debugging a non-trivial distributed consensus issue or equivalent deep study of the layer.
  • Ability to design benchmarks that isolate variables, profile systems under realistic load, identify performance bottlenecks, and explain their causes before proposing fixes.
  • Experience operating systems where outages had significant consequences, including triage under pressure, clear communication, and converting postmortems into durable structural fixes.
  • End-to-end systems thinking, including forming and defending opinions about monitoring stacks, CI versus CD versus runtime responsibilities, and when to use chaos testing.
Responsibilities
  • Build and maintain Docker and Kubernetes harnesses for repeatable stress and scale testing as a pre-release gate for core teams.
  • Run the multi-cloud footprint with explicit reliability and cost targets.
  • Consolidate CI/CD into reusable, standardized workflows.
  • Own version tagging, hotfix paths, pull request review scaffolding, and the full release lifecycle.
  • Provide a single pane of glass for metrics, traces, and logs correlated by block height and powered by OpenTelemetry.
  • Operate highly available Prometheus and the ELK stack, and define service-level objectives for every critical service.
  • Automate chaos and performance tests in CI to gate all releases and catch regressions before merge.
  • Run the bug and vulnerability workflow end to end, including auditors, Immunefi, and triage.
  • Conduct postmortems and ensure every incident produces a new test, metric, or runbook.
Desired Qualifications
  • Blockchain or Web3 infrastructure experience, especially with consensus, mempool, or storage layers.
  • Experience with multi-cloud cost programs that preserve reliability.
  • Experience integrating AI-assisted security tooling into a development pipeline.

Company Size

N/A

Company Stage

N/A

Total Funding

N/A

Headquarters

N/A

Founded

N/A