Role Overview
We're looking for a DevOps Engineer to own the infrastructure that keeps WatchMen running reliably — from AWS cloud services down to a distributed fleet of edge devices sitting on factory floors. This is a hands-on role across cloud infra, CI/CD, monitoring, and self-hosted services, not a ticket-queue ops job.
Responsibilities
- Managing AWS infrastructure via Terraform, ECS on EC2, ALB, networking, and dedicated-tenant VPC architecture
- Operating our Docker Swarm–based backend deployment and improving its reliability and scaling
- Maintaining our data infrastructure: PostgreSQL, TimescaleDB, PgBouncer, and Redpanda/Kafka event pipelines
- Building and maintaining CI/CD pipelines on self-hosted GitLab CE
- Owning observability: Grafana Cloud dashboards, OpenTelemetry tracing, and alerting for both cloud services and production incidents
- Managing Grafana Alloy fleet monitoring across our Raspberry Pi/edge device fleet deployed at customer sites
- Maintaining self-hosted internal services (MQTT broker/EMQX, Vaultwarden, n8n) behind Caddy reverse proxies
- Diagnosing infrastructure-level production issues, network, database, message broker, and container orchestration problems
- Hardening system security, including tools like Wazuh
Requirements
- 3–6 years of professional DevOps/infrastructure experience
- Strong hands-on experience with Docker and container orchestration (Docker Swarm, ECS, or Kubernetes)
- Experience with Terraform or another infrastructure-as-code tool
- Solid Linux systems administration and networking fundamentals
- Experience with CI/CD pipelines (GitLab CI, GitHub Actions, or similar)
- Comfortable debugging production incidents independently
Skills
- AWS
- Terraform
- Docker
- GitLab
- Linux