Mid-Level DevOps Engineer โ AWS, Security, and Reliability
We are looking for a mid-level DevOps engineer with hands-on experience building, securing, and operating production applications on AWS.
You should be able to independently manage infrastructure, build deployment pipelines, investigate production problems, and automate operational work. Strong experience with Terraform, Docker, and GitHub Actions is highly preferred.
Our work includes reusable cloud environments for healthcare applications, automated security assessments, and monitoring products across customer deployments. You will contribute to both our infrastructure and new product development.
We are looking for creative people who enjoy developing novel concepts, challenging assumptions, and turning ideas into working solutions. We strongly encourage thinking outside the box. You should be comfortable proposing an approach, building a prototype, testing it, and improving it based on results.
Responsibilities
- Design and maintain AWS infrastructure for development, staging, and production.
- Build reusable Terraform modules and reliable CI/CD pipelines with deployment checks and rollback procedures.
- Deploy and operate containerized applications, APIs, databases, and background workers.
- Implement access controls, network isolation, encryption, secrets management, and secure logging.
- Build monitoring, dashboards, and alerts that identify application failures and help investigate their causes.
- Configure backups, test restoration, and verify recovery against agreed objectives.
- Investigate incidents, resolve security findings, and improve system reliability and operating costs.
- Prepare architecture documentation and technical evidence for cloud reviews and customer security assessments.
- Develop reusable tools for deployment, monitoring, security checks, and evidence collection.
- Propose and prototype new approaches that simplify operations or create useful product capabilities.
Required Qualifications
- At least three years of DevOps, cloud infrastructure, or site reliability engineering experience, including two years operating production applications on AWS.
- Strong experience with Terraform, Docker, and GitHub Actions or comparable tools.
- Strong Linux troubleshooting skills and proficiency in Bash or Python.
- Practical understanding of networking, DNS, TLS, authentication, and authorization.
- Experience with cloud access management, secrets, encryption, and separation between environments.
- Ability to troubleshoot application and infrastructure failures using logs, metrics, and deployment history.
- Experience with monitoring, backups, recovery, and safe database migration practices.
- Experience with Git, code review, and controlled infrastructure changes.
- Ability to work independently and communicate clearly in English.
Preferred Qualifications
- Experience assessing cloud architecture, identifying security and reliability gaps, and collecting evidence that controls work.
- Experience implementing safeguards for healthcare applications, including sensitive-data boundaries, access auditing, and data protection.
- Experience building reusable infrastructure for separate customer accounts.
- Experience developing monitoring across multiple deployments, including useful alerts and escalation procedures.
- Familiarity with vulnerability scanning, software supply-chain security, and database operations.
- Experience with AI-assisted development tools, particularly Claude Code.
How We Work
We value initiative, curiosity, and ownership. You will work closely with developers, participate in architecture decisions, and help define what we build.
We encourage experimentation and practical solutions. Bring ideas, explain the tradeoffs, and demonstrate what works. Infrastructure changes should be reproducible, deployments verifiable, and recovery procedures tested.
We use Claude Code to accelerate implementation, troubleshooting, and documentation. You remain responsible for understanding, reviewing, and validating the resulting work.