DevOps / Platform Engineer and Server Administrator
DevOps / Platform Engineer & Server Administrator
Employment Type: Full-Time, Remote
Working Hours: U.S. Time Zone
Salary: $2,000โ$2,500 USD/month
About the Role
We are looking for an experienced DevOps / Platform Engineer & Server Administrator to manage and improve our cloud infrastructure, Linux servers, deployment platforms, CI/CD pipelines, container environments, networking, security, and production operations.
This is a hands-on infrastructure role. You will be responsible for keeping development, staging, and production environments stable, secure, scalable, observable, and easy for engineering teams to use.
The ideal candidate combines traditional Linux/System Administration skills with modern DevOps and Platform Engineering practices, including AWS, Docker, Kubernetes, Infrastructure as Code, CI/CD, monitoring, and automation.
Key Responsibilities
- Administer and maintain Linux-based production, staging, and development servers.
- Design, deploy, and maintain infrastructure on AWS.
- Build and maintain internal platforms that simplify application deployment and infrastructure management for developers.
- Manage containerized workloads using Docker and Docker Compose.
- Deploy and operate workloads using Kubernetes/EKS where appropriate.
- Design and maintain automated CI/CD pipelines.
- Automate infrastructure provisioning and configuration using Terraform and Ansible.
- Manage application deployments, releases, rollbacks, and environment configurations.
- Configure and maintain Nginx, Apache, reverse proxies, and load balancers.
- Manage DNS, domains, SSL/TLS certificates, firewalls, VPNs, and networking.
- Configure and maintain AWS VPCs, subnets, routing, security groups, IAM, and load balancers.
- Monitor server health, CPU, memory, disk utilization, network performance, and application availability.
- Implement centralized logging, monitoring, dashboards, and alerting.
- Manage backups, snapshots, retention policies, and disaster recovery procedures.
- Troubleshoot server, networking, container, database, deployment, and application infrastructure issues.
- Perform operating system updates, security patches, and infrastructure upgrades.
- Manage SSH access, IAM permissions, secrets, credentials, and production access.
- Improve infrastructure reliability, availability, scalability, and performance.
- Automate repetitive operational and administrative tasks.
- Support developers with local, development, staging, and production environments.
- Participate in incident response, root-cause analysis, and infrastructure improvements.
- Maintain infrastructure architecture and operational documentation.
Core Technical Skills
Linux & Server Administration
Strong hands-on experience with:
- Linux Administration
- Ubuntu / Debian / RHEL-based systems
- Bash / Shell scripting
- SSH
- Systemd
- Cron
- Package management
- File systems and storage
- User and permission management
- Process management
- Server performance troubleshooting
- OS patching and upgrades
- Disk and memory management
- Backup and recovery
CI/CD
Experience designing and maintaining pipelines with:
- CircleCI
- GitHub Actions
- GitLab CI
- Jenkins
Responsibilities include:
- Automated builds
- Automated testing
- Docker image builds
- Artifact management
- Security scanning
- Deployment automation
- Environment promotion
- Rollbacks
- Production releases
Networking
Strong understanding of:
- TCP/IP
- DNS
- HTTP / HTTPS
- SSL/TLS
- VPC networking
- Subnets
- Routing
- NAT
- Firewalls
- Security Groups
- VPNs
- Reverse proxies
- Load balancing
- CDN concepts
- Network troubleshooting
Monitoring & Observability
Experience with tools such as:
- Prometheus
- Grafana
- AWS CloudWatch
- Datadog
- ELK / OpenSearch
- Loki
- Sentry
The engineer will be responsible for implementing effective metrics, logs, dashboards, health checks, uptime monitoring, and alerting.
Security & Reliability
- IAM and least-privilege access
- Server hardening
- SSH security
- Secrets management
- Vulnerability scanning
- OS and dependency patching
- Container security
- Firewall management
- SSL/TLS management
- Audit logging
- Backup strategies
- Disaster recovery
- High availability
- Incident response
Platform Engineering Responsibilities
Beyond traditional DevOps, we are looking for someone who can improve the developer experience.
This includes:
- Standardizing deployment processes.
- Creating reusable infrastructure modules.
- Building self-service deployment capabilities.
- Standardizing Docker and Kubernetes environments.
- Improving CI/CD speed and reliability.
- Automating environment provisioning.
- Managing development, staging, and production environments consistently.
- Reducing manual deployment and operational work.
- Improving infrastructure observability.
- Establishing infrastructure standards and best practices.
Qualifications
- 3+ years of professional experience in DevOps, Platform Engineering, Cloud Engineering, SRE, Systems Engineering, or Linux Administration.
- Strong Linux/server administration experience.
- Strong hands-on AWS experience.
- Production experience with Docker.
- Experience with CI/CD pipelines.
- Experience with Terraform or similar Infrastructure-as-Code tools.
- Strong networking and troubleshooting knowledge.
- Experience with monitoring and observability platforms.
- Strong scripting and automation skills.
- Understanding of infrastructure and cloud security.
- Experience supporting production systems and responding to incidents.
- Ability to troubleshoot complex problems across infrastructure, applications, networking, and databases.
- Good written and verbal English communication skills.
- Ability to work independently in a remote environment.
- Availability during U.S. business hours.
Preferred Skills
- Kubernetes / EKS
- Helm
- Terraform
- Ansible
- Python
- CircleCI
- GitHub Actions
- Prometheus / Grafana
- ELK / OpenSearch
- Cloudflare
- Nginx
- Redis
- PostgreSQL / MySQL
- RabbitMQ / Kafka / AWS SQS
- Infrastructure security
- FinOps / AWS cost optimization
- High-availability architecture
- Disaster recovery planning
What We're Looking For
We are looking for someone who can operate across three areas:
Server Administration: Keep Linux servers, networking, storage, backups, access, and production environments healthy and secure.
DevOps: Automate builds, testing, deployments, infrastructure provisioning, monitoring, and operational processes.
Platform Engineering: Build standardized infrastructure and developer tooling that allows engineering teams to deploy and operate applications quickly and reliably.
The ideal candidate is comfortable taking ownership of infrastructure problems from initial investigation through resolution and long-term automation.