Actively recruiting / 9 applicants
We’re here to help you
Juliana Torrisi is in direct contact with the company and can answer any questions you may have. Email
Juliana Torrisi, RecruiterRole Overview
We are seeking a hands-on DevOps / AWS Infrastructure Engineer to play a critical role in stabilizing, strengthening, and scaling the infrastructure of a burgeoning SaaS platform in the wine industry. As customer demand increases, your expertise will be crucial in enhancing the AWS environment's reliability and operational practices, setting the stage for continued growth and multi-tenancy.
Responsibilities
- Assess the current AWS infrastructure, documenting and identifying key reliability and operational risks.
- Enhance availability and redundancy across production services to ensure seamless service delivery.
- Establish and validate robust backup, recovery, and disaster recovery procedures.
- Diagnose and resolve infrastructure-related performance issues and bottlenecks in AWS services, databases, APIs, and web application infrastructure.
- Improve monitoring, alerting, logging, and incident-response practices to ensure quick resolution of issues.
- Implement practical infrastructure safeguards and operational best practices to maintain system integrity.
- Design and implement infrastructure improvements that support increased customer volume and multi-tenant SaaS growth.
- Support AWS services, including EC2, Aurora/RDS, load balancing, networking, and storage, ensuring smooth operations.
- Collaborate with developers to address infrastructure issues intersecting with the application or database layer.
- Provide occasional ad-hoc support for urgent production issues to maintain service continuity.
- Work independently on implementation tasks, receiving high-level architectural guidance from a senior DevOps advisor.
Required Skills
- Strong hands-on experience managing production AWS environments, specifically EC2, Aurora/RDS, load balancing, networking, backups, and monitoring.
- Proven experience in improving reliability, high availability, disaster recovery, and scalability of production web or SaaS applications.
- Excellent troubleshooting skills, with a background in diagnosing infrastructure performance issues, outages, and service bottlenecks.
- Working knowledge of web application infrastructure, including HTTP APIs, DNS, SSL/TLS, load balancing, and relational databases, with basic SQL fluency.
Nice to Have
- Experience with Infrastructure as Code tools such as Terraform or CloudFormation.
- Familiarity with CloudWatch, observability platforms, logging, and alerting.
- Experience with Docker, ECS, or other containerized AWS environments.
- Background in supporting or designing infrastructure for multi-tenant SaaS applications.
- Familiarity with Aurora PostgreSQL or MySQL administration and performance monitoring.
- Experience implementing CI/CD and deployment safeguards.