Experienced DevOps Engineer for Infrastructure Optimization

Job ID: 38929115

Budget: $750 – $1,500 USD

Introduction:
We are looking for an experienced DevOps Engineer to join our team and play a critical role in enhancing the reliability, scalability, and security of our infrastructure. As a DevOps Engineer, you will work on streamlining our server and pod cluster management, ensuring seamless updates, and contributing to system design for auto-scaling, automation, and enhanced performance. You will also be responsible for ensuring high availability of our systems and integrating with AWS services, Kubernetes, and other cloud technologies.

Key Responsibilities:

* Server and Pod Cluster Management:
* Change and manage server and pod cluster passwords, ensuring minimal disruption.
* Restart servers and clusters as necessary, ensuring proper security and connectivity.
* Ensure that after a password change, the updated IP addresses are reflected in AWS Route 53 to avoid service disruptions for streaming services.

* Kubernetes and Stream Server Management:
* Update Kubernetes to the latest stable version while ensuring no interruptions to the stream servers.
* Collaborate with the team to plan and execute rolling updates to Kubernetes clusters.

* SSL Management & Automation:
* Update and renew SSL certificates using Let's Encrypt, implementing automation to ensure certificates are always up-to-date without manual intervention.
* Ensure security best practices are followed for SSL certificate management.

* Scaling and Systems Architecture Design:
* Collaborate with engineering teams to design and implement horizontal and vertical scaling solutions for the stream servers and other components.
* Develop systems to automatically scale pods when capacity is reached, ensuring minimal service disruption.
* Design cloud infrastructure to handle increasing loads and optimize resource allocation.

* Route 53 Management:
* Ensure that AWS Route 53 is properly routing traffic based on geo-location and provides high availability for our services.
* Optimize Route 53 configurations to reduce latency and enhance performance.

* GitHub Deployment Configuration:
* Switch the deployment pipeline to a different account in GitHub.
* Ensure proper configuration of GitHub Actions or other CI/CD tools for seamless deployment to production.

Skills & Qualifications:
* Proven experience with AWS services, including EC2, Route 53, and other cloud infrastructure tools.
* Solid understanding of Kubernetes, including cluster management, updates, and scaling.
* Expertise in SSL/TLS management, particularly with Let’s Encrypt for automated certificate renewal.
* Hands-on experience with CI/CD pipelines and tools, preferably GitHub Actions, Jenkins, or similar.
* Familiarity with Docker and container orchestration technologies.
* Strong knowledge of Linux/Unix systems and scripting (Bash, Python, etc.).
* Experience with Infrastructure as Code (IaC) tools such as Terraform or CloudFormation.
* Strong troubleshooting skills and ability to identify and resolve issues quickly with minimal downtime.
* Excellent communication and collaboration skills, with a proactive approach to problem-solving.

Desired Skills:
* Familiarity with Geo-location-based routing and advanced DNS management.
* Experience with auto-scaling architectures and resource optimization techniques.
* Knowledge of monitoring and alerting tools like Prometheus, Grafana, CloudWatch, or similar.
* Experience working in a fast-paced, cloud-native development environment.