US Jobs US Jobs     UK Jobs UK Jobs     EU Jobs EU Jobs


DevOps - Site Reliability Engineer

This job is based in Des Moines, Iowa.

Essential Functions 1.

Pipeline Engineering: Design, develop, and maintain robust CI/CD pipelines to automate the build, test, and deployment processes for both web-based IoT applications and firmware.

Collaborate with engineering teams to implement effective branching strategies, code reviews, and automated testing frameworks.

Continuously optimize pipeline performance and reliability to accelerate delivery cycles.

2.

Cloud Infrastructure Engineering: Manage and enhance our AWS infrastructure, including provisioning, configuration, and scaling of resources.

Automate routine tasks and implement infrastructure as code (IaC) practices to improve efficiency and consistency.

Monitor system performance and proactively identify and resolve potential issues.

Implement robust security measures to protect our cloud environment.

3.

Site Reliability Engineering Ensure high availability, performance, and reliability of our applications and infrastructure.

Respond to incidents and outages promptly, implementing effective incident response procedures.

Analyze system logs and metrics to identify potential issues and bottlenecks.

Collaborate with development teams to improve software quality and reduce failure rates.

4.

Strong proficiency in scripting languages (Python, Bash, etc.) and configuration management tools (Ansible, Puppet, Chef, etc.) 5.

Experience with CI/CD tools (Jenkins, GitHub CI/CD, GitHub Workflows and Actions, etc.) and version control systems (Git) 6.

Deep understanding of cloud platforms, particularly AWS 7.

Knowledge of containerization technologies (Docker, Kubernetes) and orchestration tools 8.

Knowledge of containerization security best practices, tools and techniques.

9.

Knowledge of storage administration NFS/EFS.

10.

Knowledge of disaster recovery best practices and tools.

11.

Knowledge of network security best practices and tools.

12.

Experience with infrastructure change management best practices.

13.

Experience with infrastructure as code (IaC) practices (Terraform, CloudFormation) 14.

Solid understanding of networking concepts (TCP/IP, DNS, load balancing) 15.

Familiarity with monitoring and logging tools (Prometheus, Grafana, ELK Stack) 16.

Strong problem-solving and troubleshooting skills 17.

A passion for automation and continuous improvement





Share Job