About the position:
We re looking for a DevOps Engineer to help us build our new Banking as a Service (BaaS) platform and maintain our day to day operations. As a DevOps you ll be improving designing building and managing a scalable infrastructure that meets our security and compliance requirements. This is a seniorlevel position based in Mexico. Work remotely
Our stack runs in AWS. We use Lambda ECS SQS SNS RDS QLDB RDS MongoDB MySQL CloudFront Elasticache EC2 instances and Application Load Balancers. Most of the code from the engineering team is handled on Ruby. We are continuously improving our monitoring deployment processes and removing toil by implementing IaC (Infrastructure as code) with Terraform.
Job Description:
Design and implement CI/CD pipelines that deliver applications and services at high velocity
Collaborate with a worldclass engineering team to propose Cloud services.
Support engineering and data teams in enhancing the scalability reliability and security of our platform
Control cost & performance optimization on cloud resources management
Identify opportunities for automation and efficiency in the endtoend product delivery cycle
Design and implement and manage performance monitoring for our systems and applications
Consult with product and engineering on the operational requirements of software solutions
Keep uptodate documentation of system design implementation and troubleshooting
Work as an embedded member of an Agile team to provide endtoend responsibility for applications and microservices.
Build automation to provision infrastructure as code to cloud platforms.
Containers management
Job Profile:
4 years of relevant work experience as a DevOps engineer or SysAdmin
4 years of experience with AWS products
Experience with AWS Networking TGW Site to Site VPN
Experience with optimizing database performance AWS RDS
Experience with infrastructure as code We use terraform
Understanding of event-driven architectures Distributed systems - How clusters are formed, Quorum management, Failure handling. 3 to 5 years of hands-on Experience in MQ or NATS broker or similar messaging solutions. Understanding of Kafka clustering would be good to have. Knows Client-Server communication aspects - sockets, TLS protocol etc Understands the concept of region and AZs. Provide L2 support production systems like application, database, middleware components, infrastructure and network components. Manage production incidents end-to-end within defined SLAs with focus on resolution rather than who caused it. Interact with various stakeholders such as Release managers, program leads, service managers, development and test leads Review operational readiness requirements such as monitoring and alerting, log rotation and resilience of the components and report the gaps Provide pre-implementation support with activities such as release notes review and implementation dry runs. Protect production components by running health checks monitoring latency and memory utilization. Automate day-to-day activities and propose changes that improve reliability Participate in CAB and provide feedback on change requests Support the DevOps team in testing the promoted pipelines and suggest automation of configuration items. Practice incident management best practices and perform RCA. Participate in disaster recovery tests and operational acceptance tests Analyze the technology stack that makes up the product and optimize recovery time objective. Work with team members spread across and time zones Share knowledge, document improvements and mentor junior resources It is good to have skills using Jenkins to orchestrate builds and link to Sonar, Maven, etc. to build out the CI/CD pipeline. Support deployments of code into multiple lower environments. Supporting current processes needed with an emphasis on automating everything as soon as possible. It is good to have skill to design, Implement, and enhance our deployment automation based on Chef. We need proven experience designing and implementing an overall release and deployment process. It is good to have skill to design and implement a Git based code management strategy that will support multiple environment deployments in parallel. Experience with automation for Branch management, code promotions, and version management. Engage in and improve the whole lifecycle of services from inception and design through deployment, operation, and refinement. Requirements MQ/EB Understanding of event-driven architectures Distributed systems - How clusters are formed, Quorum management, Failure handling. 3 to 5 years of hands-on Experience in MQ or NATS broker or similar messaging solutions. An understanding of Kafka clustering would be good to have. Knows Client-Server communication aspects - sockets, TLS protocol etc Understand the concept of region and AZs. Deployments MTF/Prod, Maintenance items (including stop/start, Disaster Recovery-related activities, etc.), CR for changes in MTF/Prod Good knowledge on Nginx Tools - Log Monitoring Tool - Splunk Application Monitoring tool - Dynatrace Ticketing incident/problem management tool - Remedy Dev-ops Basics - CI-CD Basics, Overview of Git, Bit-bucket, SonarQube, Ansible/Chef Skills - Linux & Shell Scripting ITIL / ITSM PL/SQL Troubleshooting Jenkins - CI/CD Groovy Scripting/Yaml Ansible/Chef Nginx Java / JEE Event-Driven Architectures MQ or NATS broker or similar messaging solutions. Kafka Client-server communication aspects - sockets, TLS protocol Understand the concept of region and AZs.