Contribute to the reliability of a high-transaction payment platform. Collaborate remotely with a dynamic team in a growth-oriented environment. Leverage cutting-edge AWS technologies and SRE principles.
Site Reliability Engineer
in Information Technology PermanentJob Detail
Job Description
Site Reliability Engineer Overview
- The Site Reliability Engineer ensures the reliability, scalability, and performance of the client’s payment processing platform.
- Apply engineering principles to operational challenges, automating tasks and enhancing system resilience.
- Collaborate with cross-functional teams to embed reliability into application design and deployment practices.
- Operate and optimize AWS infrastructure, ensuring secure and scalable environments.
- Monitor and improve database operations, addressing performance issues and scaling for growth.
- Develop observability stacks to monitor system health and drive reliability improvements.
- Participate in incident response, maintaining runbooks and escalation paths for swift resolutions.
- Work remotely with opportunities for collaboration and professional growth.
Site Reliability Engineer Key Responsibilities & Duties
- Read, debug, and contribute to production C#/.NET code for reliability enhancements.
- Optimize AWS workloads using EC2 Auto Scaling, ECS Fargate, and Lambda services.
- Automate deployment pipelines and disaster recovery procedures to reduce operational toil.
- Manage RDS SQL Server deployments, ensuring failover readiness and performance optimization.
- Build observability stacks using CloudWatch metrics, AWS X-Ray, and structured logging.
- Design secure AWS network topologies, configuring VPCs, subnets, and security groups.
- Conduct root cause analysis for incidents and lead post-mortems to capture lessons learned.
- Collaborate with engineering teams to share best practices and mentor on SRE principles.
Site Reliability Engineer Job Requirements
- Bachelor’s degree in Computer Science, Engineering, or equivalent experience.
- 4+ years in SRE, DevOps, or systems-focused engineering roles.
- Proficiency in C#/.NET, AWS compute services, and RDS SQL Server operations.
- Experience with infrastructure-as-code tools like CloudFormation or CDK.
- Strong scripting skills in PowerShell, Python, or Bash for automation.
- Knowledge of AWS networking, security configurations, and observability tools.
- Preferred experience in fintech, payments, or high-transaction environments.
- Openness to leveraging AI tools for diagnostics and workflow improvement.
- ShareAustin:
