Site Reliability Engineer

in Information Technology
  • Columbus, Ohio View on Map
  • Salary: $125,000.00 - $150,000.00
Permanent

Job Detail

  • Experience Level Sr Level
  • Degree Type Bachelor of Science (BS)
  • Employment Full Time
  • Working Type Remote
  • Job Reference 0000021111
  • Salary Type Annually
  • Industry Financial Services
  • Selling Points

    Contribute to the reliability of a high-transaction payment platform. Collaborate remotely with a dynamic team in a growth-oriented environment. Leverage cutting-edge AWS technologies and SRE principles.

Job Description

Site Reliability Engineer Overview

  • The Site Reliability Engineer ensures the reliability, scalability, and performance of the client’s payment processing platform.
  • Apply engineering principles to operational challenges, automating tasks and enhancing system resilience.
  • Collaborate with cross-functional teams to embed reliability into application design and deployment practices.
  • Operate and optimize AWS infrastructure, ensuring secure and scalable environments.
  • Monitor and improve database operations, addressing performance issues and scaling for growth.
  • Develop observability stacks to monitor system health and drive reliability improvements.
  • Participate in incident response, maintaining runbooks and escalation paths for swift resolutions.
  • Work remotely with opportunities for collaboration and professional growth.

Site Reliability Engineer Key Responsibilities & Duties

  • Read, debug, and contribute to production C#/.NET code for reliability enhancements.
  • Optimize AWS workloads using EC2 Auto Scaling, ECS Fargate, and Lambda services.
  • Automate deployment pipelines and disaster recovery procedures to reduce operational toil.
  • Manage RDS SQL Server deployments, ensuring failover readiness and performance optimization.
  • Build observability stacks using CloudWatch metrics, AWS X-Ray, and structured logging.
  • Design secure AWS network topologies, configuring VPCs, subnets, and security groups.
  • Conduct root cause analysis for incidents and lead post-mortems to capture lessons learned.
  • Collaborate with engineering teams to share best practices and mentor on SRE principles.

Site Reliability Engineer Job Requirements

  • Bachelor’s degree in Computer Science, Engineering, or equivalent experience.
  • 4+ years in SRE, DevOps, or systems-focused engineering roles.
  • Proficiency in C#/.NET, AWS compute services, and RDS SQL Server operations.
  • Experience with infrastructure-as-code tools like CloudFormation or CDK.
  • Strong scripting skills in PowerShell, Python, or Bash for automation.
  • Knowledge of AWS networking, security configurations, and observability tools.
  • Preferred experience in fintech, payments, or high-transaction environments.
  • Openness to leveraging AI tools for diagnostics and workflow improvement.
  • ShareAustin:

Related Jobs