Site Reliability Engineer

Employer New York Technology Partners
Location New York, New York
Type Full-time

Job Description

Candidates must be comfortable working onsite 5x a week

We are looking for a high-energy, enthusiastic Platform Engineer with a desire to refine their craft. The ideal candidate will be eager to contribute, demonstrate exceptional technical aptitude, and bring a positive attitude to our collaborative environment in the heart of Chicago.

In this role, you can expect to:

• Drive Incident Management: Lead and cultivate a culture of incident management within the team.

• Participate in On-Call Duties: Contribute to incident command and on-call rotations.

• Build Technical Skills: Develop your technical expertise while working closely with a team of skilled engineers.

• Collaborate Cross-Functionally: Engage in learning, teaching, and collaborating with different teams across the organization.

You may be a good fit for our team if you have:

• Professional AWS Experience: Proven experience in implementing and maintaining scalable and reliable infrastructure on Amazon Web Services (AWS).

• Diverse Technical Interests: Enjoy working on a range of scopes including software engineering, cloud infrastructure, DevSecOps, and SRE.

• Efficiency Improvement Skills: A track record of driving efficiency improvements in software at scale.

• Cross-Functional Collaboration: Experience working cross-functionally to promote and implement engineering culture changes.

• SaaS or Managed Software Experience: Hands-on experience with Software as a Service (SaaS) or other managed software offerings.

• Public Cloud Expertise: Expertise in one or more major public cloud platforms.

Qualifications:

• Educational Background: Degree in Computer Science, Engineering, or a related field is preferred.

• Linux Systems and CI/CD: Deep knowledge of Linux systems and CI/CD tools such as Jenkins.

• Development and Automation: Experience in developing applications, automation tools, and the necessary infrastructure for large environments. Experience developing and maintaining infrastructure as code (IaC) templates using Terraform.

• Security Proficiency: Well-versed in vulnerability management and security products.

• Troubleshooting Skills: Expertise in troubleshooting using monitoring tools.

• IT Operations Knowledge: Understanding of IT Operations best practices in always-up environments.

• Strong written and verbal communication abilities.

• Containerization Experience: Familiarity with containerization technologies such as Docker and Kubernetes.

Our Ideal Candidate:

• Develops and Automates: Has created applications and automation tools to build, deploy, monitor, and integrate data sources for our systems and applications.

• Collaborates and Engages: Enjoys working closely with multiple teams, fostering a collaborative environment.