Site Reliability Engineer

San Francisco FULL TIME $120,000 - $150,000 / Year
($10,000 - $12,500 / Month)

Job Description

Join our innovative team as a Site Reliability Engineer, where you will use your expertise to enhance our cloud infrastructure reliability. Your tasks will involve implementing monitoring solutions, optimizing system performance, and enhancing our continuous delivery processes.

Responsibilities

  • Design scalable systems and policies for operational maintenance.
  • Enhance and maintain applications using cloud technologies.
  • Analyze system performance metrics and implement improvements.
  • Lead post-mortem analysis to understand outages and propose remediation.
  • Provide on-call support and resolve issues in production environments.
  • Work across teams to establish a culture of reliability in system designs.

Requirements

Education
  • Bachelor's degree in Computer Science or related field
  • Master's degree is preferred
Experience
  • 5+ years of experience in technology or operations roles
Technical Skills
  • Infrastructure Automation
  • Scripting Languages
Soft Skills
  • Problem-solving
  • Collaboration
Certifications
  • Google Professional Cloud Architect
  • Certified Kubernetes Administrator (CKA)
Languages
  • English: Fluent

Advantageous

  • Knowledge of microservices architecture: Experience with microservices to improve scalability and resilience.
  • Strong automation background: Ability to automate operational tasks and improve workflow efficiency.

Benefits

  • Full health, dental, and vision coverage for employees and families
  • 401(k) plan with employer contribution
  • Flexible working arrangements including remote work
  • Ongoing training and development opportunities

Company Culture

  • Diversity: We pride ourselves on our diverse team and inclusive policies.
  • Empowerment: Employees are encouraged to take initiative and ownership.
  • Work-Life Balance: We support a healthy work-life balance with flexible work arrangements.
Status: Closed