Skip to main content
Posted June 26, 2018
Rigado

Site Reliability Engineer

Portland Full Time

Rigado is growing quickly in the fast-moving markets of wireless connectivity and edge computing. At Rigado, the Site Reliability Engineer works closely...

Rigado is growing quickly in the fast-moving markets of wireless connectivity and edge computing. At Rigado, the Site Reliability Engineer works closely with the software engineering teams to ensure the software we ship is resilient, highly available, and monitored. We use tools like DynamoDB, Go and MQTT to build a platform to manage thousands of embedded Linux systems at scale.

Primary Job activities include:

• Working closely with software engineering teams to ship stable, highly available software and services

• Capacity planning and estimation

• Managing infrastructure such as GitLab hosted in AWS

• Define and implement monitoring, metrics and alerts

• Work with the Director of Software to ensure we have the right work planned to ensure a stable system

• Help teams define and manage runbooks for on-call rotation

You're right for this role if you:

• Have strong experience with AWS

• Have strong experience with Linux

• Have strong experience with running SaaS platforms

• Have strong experience with distributed systems

• Have experience with highly available systems

• Have experience with incident response including keeping impacting customers updated

• Can communicate effectively in writing

• Have a strong work ethic

• Are self-directed and resourceful; and a self-starter who acts with initiative

Appreciated, but not required:

• Experience with Go

• Experience with Ubuntu Core and Snaps

• Experience with BLE

This listing expired on Aug 10. Applications are no longer accepted.

Below are some other jobs we think you might be interested in.