Salary not listed
Staff Site Reliability Engineer
DevOpsRemote — USA
Published on 2026-10-06
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.
About this role
Bluesky is seeking a Staff Site Reliability Engineer to design and manage the infrastructure for their federated social network. The role involves ensuring the reliability and performance of production systems while collaborating with engineers across teams.
About the company
Bluesky is focused on developing a decentralized social media platform that prioritizes user safety and trust. The company employs a range of professionals, including engineers and product managers, to enhance its Trust & Safety initiatives and leverage AI technologies.
Stack
LinuxGoKubernetescloud servicesdistributed systems
What you'll do
- Own reliability and operational excellence for production systems
- Improve production readiness for services and migrations
- Develop software for performance and automation
- Scale systems on bare-metal servers
- Lead incident reviews for continuous improvement
- Manage vendor relationships for quality services
What we're looking for
- 10+ years experience with high-scale production systems
- Strong fundamentals in Linux, networking, and databases
- Experience with observability systems and incident response
- Ability to write production-quality software in Go
Nice to have
- Comfortable debugging across various system layers
- Experience with capacity planning and cost management
