Salary not listed

Staff Site Reliability Engineer

DevOpsRemote — USA
Published on 2026-10-06
These details were extracted automatically from the original listing. They may not be complete or fully up to date — it's worth checking the original job post.

About this role

Bluesky is seeking a Staff Site Reliability Engineer to design and manage the infrastructure for their federated social network. The role involves ensuring the reliability and performance of production systems while collaborating with engineers across teams.

About the company

Bluesky is focused on developing a decentralized social media platform that prioritizes user safety and trust. The company employs a range of professionals, including engineers and product managers, to enhance its Trust & Safety initiatives and leverage AI technologies.

Stack

LinuxGoKubernetescloud servicesdistributed systems

What you'll do

  • Own reliability and operational excellence for production systems
  • Improve production readiness for services and migrations
  • Develop software for performance and automation
  • Scale systems on bare-metal servers
  • Lead incident reviews for continuous improvement
  • Manage vendor relationships for quality services

What we're looking for

  • 10+ years experience with high-scale production systems
  • Strong fundamentals in Linux, networking, and databases
  • Experience with observability systems and incident response
  • Ability to write production-quality software in Go

Nice to have

  • Comfortable debugging across various system layers
  • Experience with capacity planning and cost management
View original job post