Senior Site Reliability Engineer
Umbra is an American space technology company delivering advanced systems, from sensors to spacecraft, that empower customers worldwide with unmatched access to critical information from space. Our mission is simple and ambitious: redefine space—for people, systems, and missions in every domain. Umbra’s ecosystem operates through three business units: Remote Sensing (the data), Space Systems (the components), and Mission Solutions (the platforms).Together, our teams develop capabilities that deliver persistent access, resilient performance, and mission-ready solutions, advancing U.S. space leadership while keeping the world safe and informed.
About the Team
Remote Sensing – The Data
Remote Sensing is where Umbra got its start, and our agile Synthetic Aperture Radar (SAR) constellation remains the most capable on the market. We transform satellite data into real-world, actionable insights that strengthen U.S. national security and intelligence, support disaster response, and advance scientific discovery. Our team delivers data at scale with unmatched quality, persistence, and the speed and responsiveness our partners demand.
If you want to work on cutting-edge space technology that’s redefining what’s possible in remote sensing, you belong here at Umbra.
About the Job
We are seeking an experienced Senior Site Reliability Engineer to help design, build, operate, and scale the mission- and business-critical infrastructure that powers Umbra's systems. In this role, you will leverage a deep understanding of modern infrastructure, distributed systems, and the broader technology stack to drive technical excellence, make thoughtful architectural decisions, and balance long-term scalability with operational reliability.
You'll partner closely with engineering teams to improve processes, champion best practices, evaluate emerging technologies, and implement solutions that enhance the performance, resilience, and efficiency of our platforms. The ideal candidate is a collaborative technical leader who communicates effectively across technical and non-technical teams and drives meaningful improvements that have a lasting impact across the organization.
This position is based on-site in either our Arlington, VA office, Reston, VA office or Santa Barbara/Goleta, CA office.
Key Responsibilities
- Ensure the reliability and scalability of critical systems, meeting SLAs through proactive monitoring and effective incident response.
- Develop and promote new technologies and tools, conducting research and creating proofs of concept to introduce solutions that enhance the team's capabilities.
- Lead by example in fostering a culture of excellence and reliability.
- Continuously evaluate and improve team processes and workflows to increase efficiency and reduce complexity.
- Collaborate closely with cross-functional teams, product managers, and stakeholders to align on technical strategy and provide expert guidance.
- Participate in on-call rotations, providing support and resolving complex technical issues.
Requirements
Required Qualifications
- Bachelor’s degree in Computer Science or a related technical field.
- 5-8+ years in a Site Reliability Engineer or DevOps role supporting a SaaS platform, with demonstrated expertise managing distributed systems.
- Extensive experience with AWS services (EC2, S3, Lambda, VPC Networking) and deep knowledge of cloud infrastructure, networking, and security best practices.
- Proficiency running, optimizing, and scaling Kubernetes clusters in production environments.
- Experience using and writing Terraform to architect and manage production infrastructure.
- Ability to create and utilize Infrastructure-as-code (IaC), GitOps practices, and automation tools to increase reliability and reduce manual tasks.
- Proven success in leading teams or projects using Agile/Scrum methodologies.
- Expertise in infrastructure and software architecture, capable of designing and implementing large-scale, reliable systems with minimal guidance.
- Experience developing and managing comprehensive infrastructure monitoring and alerting strategies.
Desired Qualifications
- 10+ years in a Site Reliability Engineer or DevOps role supporting a SaaS platform, with demonstrated expertise managing distributed systems.
- Advanced understanding of cloud and application security, identity management, and compliance.
- Expertise in service mesh and service registration technologies, focusing on performance and reliability.
- Experience in the aerospace industry.
Benefits
- Flexible Time Off, Sick, Family & Medical Leave
- Medical, Dental, Vision, Life, LTD, STD (employer funded)
- Vol Life, Critical Illness, Accidental, Hospital Indemnity, Pet Insurance (employee funded)
- 401k with 3% non-elective company contribution
- Stock Options
- Free Parking
- Free lunch daily in office
Umbra is an Equal Opportunity Employer. We do not discriminate in hiring on the basis of sex, gender identity, sexual orientation, race, color, religious creed, national origin, physical or mental disability, protected veteran status, or any other characteristic protected by federal, state, or local law.
Employment Eligibility Verification
In compliance with federal laws, all hired persons will be required to verify their identity and eligibility to work in the United States by completing the required Employment Eligibility Verification Form (I-9 Form) upon hire.
ITAR/EAR Requirements
This position may include access to technology and/or data that is subject to U.S. export controls pursuant to ITAR and EAR. To comply with federal export controls, all persons hired must be a U.S. citizen, U.S. national, U.S. lawful permanent resident, refugee or asylee as defined by 8 U.S.C. § 1324b(a)(3), or must otherwise be eligible to obtain the required authorizations from the U.S. Department of State and/or U.S. Department of Commerce as applicable.
Pay Transparency
This job posting may cover multiple career levels. To ensure greater transparency, we provide base salary ranges for all roles, regardless of location. Our standard pay ranges are based on the role’s function and level, benchmarked against similar growth-stage companies. Compensation may vary based on geographical location, as certain regions may have different cost-of-living factors. The final offer will also be influenced by the candidate's skills, responsibilities, and relevant experience.
Compensation Range
The Compensation Range for this role is $150,000 - $180,000 DOE.
Skills
As published by workable · 8 questions
Basics
First name, Last name, Email, Phone, Address, Resume, Are you legally authorized to work in the US for any employer? , To conform to U.S. Government space technology export regulations, including the International Traffic in Arms Regulations (ITAR) you must be a U.S. citizen, lawful permanent resident of the U.S., protected individual as defined by 8 U.S.C. 1324b(a)(3), or eligible to obtain the required authorizations from the U.S. Department of State. , Cover letter, How did you hear about Umbra? , What is the highest level of education that you completed?, What is your desired base pay?, When are you available to start?, Various federal laws restrict the employment and post-government employment (PGE) activities of federal government employees. Please identify your employment status with the United States Federal Government, including civilian or military. Umbra may request additional detail on the nature of any U.S. Federal Government employment and may require provision of an agency ethics opinion letter., Similar to federal laws there may be state and local laws that regulate the post-government employment (PGE) activities of current or former state or local government (municipal government) employees. These regulations may vary from state to state but may address activities related to contracting, procurement, regulatory or legislative roles, or other public employment within a municipal government. Please select your employment status with respect to municipal governments., If you answered "I have never..." to both of the above questions, please select N/A. If you answered with something else, answer the following question: Did you in the past, or are you now, participating in, supervising, or otherwise responsible for any active or pending government procurement or contract administration matter that does or might involve Umbra Lab, Inc. and/or a matter that has a direct or potential effect on the financial or business interests of Umbra Lab, Inc.?, If you answered "I have never..." to both of the above questions, please select N/A. If you answered with something else, answer the following question: To the best of your knowledge, is any member of your immediate family (including spouse, parents, siblings, and children), member of your household, or person with whom you have a similarly close personal relationship a current employee of the U.S. Federal Government, a municipal government, or a candidate for political office?
Pick from a list (8)
- How many years of professional experience do you have working as a Site Reliability Engineer, DevOps Engineer, or Platform Engineer supporting production SaaS platforms?
- Which AWS services have you used in production? (Select all that apply.)
- Which best describes your experience managing Kubernetes in production?
- Which best describes your experience with Terraform?
- Which best describes your experience designing production infrastructure?
- Do you have experience in the aerospace industry?
- Do you currently reside in (or within commuting distance to) either the Arlington, VA area, Reston, VA area, or Santa Barbara, CA area?
- If you don't currently reside in (or within commuting distance to) the Santa Barbara, CA, Arlington, VA, or Reston, VA areas, are you able to relocate to either area, as this is primarily an in-office role?