Role Overview
Specter is hiring a mid-level Site Reliability Engineer. This is a full-time role in San Francisco. Part of Specter's Lifecycle hiring, posted 2 weeks ago. Full responsibilities, required qualifications, and the apply link are listed in the description below.
Salary Context
Salary is not disclosed in this posting. Market median for Mid-level Lifecycle roles is $105k-$145k (based on 329 comparable listings). Many employers share specifics during the interview process or after an initial screen.
Resume Keywords to Include
Make sure these keywords appear in your resume to improve ATS scoring
Job description
Company Background
Specter's mission is to help automate the physical world.
Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical activity captured on camera, from security incidents (e.g. perimeter intrusion, theft, LPR), to safety monitoring (e.g. PPE detection, injured people), to operational efficiency (e.g. material tracking, congestion monitoring). We offer both long range wireless (1km range) and wired sensor variants to suit any deployment.
Our co-founders Xerxes and Philip are passionate about empowering our partners in the fast approaching world of physical AI and robotics. We are a small, fast growing team who hail from Anduril, Tesla, Uber, and the U.S. Special Forces.
The Role
We're hiring a Site Reliability Engineer to own the operational health of our connected sensor platform — spanning a live fleet of edge hardware deployed at customer sites and the cloud infrastructure behind it.
This is a high-ownership role at the intersection of ops and platform engineering. You'll drive reliability across our sensor fleet — triaging issues in the field, building the systems that prevent them from recurring, and owning the observability that keeps us ahead of problems as we scale.
You set your own priorities across all three:
Responsibilities
Reactive — Triage & Recovery
- Debug and triage issues across a live fleet of diverse Linux-based sensor nodes and edge appliances deployed at customer sites.
- SSH into field hardware to diagnose, patch, and recover systems — often with limited remote access and incomplete information.
- Own site bring-ups end to end; be the person who gets things back online.
Systems Builder — Close the Loop
- Build and maintain fleet management systems: OTA update pipelines, device health tracking, remote diagnostics, and lifecycle tooling.
- Identify repeat fires and eliminate them — build tooling, pre-deployment checks, and root cause processes that prevent recurrence.
- Automate toil relentlessly: if you're doing something twice, you should be scripting it.
- Collaborate with embedded systems, and platform teams to define reliability and deployment requirements.
Observability Owner — Fleet Visibility
- Design and implement observability (logging, metrics, alerting) across edge devices and cloud infrastructure (AWS).
- Surface and close telemetry gaps; build fleet-wide visibility that enables data-driven reliability decisions.
- Develop runbooks, incident response procedures, and participate in on-call rotations.
Qualifications
- Strong Linux systems administration — comfortable working over SSH in production, not just dev environments.
- Experience with edge or on-prem hardware alongside cloud infrastructure.
- Solid networking fundamentals: DNS, firewalls, VPNs, subnets, secure remote access.
- Scripting or programming in Python, Go, or Bash for operational tooling.
- Familiarity with containerization (Docker, Kubernetes a plus).
- Embedded systems experience — reading firmware logs, understanding hardware-software boundaries, and reasoning about what's happening below the OS is a meaningful edge in this role.
- Deeper cloud experience (AWS infrastructure, IAM, networking, observability tooling) is a strong plus for owning the cloud side of the fleet.
- Rust or C experience — we have firmware in both; being able to read and reason about low-level code accelerates triage significantly.
About Specter
Frequently Asked Questions
How do I apply for the Site Reliability Engineer position at Specter?
Use the Apply button above to submit your application directly to Specter. Most applications take less than 5 minutes if your resume and contact details are ready, and you'll be routed to the employer's official application system to finish.
Where is the Site Reliability Engineer position at Specter located?
This position is based in San Francisco. Specter has not indicated remote or hybrid options for this role, so candidates should plan for on-site work.
What does a Site Reliability Engineer at Specter earn?
Specter has not disclosed a salary range in this posting. Many employers share specifics later in the interview process; you can also ask during a recruiter screen if compensation transparency is important to you.
When was the Site Reliability Engineer role at Specter posted?
This role was posted on July 2, 2026 (20 days ago). It's still listed as actively hiring; we re-confirm openings against the source system multiple times per day and remove closed roles.
Similar Jobs
Regulatory Affairs Manager I
Alcon
Project Manager (18 Month Contract)
TMX Group
Senior Reliability Test Engineer
Harbingermotors
Customer Experience Manager (Dealer & Fleet)
Harbingermotors
Systems Engineer, ADAS
Harbingermotors
More Jobs at Specter
View all →AI-powered job search
Get every job scored to your resume
Upload your resume and get jobs ranked, your resume tailored, and employee contacts found automatically.
Get started freeNo credit card to start