Founding SRE Engineer - Burnaby - Opus

Opus Burnaby

6 days ago

Description

OpusClip is the world's No.1 AI video agent, built for authenticity on social media.

We envision a world where everyone can authentically share their story through video, with no expertise needed. Within just 18 months of our launch, over 10 million creators and businesses have used OpusClip to enhance their social presence.

We have raised $50 million in total funding and are fortunate to have some of the most supportive investors, including SoftBank Vision Fund, DCM Ventures, Millennium New Horizons, Fellows Fund, AI Grant, Jason Lemkin (SaaStr), Samsung Next, GTMfund, Alumni Ventures, and many more.

Check out our latest coverage by Business Insider featuring our product and funding milestones, and our recognition as one of The Information's 50 Most Promising Startups in 2024.

Headquartered in Palo Alto, we are a team of 100 passionate and experienced AI enthusiasts and video experts, driven by our core values:

Be a Champion Team
Prioritize Ruthlessly
Ship fast, Quality Follows
Obsess over customers

Be a part of this exciting journey with us

The Mission

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. You will stabilize our processing clusters, design isolated environments for our largest Enterprise partners, and serve as the technical bridge between infrastructure and our 15M+ users.

You will engineer the infrastructure strategy that underpins our trust and reliability in the market. You will help set up oncall rotation and own the full incident lifecycle from minimizing Time-to-Detect to tracking post mortem actions ensuring that our high-velocity growth never compromises our performance.

Key Responsibilities

Infrastructure Architecture & Cluster Operations

Architect Dedicated Environments: Lead the design and implementation of high-throughput, isolated processing clusters for Enterprise clients. You will build the "paved road" to ensure strict High Availability (HA) without noisy neighbor interference.
Scale Production: Drive general improvements in our Temporal clusters and production Kubernetes environments. You will operationalize scaling strategies that support both self-serve consumers and high-touch Enterprise contracts.
Technical Execution: Be hands-on with the stack to optimize resource allocation, reduce latency, and enforce isolation strategies for critical accounts.

Monitoring, Alerting & Detectability

Beat the Customer to the Alert: Overhaul our Datadog observability suite to aggressively reduce Time-to-Detect (TTD). You ensure we identify latency spikes and stalled projects before users do.
Threshold Tuning: tune alert thresholds to eliminate noise and focus on "symptom-based" alerts that reflect the actual user experience.
External SLO Ownership: Define and report on Service Level Objectives (SLOs), acting as the internal guarantor that we are meeting the targets we sold.

Incident Command & "Extreme Ownership"

First Responder & Mitigation: Serve as the first line of defense during outages. You will own immediate mitigation, including cluster debugging and manual scaling intervention if Horizontal Pod Autoscalers (HPA) fail or lag.
Drive Recovery Metrics: You are accountable for shortening Time-to-Mitigation (TTM) and Time-to-Recover (TTR). Your priority is to stop the bleeding first, then fix the wound.
Root Cause Analysis: Lead the post-mortem process to determine Time-to-Root Cause and implement systemic fixes. You will translate these technical findings into clear updates for Customer Experience (CX) and Leadership.
Accountability: Work collaboratively with Engineering Owners to track improvements against the reliability roadmap. You are responsible for flagging risks early and resetting expectations on platform performance when necessary.
Cross-Functional Bridge: Serve as the primary technical voice to the Customer Experience (CX), Sales, and Leadership teams. You will translate technical constraints and roadmaps into clear, honest updates for stakeholders.

Qualifications

Production K8s & Temporal: Expert-level ability to debug Kubernetes internals (HPA logic, node scaling) and operate stateful workflow engines (Temporal) at scale.
Incident Command: Proven track record as a primary first responder, demonstrating the ability to aggressively reduce Time-to-Mitigation (TTM) and Time-to-Recover (TTR).
Observability Architecture: Experience architecting Datadog SLOs and tuning alerts to distinguish system noise from actual user pain.
Automation: Strong proficiency in Python or Bash to automate manual recovery and operational tasks.
Bonus: Experience scaling GPU/video rendering workloads or thriving in early-stage, high-velocity startups.

The Tech Stack

Orchestration & Compute: Kubernetes (GKE), Docker, Horizontal Pod Autoscaling (HPA).
Workflow Engine: Temporal
Observability: Datadog (APM, Custom Metrics, Alerting).
Infrastructure as Code: Terraform or similar IaC tools.
Scripting & Backend: Python (primary), Bash.
Data & Storage: Redis, Milvus (Vector DB), Postgres.

EEO

OpusClip is proud to be an equal opportunity employer. We do not discriminate in hiring or any employment decision based on race, color, religion, national origin, age, sex (including pregnancy, childbirth, or related medical conditions), marital status, ancestry, physical or mental disability, genetic information, veteran status, gender identity or expression, sexual orientation, or other applicable legally protected characteristics. OpusClip considers qualified applicants with criminal histories, consistent with applicable federal, state and local law. Opus Clip is also committed to providing reasonable accommodations for qualified individuals with disabilities and disabled veterans in our job application procedures.

#J-18808-Ljbffr

Founding SRE Engineer
3 weeks ago

Only for registered members Burnaby

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · ...
Founding SRE Engineer
1 week ago

Only for registered members Burnaby

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · Production K8s & Temporal: Expert-level ability to debug Kubernetes internals (HPA logic, node scaling) and operate stateful workflow engines (Temporal) at scale. · Incid ...
Founding SRE Engineer
3 weeks ago

Only for registered members Burnaby $160,000 - $235,000 (CAD)

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · You will stabilize our processing clusters, design isolated environments for our largest Enterprise partners, · and serve as the technical bridge between infrastructure a ...
Founding SRE Engineer
3 weeks ago

Only for registered members Burnaby Full time

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · ...
Founding SRE Engineer
3 weeks ago

Only for registered members Burnaby, BC

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · You will stabilize our processing clusters, design isolated environments for our largest Enterprise partners, and serve as the technical bridge between infrastructure and ...
DevOps / SRE Engineer
1 month ago

Only for registered members Vancouver

We are seeking a talented and driven DevOps / SRE Engineer to help scale, manage and secure our cloud infrastructure as we continue to grow. · Design, build, and maintain scalable and secure cloud infrastructure. · Develop and optimize CI/CD pipelines to enable efficient software ...
DevOps / SRE Engineer
1 month ago

Only for registered members Vancouver, British Columbia

We are seeking a talented and driven DevOps / SRE Engineer to help scale manage and secure our cloud infrastructure as we continue to grow. · ...
Senior Software Engineer, Core Services SRE
5 days ago

Only for registered members Vancouver Full time

The Core Services Site Reliability Engineering (Core Services SRE) team sets the foundations and maintains the principles of reliable engineering across StackAdapt's core service teams. · ...
SRE Specialist
5 days ago

Only for registered members Burnaby, BC, Canada

We are the SSP (Support Systems and Processes) SRE Automation Team at Fortinet and passionate about building improving and maintaining various information systems that serve our employees worldwide as well as consumer-facing services with high traffic volumes around the world. · ...
SRE Specialist/DevOps Developer
5 days ago

Only for registered members Burnaby, BC, Canada

++Join Fortinet as we continue to shape the future of cybersecurity. We seek a dynamic SRE Specialist/DevOps Developer for our rapidly growing business. · + · ,+>+ · +,100% company paid medical, dental, and vision coverage,+a Health Spending Account,+a Personal Spending Account, ...
Senior Back-end Engineerr
5 days ago

Only for registered members Burnaby

+We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · +Be a Champion Team · Prioritize Ruthlessly · Ship fast, Quality Follows · Obsess over customers · ,+,+ · Comprehensive medical · vision, · and dental coverage support y ...
Senior Back-end Engineerr
1 week ago

Only for registered members Burnaby

We envision a world where everyone can authentically share their story through video, with no expertise needed. We have raised $50 million in total funding and are fortunate to have some of the most supportive investors. Check out our latest coverage by Business Insider featuring ...
SRE Specialist
5 days ago

Only for registered members Burnaby, BC, Canada

We are recruiting a Site Reliability Engineer in OpenStack to join our FortiStack team. · This role would represent a great fit for Openstack specialists or IT professionals with a combination of DevOps/GitOps, virtualization, Openstack, storage and networking experience. · ...
SRE Specialist
5 days ago

Only for registered members Burnaby, BC, Canada

We are a service-focused team managing high-traffic consumer-facing systems deployed globally. Our responsibility spans the full lifecycle of services running on top of OpenStack Kubernetes and physical/virtual infrastructure. We own both the operational stability of the services ...
SRE Specialist
4 weeks ago

Only for registered members Burnaby $112,000 - $124,000 (CAD)

Job summary:We are the SSP (Support Systems and Processes) SRE Automation Team at Fortinet and passionate about building, · improving, and maintaining various information systems that serve our employees worldwide, · as well as consumer-facing services with high traffic volumes a ...
Senior Cloud Architect
1 week ago

Only for registered members Burnaby Full time $200,000 - $225,000 (CAD)

This position offers a competitive salary plus variable compensation based on performance targets and business objectives. Tantalus also offers generous benefits. · Tantalus Systems (TSX: GRID) is a technology company dedicated to helping utilities modernize their distribution gr ...
Senior DevOps Developer
1 month ago

Only for registered members Burnaby

Join Fortinet as we continue to shape the future of cybersecurity. We are seeking a dynamic DevOps Developer to contribute to our rapidly growing business . As a Senior DevOps Developer your responsibilities will include maintaining and improving our DevOps infrastructure (on-pre ...
Senior Back-end Engineerr
1 week ago

Only for registered members Burnaby $160,000 - $235,000 (CAD)

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. · You will stabilize our processing clusters, design isolated environments for our largest Enterprise partners, and serve as the technical bridge between infrastructure and ...
Senior Back-end Engineer
13 hours ago

Only for registered members Burnaby $160,000 - $235,000 (CAD)

We envision a world where everyone can authentically share their story through video, · with no expertise needed.We have raised $50 million in total funding and are fortunate to have some of the most supportive investors. · ...
Senior Back-end Engineerr
1 week ago

Only for registered members Burnaby Full time $160,000 - $235,000 (CAD)

We are looking for a hands-on Founding SRE to own the reliability and scalability of our platform. · ...
SRE Specialist
1 month ago

Only for registered members Burnaby Full time $100,800 - $138,300 (CAD)

Fortinet is recruiting a Site Reliability Engineer in OpenStack to join our FortiStack team. · ...

Founding SRE Engineer
Only for registered members Burnaby
Founding SRE Engineer
Only for registered members Burnaby
Founding SRE Engineer
Only for registered members Burnaby
Founding SRE Engineer
Full time Only for registered members Burnaby
Founding SRE Engineer
Only for registered members Burnaby, BC
DevOps / SRE Engineer
Only for registered members Vancouver
DevOps / SRE Engineer
Only for registered members Vancouver, British Columbia
Senior Software Engineer, Core Services SRE
Full time Only for registered members Vancouver
SRE Specialist
Only for registered members Burnaby, BC, Canada
SRE Specialist/DevOps Developer
Only for registered members Burnaby, BC, Canada
Senior Back-end Engineerr
Only for registered members Burnaby
Senior Back-end Engineerr
Only for registered members Burnaby
SRE Specialist
Only for registered members Burnaby, BC, Canada
SRE Specialist
Only for registered members Burnaby, BC, Canada
SRE Specialist
Only for registered members Burnaby
Senior Cloud Architect
Full time Only for registered members Burnaby
Senior DevOps Developer
Only for registered members Burnaby
Senior Back-end Engineerr
Only for registered members Burnaby
Senior Back-end Engineer
Only for registered members Burnaby
Senior Back-end Engineerr
Full time Only for registered members Burnaby
SRE Specialist
Full time Only for registered members Burnaby