Senior Back-end Engineer Job at OpusClip (Burnaby)

Job Description

We are looking for a hands-on Founding SRE to own the reliability and scalability of the OpusClip platform. You will stabilize our processing clusters, design isolated environments for our largest Enterprise partners, and serve as the technical bridge between infrastructure and our 15M+ users. You will engineer the infrastructure strategy that underpins our trust and reliability in the market. You will help setup oncall rotation and own the full incident lifecycle from minimizing Time-to-Detect to tracking post mortem actions ensuring that our high-velocity growth never compromises our performance.

Job Responsibility

Architect Dedicated Environments: Lead the design and implementation of high-throughput, isolated processing clusters for Enterprise clients
Scale Production: Drive general improvements in our Temporal clusters and production Kubernetes environments
Technical Execution: Be hands-on with the stack to optimize resource allocation, reduce latency, and enforce isolation strategies for critical accounts
Beat the Customer to the Alert: Overhaul our Datadog observability suite to aggressively reduce Time-to-Detect (TTD)
Threshold Tuning: tune alert thresholds to eliminate noise and focus on 'symptom-based' alerts
External SLO Ownership: Define and report on Service Level Objectives (SLOs)
First Responder & Mitigation: Serve as the first line of defense during outages
Drive Recovery Metrics: You are accountable for shortening Time-to-Mitigation (TTM) and Time-to-Recover (TTR)
Root Cause Analysis: Lead the post-mortem process to determine Time-to-Root Cause and implement systemic fixes
Accountability: Work collaboratively with Engineering Owners to track improvements against the reliability roadmap
Cross-Functional Bridge: Serve as the primary technical voice to the Customer Experience (CX), Sales, and Leadership teams

Requirements

Production K8s & Temporal: Expert-level ability to debug Kubernetes internals (HPA logic, node scaling) and operate stateful workflow engines (Temporal) at scale
Incident Command: Proven track record as a primary first responder, demonstrating the ability to aggressively reduce Time-to-Mitigation (TTM) and Time-to-Recover (TTR)
Observability Architecture: Experience architecting Datadog SLOs and tuning alerts to distinguish system noise from actual user pain
Automation: Strong proficiency in Python or Bash to automate manual recovery and operational tasks

Nice to have

Experience scaling GPU/video rendering workloads or thriving in early-stage, high-velocity startups

What we offer

Comprehensive medical, vision, and dental coverage
Flexible paid time off
Generous equipment, software, and office furniture budget
Access to the latest technologies and tools
Free lunch and dinner, plus a wide variety of global snacks and beverages for onsite positions in North America
Visa sponsorship available, subject to eligibility
Competitive equity packages (ISOs)
Performance-based bonus/commission

OpusClip - All Job Offers

Select Country

Senior Back-end Engineer

Job Description

Job Responsibility

Requirements

Nice to have

What we offer

Looking for more opportunities?

Senior Back-end Engineer

Senior Back-End Engineer

Senior Back-End Engineer

Senior Staff Principal Back-end and Front-end Engineer

Senior Back-end Engineer (Machine Learning)

Senior Software Engineer, Back-End

Senior Software Engineer (Back-end)

Senior Software Engineer, Back-end

Back-End Senior Software Engineer

Our AI answers in your language