Job Opportunities API

The Public Ledger of Openings

← Back to the ledger

Staff/Senior Software DevOps Engineer

Olix
CompanyOlix
CategoryEngineering
LocationToronto
RemoteOn-site (inferred)
EmploymentNot stated
LevelSenior
SalaryNot stated by the employer
Posted15 Jul 2026
Last verified12 Aug 2026
SourceThe employer's own careers page (company_site)
Applications are handled by the employer, not by us.Apply on the employer's site →
Description
OLIX AI is building the DX-1 decode accelerator, a next-generation hardware system for AI inference. This Staff/Senior Software DevOps Engineer role owns the entire build, test, and CI infrastructure that the compiler, runtime, simulator, and framework teams depend on, managing a large test suite that runs across scarce hardware resources including simulation compute and hardware-in-the-loop testing platforms. What You'll Do • Design, build, and own CI pipelines across PR, merge, and nightly lanes that gate the entire software stack, balancing feedback speed against coverage and cost • Scale and optimize a large test suite through staged lanes, parallelization with real test isolation, and content-addressed caching to maintain fast, affordable feedback as the suite and team grow • Manage heterogeneous CI runner fleets spanning cloud and self-hosted machines (VMs, containers, bare-metal), providing fair and monitored shared access to scarce and expensive hardware resources • Establish performance regression baselines and turn CI and test signal into CI-health and product-readiness dashboards that inform real engineering decisions • Own software observability infrastructure including metrics collection, storage, and dashboard systems; define standards for hermetic, reproducible builds and fail-closed system behavior What You Need • Demonstrated end-to-end ownership of a large-scale build/test infrastructure, CI/CD system, or release engineering project • Expertise scaling large test suites using staged lanes, parallelism with real isolation, content-addressed caching, and cost-aware optimization • Experience managing heterogeneous CI runner fleets across cloud and on-premises infrastructure, including custom accelerators, FPGA/prototype hardware, and lab automation • Strong performance-analysis skills: trustworthy regression baselines, determinism handling, sound metric aggregation, and fast bisection capabilities • Strong scripting and systems programming (Python plus a systems language), fluency with containers and Linux, and cloud infrastructure experience (AWS or similar) Nice to Have • GitHub Actions or comparable CI platforms at scale; scaling CI runner fleets on cloud infrastructure (e.g. AWS) • Hardware-in-the-loop or lab automation experience for custom silicon or FPGA bring-up • Time-series and observability stack experience (Prometheus, Grafana, Datadog, columnar warehouses) • Adjacent depth in HPC/cluster batch scheduling, release engineering, or developer-productivity platforms Competitive salary commensurate with experience, skills, and location; meaningful stock options; annual living-local bonus if residence within 20 minutes of office; employer-contributed retirement plans