Staff/Senior Software DevOps Engineer
Olix
| Company | Olix |
| Category | Engineering |
| Location | Toronto |
| Remote | On-site (inferred) |
| Employment | Not stated |
| Level | Senior |
| Salary | Not stated by the employer |
| Posted | 15 Jul 2026 |
| Last verified | 12 Aug 2026 |
| Source | The employer's own careers page (company_site) |
Description
OLIX AI is building the DX-1 decode accelerator, a next-generation hardware system for AI inference. This Staff/Senior Software DevOps Engineer role owns the entire build, test, and CI infrastructure that the compiler, runtime, simulator, and framework teams depend on, managing a large test suite that runs across scarce hardware resources including simulation compute and hardware-in-the-loop testing platforms.
What You'll Do
• Design, build, and own CI pipelines across PR, merge, and nightly lanes that gate the entire software stack, balancing feedback speed against coverage and cost
• Scale and optimize a large test suite through staged lanes, parallelization with real test isolation, and content-addressed caching to maintain fast, affordable feedback as the suite and team grow
• Manage heterogeneous CI runner fleets spanning cloud and self-hosted machines (VMs, containers, bare-metal), providing fair and monitored shared access to scarce and expensive hardware resources
• Establish performance regression baselines and turn CI and test signal into CI-health and product-readiness dashboards that inform real engineering decisions
• Own software observability infrastructure including metrics collection, storage, and dashboard systems; define standards for hermetic, reproducible builds and fail-closed system behavior
What You Need
• Demonstrated end-to-end ownership of a large-scale build/test infrastructure, CI/CD system, or release engineering project
• Expertise scaling large test suites using staged lanes, parallelism with real isolation, content-addressed caching, and cost-aware optimization
• Experience managing heterogeneous CI runner fleets across cloud and on-premises infrastructure, including custom accelerators, FPGA/prototype hardware, and lab automation
• Strong performance-analysis skills: trustworthy regression baselines, determinism handling, sound metric aggregation, and fast bisection capabilities
• Strong scripting and systems programming (Python plus a systems language), fluency with containers and Linux, and cloud infrastructure experience (AWS or similar)
Nice to Have
• GitHub Actions or comparable CI platforms at scale; scaling CI runner fleets on cloud infrastructure (e.g. AWS)
• Hardware-in-the-loop or lab automation experience for custom silicon or FPGA bring-up
• Time-series and observability stack experience (Prometheus, Grafana, Datadog, columnar warehouses)
• Adjacent depth in HPC/cluster batch scheduling, release engineering, or developer-productivity platforms
Competitive salary commensurate with experience, skills, and location; meaningful stock options; annual living-local bonus if residence within 20 minutes of office; employer-contributed retirement plans