Back to job search
WH
Web Hosting Canada (WHC)Verified Job Source

DevOps Lead

The role involves owning the performance, reliability, and operational health of the server fleet through hands-on technical leadership. Key duties include automating server configurations via Git, ensuring robust backup and recovery processes, and maintaining security standards for PCI DSS and SOC 2 compliance.

  • Hybrid
  • Montréal, QC
  • Posted Jul 21, 2026
  • 1 position

More jobs you can apply to directly

Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.

Job summary

The Opportunity WHC is looking for a hands-on, technically deep DevOps Lead to own the performance, reliability, and operational health of our server fleet. Reporting to the Director of Technology, you’ll keep our servers fast and available, bring order and consistency to how we run the fleet, and make sure automation, security, and backups are solid so problems stop repeating. This is a doer-first role: an operator who leads by doing and sets the standard by example, not a manager directing from a distance. You’ll set standards for a small team of DevOps and system administration engineers, but the day-to-day is technical ownership, not headcount management. It is not an on-call role; our system administrators are the 24/7 first responders, and you’re the next level up. What You’ll Do • Keep our servers fast and available: plan for growth, and find and resolve performance problems across CPU, disk, databases, and web servers before they become incidents. • Bring the fleet to a known, consistent state: reduce the one-off snowflakes, close the gaps, and make “how a WHC server should look” something the whole team can point to and trust. • Treat Git as the single source of truth for server configuration; automate routine work so changes are safe and repeatable, and add checks that flag when a server has drifted from its expected state. • Keep our servers hardened and get real answers on how well our current tooling is actually protecting us, and support our PCI DSS and SOC 2 work. • Ensure backups are reliable and that we can rebuild a full server within our target recovery time, including a repeatable, documented server-migration capability. Test the restore; don’t trust the document. • Build and maintain a strong library of operational runbooks, keeping our procedures current and trusted so the team always has good documentation to work from. • Be the main day-to-day operational contact for our infrastructure vendors: support, escalations, and operational issues. • Look for simpler, safer, more efficient ways to run our systems, including the careful, practical use of AI in operations work. • Coach the engineers around you, set clear expectations, and partner well with Development and System Administration. Lead by owning outcomes, not by counting hours. What You Bring • Deep, hands-on Linux skills across a fleet of servers, not only cloud or container environments. You’re comfortable operating many machines at once and making them consistent. • Solid scripting and automation (Bash, plus Python or Perl), and real experience managing server configuration in Git and automating changes safely. • Production MySQL/MariaDB: replication, tuning, backup, and restore. • Experience with backups and disaster recovery, including meeting a defined recovery time. • A monitoring and performance mindset: you find the bottleneck rather than waiting for the ticket. • Security-by-default instincts for hardening and running servers. • Excellent written and spoken English; you'll communicate clearly with our team and vendors. • Ownership: you take a problem to done, document your work, and you don’t become the bottleneck for everyone else. • Clear, calm communication: you can steady a team during an incident and explain a technical issue in plain language. • Substantial experience in Linux systems or operations roles, including some experience setting standards for or leading a small team. Years matter less than range. Nice to Have • Web hosting at scale: cPanel/WHM, CloudLinux, LiteSpeed, shared hosting. • WHMCS, or a similar billing and automation platform. • FreeIPA, or another LDAP/Kerberos identity system. • Self-hosted virtualization (Proxmox, LXD, KVM) rather than only public cloud. • PCI DSS or SOC 2 experience. • Comfort with AI tools, and experience using AI safely and practically in operations work. This is an important plus. • French, and experience working with Canadian data. Why Join WHC? • A collaborative team culture where your work is visible, your decisions matter, and your impact is meaningful. • A Canadian, independent company with a clear market identity, a loyal customer base, and a strong focus on helping Canadians succeed online. • An AI-forward environment where we’re actively exploring how AI can improve how we build, operate, and support customers. • Competitive compensation and benefits with a flexible hybrid work model. • Access to training, mentorship, and career advancement opportunities. • Frequent gatherings, lively 5à7s, a vibrant social club, and memorable holiday parties. • A bright, renovated office in Little Italy with a fully stocked kitchen and a game room featuring ping-pong, foosball, Nintendo, Nerf guns, and a library. • Certified a Great Place to Work® for five years running, because we believe work should be fun, fulfilling, and rewarding. Ready to Make an Impact? If you’re a hands-on infrastructure operator who takes ownership, brings order to complexity, and cares about running a fleet you can defend, we want to hear from you. Apply today and help us keep the platform behind WHC fast, reliable, and ready for what’s next. WHC is an equal-opportunity employer. We welcome and encourage applications from all qualified candidates, including those with diverse backgrounds and abilities.

What you’ll do

The role involves owning the performance, reliability, and operational health of the server fleet through hands-on technical leadership. Key duties include automating server configurations via Git, ensuring robust backup and recovery processes, and maintaining security standards for PCI DSS and SOC 2 compliance.

Requirements

Candidates must have deep Linux expertise, strong scripting skills in Bash and Python, and production experience with MySQL/MariaDB. Previous experience setting standards for a small team and a strong mindset for monitoring and security are essential.

Benefits

• Competitive compensation • Flexible hybrid work model • Training • Mentorship • Career advancement opportunities • Social club • Office game room • Stocked kitchen

Listed skills

  • MySQLPreferred
  • GitPreferred
  • PythonPreferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • Linux Administration
  • Bash Scripting
  • Python
  • Perl
  • MySQL
  • MariaDB
  • Git
  • Infrastructure Automation
  • Disaster Recovery
  • Server Hardening
  • Performance Tuning
  • PCI DSS
  • SOC 2
  • Monitoring
  • Technical Leadership
  • Operational Runbooks

Job areas

  • Technology
  • Software
  • Engineering
  • Management & Leadership
  • Security & Safety

Additional details

Minimum experience
5+ years
Posting language
English
Working hours
40 hours per week