Back to job search
BA
Boson AIVerified Job Source

Datacenter Technician

Maintain and repair physical AI infrastructure including Supermicro servers, NVIDIA GPUs, and high-speed networking equipment. Ensure maximum uptime through proactive maintenance, hardware replacements, and precise asset tracking.

  • On-site
  • Barrie, ON
  • Posted Aug 28, 2026
  • Apply by Sep 27, 2026
  • 1 position

More jobs you can apply to directly

Similar opportunities posted by employers hiring on Jobs.ca, with no external application form.

Job summary

Boson AI is an early-stage startup building large language tools for everyone to use. Our founders (Alex Smola, Mu Li), and a team of Deep Learning, Optimization, NLP, AutoML and Statistics scientists and engineers are working on high quality generative AI models for language and beyond. You keep the machines running We are looking for a Datacenter Hardware Technician to keep the physical infrastructure behind our AI research running in Barrie, ON. You will work hands-on with the latest NVIDIA GPUs, thousands of disks, terabit networking and hundreds of Supermicro servers — racking them, repairing them, keeping firmware current, and making sure every machine that should be online is online. Our researchers train models around the clock, so the difference between a good day and a bad one is often a technician who noticed a marginal cable, logged a serial number correctly, or caught a failing drive before it took a training job down with it. This is careful, methodical work, and we treat it that way. This role is for our Barrie datacenter. You must live within 50 km of the site, such that you can be onsite on short notice. Workload can be bursty, i.e. periods of smooth sailing mixed with periods of intense work during hardware failures, upgrade and maintenance cycles. A day in the life Install, rack, cable and commission new Supermicro servers, storage and network equipment Diagnose and repair hardware failures — replace DIMMs, drives, power supplies, fans, GPUs, cables and mainboards — and drive RMA cases with vendors through to resolution Install and update drivers, BIOS and firmware across servers, NICs, HBAs and switches, keeping fleet versions consistent and documented Test and install network connections, including structured cabling, optics and link validation, and troubleshoot physical-layer faults Perform preventive maintenance: inspections, cable management, airflow and filter checks, and spare-parts inventory Keep accurate records of every asset, serial number, part swap and rack location Respond to hardware failure alerts, and escalate to the SRE team when a fault is not purely physical Follow runbooks and ESD and safety procedures precisely — and improve them where they are unclear If you take pride in a tidy rack, a clean cable run, and a fleet where every machine is accounted for, we'd love to hear from you. We may use artificial intelligence (AI) tools to support parts of the hiring process, such as reviewing applications, analyzing resumes, or assessing responses and identifying potential inconsistencies or verification signals in application materials based on available information. These tools assist our recruitment team but do not replace human judgment. Final hiring decisions are ultimately made by humans. If you would like more information about how your data is processed, please contact us. Minimum Qualifications 2+ years of hands-on experience with server hardware in a datacenter, colocation, IT operations or equivalent environment Practical experience replacing and troubleshooting server components: drives, memory, power supplies, fans, GPUs and mainboards Experience installing drivers, BIOS and firmware updates on server hardware Experience testing and installing network connections, including structured cabling Comfortable on the Linux command line — navigating the filesystem, reading logs, running diagnostics Familiar with remote access and operational practices: SSH, VPN, and out-of-band management (IPMI, BMC, Redfish) Exceptionally organized and meticulous. Accurate records, careful labelling and consistent procedure matter more here than raw speed Able to work onsite in Barrie, ON, living within 50 km of the site Comfortable with the physical demands of the role: lifting and racking equipment, working in cold aisles, and occasional after-hours or on-call response Clear written communication — your notes are what the next person relies on Preferred Qualifications Direct experience with Supermicro servers, chassis and IPMI tooling Experience with GPU servers (NVIDIA H100, A100 or similar) and their power and cooling requirements Experience with high-speed networking: 100Gb+ Ethernet, InfiniBand, optics and DAC cabling Familiarity with DCIM or asset-management tools such as NetBox Basic scripting in Bash or Python to automate repetitive checks Experience with power distribution units, structured cabling standards, and rack and power design Experience handling RMA workflows with hardware vendors Relevant certifications (CompTIA Server+, A+, Network+) or an equivalent hands-on track record

What you’ll do

Maintain and repair physical AI infrastructure including Supermicro servers, NVIDIA GPUs, and high-speed networking equipment. Ensure maximum uptime through proactive maintenance, hardware replacements, and precise asset tracking.

Requirements

Requires 2+ years of datacenter hardware experience and proficiency with Linux and remote management tools. Candidates must live within 50 km of Barrie, ON, and be capable of performing physical labor including lifting equipment.

Listed skills

  • Hardware DiagnosticsPreferred
  • Preventive MaintenancePreferred
  • Asset ManagementPreferred
  • VPNPreferred

Other relevant skills

Identified from the job description. Confirm important requirements above.

  • Server Hardware Maintenance
  • GPU Troubleshooting
  • Network Cabling
  • Linux Command Line
  • Firmware Updates
  • BIOS Configuration
  • IPMI
  • BMC
  • Redfish
  • SSH
  • VPN
  • Asset Management
  • RMA Processing
  • Hardware Diagnostics
  • Structured Cabling
  • Preventive Maintenance

Job areas

  • Technology
  • Engineering
  • Trades
  • Data & Analytics
  • Software

Additional details

Minimum education
Professional degree
Minimum experience
2+ years
Apply by
Sep 27, 2026
Posting language
English
Working hours
40 hours per week
Seniority
Not Applicable
Application method
Direct apply is available