Technical Lead Manager, Infrastructure Hardware (Server and Network Systems)
cerebras · United States and Canada
Job description
About the role
As a Technical Lead Manager for Infrastructure Hardware, you will drive end‑to‑end execution of server and network platform programs for Cerebras AI clusters. You will own the technical delivery of new product introductions, platform refreshes and major configuration changes, working across OEM/ODM partners, internal software teams and data‑center stakeholders.
Key responsibilities
- Own end‑to‑end technical execution for server systems and network equipment, including NPIs and platform refreshes.
- Gather requirements, evaluate trade‑offs and translate them into executable plans with clear milestones and readiness gates.
- Lead OEM/ODM, switch‑vendor and component‑vendor engagements, including RFI/RFP, technical evaluations and roadmap alignment.
- Partner with Compute, Server Platform and Network Architects to define qualification plans, acceptance criteria and rollout strategies.
- Drive NPI qualification, lab/staging validation, issue resolution and go/no‑go decisions.
- Manage risk, change and versioning for production rollouts and ensure operational readiness with deployment teams.
Required profile
- B.S. or M.S. in Computer Science, Electrical/Computer Engineering or equivalent experience.
- 8+ years of technical leadership or systems‑engineering experience on server, network or infrastructure platforms.
- Proven ability to lead complex hardware programs across OEM/ODM, switch vendors and component suppliers.
- Strong multi‑team execution skills, including integrated planning, risk management and executive communication.
- Experience with AI/ML, HPC or performance‑sensitive distributed infrastructure is a plus.
Required skills
- Server architecture: CPU/NUMA, memory bandwidth, PCIe, NIC, storage I/O.
- Networking fundamentals: leaf‑spine fabrics, switch platforms, optics, high‑performance interconnects.
- Linux server fleet management: provisioning, firmware/BIOS, drivers, field triage.
What we offer
- Opportunity to build a breakthrough AI platform beyond GPU constraints.
- Work on one of the fastest AI supercomputers in the world.
- Publish and open‑source cutting‑edge AI research.
- Job stability combined with startup vitality and a non‑corporate culture.
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Canada.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 2 hours ago
Expires 1 month from now
3 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
cerebras
United States and Canada
Related job offers
-
Principal AI Security Engineer
cerebras United States and Canada -
Software Engineer, Kernel Reliability
cerebras United States and Canada -
Applied Machine Learning Research Scientist
cerebras United States and Canada -
Senior Backend Engineer (Remote, Canada)
shortstory Canada (Remote) -
Technical Support Engineer
sentry Toronto