Hardware Systems Architect na Anthropic

Presencial - San Francisco, CA; New York City, NY; Seattle, WA

Candidatar-se
Ver mais vagas na Anthropic

Anthropic trains and serves frontier AI models on some of the largest and most diverse accelerator fleets in the world, and the hardware those models run on is one of the most direct levers we have on capability, cost, and reliability. Our Hardware Systems team owns the system‑level architecture of that compute — from the package boundary out through boards, racks, interconnect, power, cooling, and the datacenter interface. We’re hiring senior hardware systems architects and technical leads who can work broadly across the hardware stack and go deep where it’s needed, and who have the judgment to make and own directional calls — what to build, what to buy, what to co‑design with a partner.

Requirements

Skills

  • Deep, hands‑on expertise in at least one core hardware‑systems domain (e.g., interconnect, power, thermal) for large‑scale compute or networking systems
  • Experience owning hardware system architecture at scale — machine, rack, row, and cluster — and authoring specifications and requirements
  • Experience taking hardware from architecture through bring‑up and high‑volume production deployment
  • Experience working with external hardware vendors and partners — reviewing their designs and holding them to a specification
  • Ability to reason about trade‑offs across adjacent hardware domains (e.g., interconnect ↔ power ↔ thermal ↔ mechanical)
  • Track record of technical leadership — owning directional decisions and their consequences, and driving alignment across teams and partners
  • Comfort operating with high autonomy and limited process in a fast‑moving, ambiguous environment
  • Strong written communication
  • 8+ years in hardware systems architecture or engineering for hyperscale, HPC, AI/ML, or high‑end networking platforms
  • Experience with AI accelerator systems (GPU, TPU, or other custom ASIC platforms) and their scale‑up / scale‑out fabrics
  • Familiarity with the chip‑package‑system interface and co‑design
  • Experience establishing a new function, program, or engineering practice
  • Comfort working with — and curiosity about pushing — AI tools as part of an engineering workflow
  • Degree in EE, CE, ME, CS, or a related field, or equivalent experience
  • Depth in one or more of: high‑speed SerDes, PCIe/CXL, Ethernet/InfiniBand, and optical interconnect; BMC and platform firmware, hardware root of trust, and secure boot; 48V/HVDC power delivery and direct liquid cooling; fleet‑scale reliability engineering and RAS architecture; signal and power integrity

Responsibilities

  • Own system path‑finding and architecture for your domain across the breadth of compute Anthropic deploys — from concept and requirements through spec, design review, and deployment — and act as a reviewer and thought partner across the others
  • Write and own high‑level requirements and specifications — system, board, interface, rack — that partners and internal teams build against
  • Review partner and vendor designs against our requirements, make the build / buy / co‑design calls with supply chain and partnership teams, and surface misalignment or schedule risk early enough to act on it
  • Drive the interconnect and networking architecture that ties our systems together — scale‑up and scale‑out fabrics, NICs and switches, optics, and topology
  • Drive board‑, chassis‑, and system‑level architecture with our partners — major component placement, PCB architecture, connector and cable design, and power and thermal budgeting — at the chassis, rack, and data‑hall level
  • Develop performance simulations, hardware models, and “what‑if” scenarios to evaluate and choose between candidate hardware architectures
  • Define validation, bring‑up, and qualification strategy for new platforms — including the telemetry and reliability targets they need to hit in the fleet — and track the technology roadmaps and vendor landscape that shape what we deploy in the short and long term
  • Partner with ML performance and infrastructure software teams so hardware decisions land well for the workloads that actually run on them
  • Help shape how hardware engineering operates at Anthropic — develop and apply AI‑assisted approaches to design, spec review, and validation; establish review forums, spec standards, and partner engagement

Technologies

InterconnectPowerCoolingMechanicalSignal integrityData planeControl planeSystem managementReliabilityServiceabilityNetworkingComputeStorageAccelerator systemsHigh‑speed SerDesPCIeCXLEthernetInfiniBandOptical interconnectBMCPlatform firmwareRoot of trustSecure boot48V/HVDC power deliveryDirect liquid coolingReliability engineeringRAS architectureAI tools

Compartilhar vaga

Descubra se seu currículo está pronto para esta vaga

Veja como nossa IA pode otimizar seu currículo e aumentar suas chances de conseguir esta posição.