Senior Director, AI Platforms

Trace3 - Irvine, CA

Hiring: Senior Director, AI Platforms Company: Trace3 Location: Irvine, CA Job Posted Time: 2026-09-03 14:28:32 Target Skills & Keywords : Fine-tuning, Kubernetes, MLOps, Product Management, RAG About the job Experience: •12+ years of progressive experience in enterprise infrastructure, with at least 5+ years hands-on designing and implementing AI, ML, or HPC environments at production scale. •5+ years of people leadership experience with direct accountability for building and growing technical presales or architecture teams. Required Skills: •Build, hire, coach, and retain a team of presales solutions architects covering AI infrastructure and HPC, orchestration and observability, and AI-specific cybersecurity. •Own team capacity planning, deal support coverage, and utilization targets against pipeline volume and velocity. •Establish role definitions, technical career paths, certification roadmaps, and performance standards aligned with broader OCTO presales norms. •Partner with Data & Analytics, Security Solutions, Digital, and Innovation teams to present a single technical front to customers on AI opportunities. •Serve as the primary technical architect for Trace3's largest and most complex AI infrastructure opportunities, including greenfield AI factory builds, training and inference cluster designs, and edge AI deployments. •Architect end-to-end designs covering GPU compute, CPU head nodes, scale-out and parallel storage, east-west and north-south networking, power, cooling, and facility integration. •Define reference architectures, bill-of-materials templates, and repeatable deal patterns that accelerate presales velocity and improve design quality. •Lead architectural validation, risk assessment, and commercial shaping for strategic pursuits. Qualifications: •Bachelor's degree in Computer Science, Electrical Engineering, or a related technical field; advanced degree preferred. •Demonstrated delivery experience on multi-rack, GPU-dense factory deployments using at least two of the following platform families: NVIDIA DGX or HGX, Cisco AI infrastructure, Dell PowerEdge XE, Super Micro, Lenovo ThinkSystem, HPE ProLiant Compute DL or Cray. •Deep technical fluency across GPU compute, high-performance networking (InfiniBand, RoCE, NVIDIA Spectrum-X), parallel and scale-out file systems (VAST, WEKA, DDN, Pure, NetApp ONTAP AI), and AI-optimized cooling and power design. •Working command of the orchestration and operations stack including Kubernetes, Slurm, Run:ai, NVIDIA Base Command, Mission Control, and mainstream observability and MLOps tooling. •Operational familiarity with AI-specific security considerations including model protection, data governance, inference security, and alignment with emerging regulatory frameworks. •Executive presence with proven ability to engage CxO-level stakeholders at customers and OEM partners. •Strong track record in a VAR, OEM, hyperscaler, or large enterprise setting with multi-million dollar annual opportunity ownership. •Excellent written and verbal communication skills, including the ability to simplify complex AI infrastructure concepts for non-technical audiences. •Willingness to travel up to 50% to various locations around the United States as business needs require; valid driver’s license required. Compensation: •Comprehensive medical, dental and vision plans for you and your dependents Interested candidates, please apply directly through the job posting on company's career page or try via AI auto apply on this platform. Don't miss this opportunity to join a forward-thinking team!