inhousefyi
← Back to listings

AI Infrastructure Solutions Architect

SUBMERRemote · Posted 24 days ago
Full-timeRemote
Apply now

Description

Location & work modality: India ( Remote)

About Submer

Submer enables organizations scaling AI to overcome the limits of traditional datacenters across power, compute density and efficiency.

We design and deliver scalable, high-density AI datacenter infrastructure built around industry-leading liquid cooling, supporting everything from early AI deployments to full-scale production environments.

What Impact you will have

We are looking for AI Infrastructure Solutions Architect

Department: Solutions Engineering / Presales

Position Summary

The AI Infrastructure Solutions Architect is responsible for designing end-to-end AI infrastructure solutions that enable customers to deploy scalable, high-performance AI and HPC environments. The role combines technical presales, solution architecture, infrastructure sizing, and deployment planning to deliver commercially viable and operationally sound AI infrastructure solutions.

The architect serves as the technical authority for compute, networking, storage, AI software stack, and infrastructure integration while working closely with Sales, Delivery, Product Management, OEM partners, and the Facilities Presales & Design teams to ensure that the AI infrastructure requirements are fully aligned with the data center's power, cooling, space, and physical infrastructure capabilities.

This role owns the AI infrastructure architecture and complements the facility design process.

What you will do

Key Responsibilities

  1. AI Infrastructure Presales
  • Partner with Sales to qualify AI infrastructure opportunities.
  • Engage with customers to understand AI workloads, performance expectations, scalability requirements, and operational objectives.
  • Conduct technical discovery workshops and architecture discussions.
  • Prepare technical proposals, presentations, and solution demonstrations.
  • Support RFI, RFP, and RFQ responses.
  • Serve as the trusted technical advisor throughout the sales lifecycle.

  1. AI Infrastructure Solution Architecture

Design complete AI infrastructure solutions including:

  • GPU compute clusters
  • AI Factory infrastructure
  • HPC platforms
  • High-performance storage
  • AI networking fabrics (Ethernet, InfiniBand, RoCE)
  • Kubernetes and container platforms
  • AI software stack integration
  • Cluster management
  • Security architecture
  • Monitoring and observability
  • Automation platforms
  • Scalability and expansion planning

Develop:

  • Solution architecture
  • Infrastructure sizing
  • Bill of Materials (BoM)
  • Reference architectures
  • Technical specifications
  • Solution documentation
  • Deployment architecture

  1. Infrastructure Integration with Data Center Design

Work closely with Facilities Presales, Design, and Engineering teams to ensure the AI infrastructure solution is fully compatible with the proposed data center environment.

Provide technical inputs related to:

  • Rack power density
  • Rack layouts
  • GPU server deployment
  • Network topology
  • Cable density
  • Space requirements
  • Liquid cooling interfaces
  • CDU connectivity requirements
  • Infrastructure dependencies
  • Expansion strategy

Validate that the AI infrastructure can be successfully deployed within the facility constraints without owning the facility design itself.

  1. Solution Validation & Delivery Readiness

Collaborate with Delivery and Engineering teams to ensure solution feasibility.

Support:

  • Technical design reviews
  • OEM interoperability validation
  • Deployment planning
  • Installation readiness
  • Factory acceptance planning
  • Site acceptance planning
  • Commissioning support
  • Technical handover to delivery teams

Maintain ownership of the AI infrastructure solution throughout the project lifecycle.

  1. Operational Architecture

Ensure the proposed solution is designed for long-term operational success by considering:

  • Serviceability
  • Scalability
  • High availability
  • Redundancy
  • Cluster management
  • Monitoring
  • Capacity management
  • Firmware lifecycle
  • Upgrade strategy
  • Infrastructure observability
  • Operational support requirements

  1. Cross-Functional Collaboration

Work closely with:

  • Sales
  • Product Management
  • Delivery
  • Professional Services
  • Facilities Presales
  • Facilities Design & Engineering
  • OEM partners
  • System Integrators
  • Customer Infrastructure teams
  • AI Engineering teams

Act as the primary technical interface for all AI infrastructure-related discussions.

Technical Expertise

AI Compute

  • NVIDIA GPU platforms
  • AMD GPU platforms
  • Intel AI platforms
  • GPU cluster design
  • AI Factory architectures
  • HPC infrastructure

Networking

  • Ethernet (100/200/400/800G)
  • InfiniBand
  • RoCE
  • Spine-Leaf architectures
  • AI fabric design

Storage

  • Parallel file systems
  • High-performance NAS
  • Object storage
  • NVMe-over-Fabrics
  • AI data pipelines

Software

  • Kubernetes
  • Docker
  • Slurm
  • NVIDIA AI Enterprise
  • GPU scheduling
  • Cluster management
  • Infrastructure automation

Infrastructure

  • Server platforms
  • Rack integration
  • Infrastructure sizing
  • High-density deployments
  • Liquid-cooled server technologies
  • Infrastructure monitoring
  • DCIM integration

Qualifications

  • Bachelor's degree in Computer Science, Information Technology, Electronics, Electrical Engineering, or a related discipline.
  • Master's degree is preferred.
  • 8–15 years of experience in enterprise infrastructure, HPC, AI infrastructure, or solution architecture.
  • At least 5 years in a customer-facing Presales or Solutions Architecture role.
  • Experience designing GPU-based AI infrastructure and high-density compute environments is highly desirable.

Preferred Certifications

  • NVIDIA Certified Professional (or equivalent)
  • AWS Solutions Architect
  • Microsoft Azure Solutions Architect
  • Red Hat OpenShift
  • VMware VCP
  • Cisco CCNP/CCIE
  • Kubernetes (CKA/CKAD)
  • ITIL Foundation

Core Competencies

Technical

  • AI Infrastructure Architecture
  • GPU Compute Platforms
  • HPC Infrastructure
  • Infrastructure Sizing
  • Solution Design
  • AI Networking
  • Storage Architecture
  • Kubernetes
  • AI Platform Integration

Business

  • Technical Presales
  • Solution Consulting
  • Proposal Development
  • Technical Bid Management
  • Customer Engagement
  • Executive Presentations

Collaboration

  • Cross-functional leadership
  • Stakeholder management
  • Technical mentoring
  • OEM engagement
  • Customer relationship management

What we offer

  • Attractive compensation package reflecting your expertise and experience.
  • A great work environment characterized by friendliness, international diversity, flexibility, and a hybrid-friendly approach.
  • You´ll be part of a fast-growing scale-up with a mission to make a positive impact, offering an exciting career evolution.

Our job titles may span more than one job level. The actual base pay is dependent on a number of factors, such as transferable skills, work experience, business needs and market demands.

Our inclusive responsibility

Submer is committed to creating a diverse and inclusive environment and is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, gender, gender identity or expression, sexual orientation, national origin, genetics, disability, age, veteran status, or any other protected category under applicable law.

Similar jobs

SUBMERHouston, Texas, United States

Est. 125,000 USD

Technical Program Manager — Data Center Delivery (Houston, TX - USA) Location: Houston, TX Type of Contract: Full Time / Permanent Travel Requirements: 50% About Submer Submer designs and delivers end-to-end AI datacente…

Full-time
SUBMERHouston, Texas, United States

Est. 115,000 USD

About Submer Submer designs and delivers end-to-end AI datacenter infrastructure built around industry-leading liquid cooling. We help organizations scale AI beyond the limits of traditional datacenters by enabling highe…

Full-time
SUBMERHouston, Texas, United States

Est. 140,000 USD

**Unfortunately, we are not able to provide work authorization in the US for this position. We are seeking a Senior AI Data Center Electrical Engineer to lead the electrical power design, implementation, and optimization…

Full-time
SUBMERHouston, Texas, United States

Est. 120,000 USD

About Submer Submer designs and delivers end-to-end AI datacenter infrastructure built around industry-leading liquid cooling. We help organizations scale AI beyond the limits of traditional datacenters by enabling highe…

Full-time
SUBMERHouston, Texas, United States

Est. 115,000 USD

Finance Manager – Modular Data Center Infrastructure Location: Houston, Texas (Hybrid or Open to candidates willing to travel)Travel: Up to 20% About Submer Submer designs and delivers end-to-end AI datacenter infrastruc…

Full-time
Firmus TechnologiesSan Francisco, California, United States

Est. 165,000 USD

Firmus Technologies Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure from model to grid. Founded in Australia in 2019, our mission is to create the most effic…

Full-time
SUBMERMumbai, India

Location & work modality: Mumbai (Hybrid) About Submer Submer enables organizations scaling AI to overcome the limits of traditional datacenters across power, compute density and efficiency. We design and deliver sca…

Full-time
SUBMERRubí, Catalonia, Spain

Est. 65,000 EUR

Location: HQ - Rubí (Barcelona) Start: ASAP Type of Contract: Full Time / Permanent About Submer Submer designs and delivers end-to-end AI datacenter infrastructure built around industry-leading liquid cooling. We help o…

Full-time
SUBMERTaipei, Taiwan

Location & work modality: Taipei, Taiwan (Hybrid) Start: ASAP Type of Contract: Full Time / Permanent About Submer Submer enables organizations scaling AI to overcome the limits of traditional datacenters across powe…

Full-time
SUBMERHouston, Texas, United States

Est. 124,000 USD

Title: Senior Test & Validation Engineer Location: Houston, USA [Hybrid] Type of Role: Permanent / Full time Level of Experience: Mid - Senior (5+ years) Travel Requirements: 10% About Submer At Submer, we’re committ…

Full-time
NebiusRemote

Est. 285,000 USD

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
Firmus TechnologiesSydney, New South Wales, Australia

Firmus Technologies Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure across Asia Pacific. Founded in Australia in 2019, our mission is to create the most effi…

Full-time
Firmus TechnologiesSingapore, Singapore

Firmus Technology Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure across Asia Pacific. Founded in Australia in 2019, our mission is to create the most effici…

Full-time
NebiusRemote

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
NebiusRemote

Est. 234,900 USD

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
ArmadaRemote

About the Company Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company's Most Innovative Companies…

Full-timeRemote
Together AISan Francisco, California, United States

Est. 230,000 USD

About The Role Together AI is building its infrastructure footprint at scale, and this role is central to making that happen. As an Infrastructure Design Engineer, you will own the design, planning, and technical executi…

Full-time
NebiusRemote

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
NebiusRemote

Est. 120,000 EUR

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
ArmadaRemote

Est. 110,000 USD

About the Company Armada is the hyperscaler for the edge, delivering modular AI infrastructure from first deployment to AI factory with speed, scale and sovereignty. Named one of Fast Company's Most Innovative Companies…

Full-timeRemote
SUBMERMumbai, India

Location & work modality: India , Mumbai About SubmerAt Submer, we are redefining how data centers are built, integrated, and operated, with sustainability, efficiency, and innovation at the core. Our technology is d…

Full-time
DigitalOceanRemote

Est. 242,500 USD

Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, natura…

Full-timeRemote
NebiusRemote

Est. 155,000 USD

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
Myriad360Remote

Est. 170,000 USD

Who You Are You are a seasoned Senior Presales Solutions Architect with deep data center, network infrastructure, and high-performance computing experience. You are comfortable leading customer discovery, reviewing techn…

Full-timeRemote
Scale AILondon, United Kingdom

Est. 120,000 GBP

At Scale AI, our mission is to accelerate the development of AI applications. For 10 years, Scale has been the leading AI data foundry, helping fuel the most exciting advancements in AI, including: generative AI, defense…

Full-time
NebiusRemote

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
DigitalOceanBengaluru, India

Dive in and do the best work of your career at DigitalOcean. Journey alongside a strong community of top talent who are relentless in their drive to build the simplest scalable cloud. If you have a growth mindset, natura…

Full-time
Astera LabsResearch Triangle Park, North Carolina, United States

Est. 222,500 USD

Astera Labs (NASDAQ: ALAB) provides rack-scale AI infrastructure through purpose-built connectivity solutions. By collaborating with hyperscalers and ecosystem partners, Astera Labs enables organizations to unlock the fu…

Full-time
NebiusRemote

About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to…

Full-timeRemote
BiohubRedwood City, California, United States

Est. 286,000 USD

Biohub is the first large-scale initiative bringing frontier AI models, massive compute, and frontier experimental capabilities under one roof. We're building a general-purpose system to accelerate scientific discovery,…

Full-time