Senior Solutions Architect
About the Company
Firmus Technologies is a global leader pioneering the development and operation of efficient AI infrastructure from model to grid. Founded in Australia in 2019, our mission is to create the most efficient AI infrastructure by combining cutting-edge technology with a steadfast commitment to sustainability. We design, build, and operate a new class of digital infrastructure—the AI Factory—pushing the boundaries of multi-generational liquid cooling systems, energy management, AI software orchestration, and construction. Our model-to-grid approach ensures every watt counts, delivering low-cost AI tokens globally.
Our large-scale GPU cloud platform, Firmus AI Cloud, is purpose-built to deliver energy-efficient AI compute at scale. It empowers developers, enterprises, educational institutions, and government users to train and deploy AI models with unmatched efficiency and cost savings.
Why You’ll Love Working Here
At Firmus, you’ll work at the intersection of sustainability and artificial intelligence in a fast-paced environment powered by next-generation technology. You’ll help transform an entire industry and see the impact of your work first-hand as we democratize AI tools more sustainably and affordably. Our team consists of true innovators and leaders in their fields, offering opportunities to work closely with founders, build a strong network, and collaborate with diverse, authentic individuals. We are proud to be an equal opportunity employer.
About the Role
As an AI Infrastructure Solutions Architect at Firmus, you will assist customers in designing and implementing advanced AI cluster solutions. This role combines solution architecture skills with AI infrastructure expertise to support organizations’ AI initiatives. You will serve as the crucial link between our customers and product teams as we scale the Firmus AI Factory offering globally.
Responsibilities
- Solution Design
- Architect scalable and high-performance AI infrastructure solutions tailored to customer needs.
- Work with internal Firmus Engineering and Technology teams to develop detailed technical specifications and implementation plans for AI clusters.
- Optimize designs for performance, reliability, and cost-effectiveness.
- Collaborate closely with global sales teams, providing deep technical expertise during the sales cycle to address customer requirements and challenges.
- Identify relevant technology capability gaps in the platform that impact service delivery.
- Customer Engagement
- Collaborate closely with customers to understand their AI requirements and objectives.
- Support the Firmus Sales Function to ensure successful closure of RFPs, Proposals, and service contracts.
- Develop effective SLAs in collaboration with Sales, Operations, and Engineering teams to ensure successful service delivery.
- Provide expert guidance on AI infrastructure best practices and emerging technologies.
- Present complex technical concepts to both technical and non-technical stakeholders.
- Lead AI infrastructure projects from conceptualization to implementation.
- Assist with the successful deployment and integration of AI solutions, guiding customers through the implementation process and optimizing their AI infrastructure for peak performance.
- Coordinate with cross-functional teams to ensure seamless integration of AI solutions.
- Work with and present to C-level stakeholders.
- Technical Expertise
- Understand training and inference requirements of customers and their use cases.
- Provide technical mentorship to junior team members.
- Maintain in-depth knowledge of advanced networking technologies, including NVIDIA InfiniBand, Spectrum Ethernet Platform, and RDMA over Converged Ethernet (RoCE).
- Work with AI accelerator hardware, software, and infrastructure components (NVIDIA CPU-GPU architectures, AMD Instinct-based accelerators, Intel Gaudi, among others).
- Understand and leverage AI-specific technologies such as NVIDIA ConnectX NICs and BlueField DPUs.
- Stay at the forefront of AI technology trends, advocating for our AI solutions and educating customers and partners on their benefits and applications.
Requirements
- Bachelor’s Degree in Computer Science, Engineering, or a related field.
- 8+ years of working experience in IT infrastructure, with at least 3+ years focused on AI or high-performance computing systems.
- Extensive experience in designing and implementing AI infrastructure solutions.
- Deep understanding of high-performance computing and AI workloads.
- Proficiency in architecting solutions using NVIDIA InfiniBand, Spectrum Ethernet Platform, and related technologies.
- Proficiency in programming languages relevant to AI and ML, such as Python, C++, and familiarity with AI frameworks like TensorFlow and PyTorch.
- Solid experience with cloud architectures and services, particularly in GPU-accelerated computing environments.
- Strong problem-solving skills and ability to troubleshoot complex infrastructure issues.
- Demonstrated ability to design scalable and sustainable AI solutions, with a track record of successful customer engagements, project deliveries, and implementations.
Location & Reporting
This role is based in the Bay Area, California, USA, and reports to the CTO in the interim. Employment basis is full-time.
Diversity
At Firmus, we are committed to building a diverse and inclusive workplace. We encourage applications from candidates of all backgrounds who are passionate about creating a more sustainable future through innovative engineering solutions.