Enter a job title or keyword

Senior Solutions Architect AI Infrastructure & Edge Computing ISV Partners , WWPS Global ISV Partners

Amazon


Job Location:

Arlington, TX - USA

Yearly Salary: USD 153600 - 207800
Posted: 13 September 2026 (14 hours ago)
Application Deadline: 11 December 2026
Vacancies: 1 Vacancy

Job Summary

Amazon Web Services (AWS) is seeking a Senior Solutions Architect to join the Worldwide Public Sector ISV Partner Solutions Architecture team focused on accelerated computing edge AI and the open weight model this role you will serve as the technical thought leader for strategic partners building GPU-accelerated AI infrastructure and deploying open foundation models at the cloud core the tactical edge and in sovereign/air-gapped environments where hosted APIs are not an option.

You will work directly with NVIDIA and adjacent ISV partners to architect full-stack AI compute solutions on AWS from multi-node GPU training clusters to lightweight inference at the disconnected edge. You will guide partners through model optimization deployment orchestration and infrastructure design patterns that enable mission-critical AI in defense intelligence community and national security contexts. This includes scoping sequencing and architecting solutions that create measurable mission value while driving clarity across internal AWS teams and external agency stakeholders.

This position requires that the candidate selected be a US Citizen and obtain and maintain an active TS/SCI security clearance.


Key job responsibilities
You will own the technical relationship with accelerated computing and edge AI partners adopting AWS from first whiteboard session to production at scale. Specifically you will:
Design and architect GPU-accelerated AI infrastructure on AWS including multi-node training clusters with EFA networking high-performance storage and container orchestration for large-scale model training and inference.
Guide partners in deploying open weight foundation models (Llama Mistral Falcon Nemotron Gemma to list a few) across AWS environments leveraging NVIDIA NIM microservices Triton Inference Server TensorRT-LLM and vLLM for optimized inference at cost and latency targets.
Architect edge and disconnected compute solutions for DDIL (denied degraded intermittent limited) environments enabling AI inference on NVIDIA Jetson IGX and embedded GPU platforms integrated with AWS hybrid services (Outposts ECS Anywhere).
Lead sovereign AI and air-gapped deployment architecture for IL4/IL5/IL6/SCIF environments where partners must run open models entirely on-premise or in isolated AWS regions without reliance on external API endpoints.
Drive model optimization and right-sizing engagements including quantization distillation pruning and adapter-based fine-tuning (LoRA/QLoRA) to help partners achieve production-grade performance within infrastructure constraints.
Serve as the embedded technical advisor for partner engineering teams conducting architecture reviews Well-Architected assessments and proof-of-concept builds that accelerate partner product roadmap delivery on AWS.
Build executive and working-level relationships with partner CTOs VP Engineering and mission-focused government stakeholders (DoD IC federal civilian) translating infrastructure capabilities into mission value.
Develop and publish reusable technical content: reference architectures blog posts deployment guides and benchmark reports for GPU workloads edge inference patterns and open model deployment on AWS.
Identify and drive AWS product feature requests (EC2 Bedrock Custom Model Import Amazon SageMaker AI AWS Neuron and AWS Nitro Enclaves) based on partner field signal to influence the service roadmap.
Establish repeatable benchmarks to evaluate model accuracy latency throughput GPU utilization reliability and cost across cloud and edge environments ensuring solutions meet mission and production requirements.
Travel approximately 30% of the time for partner site visits customer engagements and industry events.


A day in the life
Your morning starts with a design session alongside a partner engineering team whiteboarding a multi-node GPU training architecture on P5 instances with EFA networking and FSx for Lustre sizing the cluster for a Llama 3 70B fine-tuning workload that must complete within their sprint cycle.

Mid-morning you join a call with the partners edge deployment team to walk through an inference architecture for NVIDIA Jetson AGX Orin devices operating in a disconnected forward-deployed environment. You map out the model optimization pipeline: quantization via TensorRT-LLM container packaging and an OTA update mechanism that syncs model weights when connectivity is available through S3 and ECS Anywhere.

After lunch you draft a reference architecture showing how the partners threat detection platform can serve a Mistral 7B model via NVIDIA NIM behind an air-gapped EKS cluster in an IL5 environment with Nitro Enclaves handling sensitive data processing. You coordinate with AWS Networking and Security specialists to validate the VPC design and encryption-at-rest patterns.

Late afternoon you prep for a joint customer briefing with the partners CTO and a DoW program office. You build a technical deep-dive showing total cost of ownership: comparing on-demand GPU instances vs. reserved capacity vs. edge inference at the point of need with latency and throughput benchmarks from your recent proof-of-concept.

You close the day reviewing a product feature request youre submitting to the EC2 team based on partner feedback about GPU memory requirements for serving multiple concurrent open models and you update your SFDC activities to capture the weeks technical engagements.


About the team
You will join the WWPS ISV Partner Solutions Architecture team a group of senior technologists who serve as the embedded technical bridge between AWS and our most strategic independent software vendor partners. We help partners build modernize and scale their solutions on AWS to drive mission outcomes for public sector customers worldwide. Our team operates at the intersection of deep technical architecture partner engineering and go-to-market execution.

About the team
Diverse Experiences
AWS values diverse experiences. Even if you do not meet all of the preferred qualifications and skills listed in the job description we encourage candidates to apply. If your career is just starting hasnt followed a traditional path or includes alternative experiences dont let it stop you from applying.

Why AWS
Amazon Web Services (AWS) is the worlds most comprehensive and broadly adopted cloud platform. We pioneered cloud computing and never stopped innovating thats why customers from the most successful startups to Global 500 companies trust our robust suite of products and services to power their businesses.

Inclusive Team Culture
AWS values curiosity and connection. Our employee-led and company-sponsored affinity groups promote inclusion and empower our people to take pride in what makes us unique. Our inclusion events foster stronger more collaborative teams. Our continual innovation is fueled by the bold ideas fresh perspectives and passionate voices our teams bring to everything we do.

Mentorship & Career Growth
Were continuously raising our performance bar as we strive to become Earths Best Employer. Thats why youll find endless knowledge-sharing mentorship and other career-advancing resources here to help you develop into a better-rounded professional.

Work/Life Balance
We value work-life harmony. Achieving success at work should never come at the expense of sacrifices at home which is why we strive for flexibility as part of our working culture. When we feel supported in the workplace and at home theres nothing we cant achieve.


- 3 years of design implementation or consulting in applications and infrastructures experience
- Experience in developing and deploying LLMs in production on GPUs Neuron TPU or other AI acceleration hardware or experience with PyTorch JIT compilation and AOT tracing
- 8 years of specific technology domain areas (e.g. systems engineering infrastructure GPU/accelerated computing networking security cloud architecture) experience
- Experience with containerized and distributed computing architectures (Kubernetes ECS/EKS or equivalent)

- Experience working with end user or developer communities
- Experience with partner sales and/or alliance development in the software/technology industry
- Deep familiarity with the NVIDIA ecosystem: CUDA TensorRT Triton Inference Server NIM microservices Jetson/IGX edge platforms and NVIDIA AI Enterprise stack
- Hands-on experience deploying open weight models (Llama 2/3/4 Mistral Falcon Phi Gemma) with optimization frameworks (vLLM TensorRT-LLM quantization LoRA fine-tuning)
- Experience architecting for edge disconnected or DDIL environments with hybrid cloud integration (AWS Outposts ECS Anywhere)
- 5 years of infrastructure architecture including high-performance networking (InfiniBand RDMA EFA) distributed storage and cluster management
- Experience in the Defense or Intelligence Community sector with understanding of classification levels (IL4/5/6) and compliance frameworks (NIST 800-53 CMMC FedRAMP ITAR)

Amazon is an equal opportunity employer and does not discriminate on the basis of protected veteran status disability or other legally protected status.

Our inclusive culture empowers Amazonians to deliver the best results for our customers. If you have a disability and need a workplace accommodation or adjustment during the application and hiring process including support for the interview or onboarding process please visit for more information. If the country/region youre applying in isnt listed please contact your Recruiting Partner.

The base salary range for this position is listed below. Your Amazon package will include sign-on payments and restricted stock units (RSUs). Final compensation will be determined based on factors including experience qualifications and location. Amazon also offers comprehensive benefits including health insurance (medical dental vision prescription Basic Life & AD&D insurance and option for Supplemental life plans EAP Mental Health Support Medical Advice Line Flexible Spending Accounts Adoption and Surrogacy Reimbursement coverage) 401(k) matching paid time off and parental leave. Learn more about our benefits at FL Miami - 153600.00 - 207800.00 USD annually
USA NY New York - 169000.00 - 228600.00 USD annually
USA TX Austin - 153600.00 - 207800.00 USD annually
USA TX Dallas - 153600.00 - 207800.00 USD annually
USA TX Houston - 153600.00 - 207800.00 USD annually
USA VA Arlington - 153600.00 - 207800.00 USD annually
USA VA Herndon - 153600.00 - 207800.00 USD annually


Required Experience:

Senior IC


About Company

Company Logo

Free shipping on millions of items. Get the best of Shopping and Entertainment with Prime. Enjoy low prices and great deals on the largest selection of everyday essentials and other products, including fashion, home, beauty, electronics, Alexa Devices, sporting goods, toys, automotive ... View more

View Profile View Profile