Staging environment
Hero image

Empowering GenAI on AWS EKS: Deploy, Experiment, and Scale Advanced AI Workflows

Unlock the power of AI with hands-on expertise in vLLM, JARK Satck, RayServe, BioNeMo, Kubernetes, Amazon EKS, and Terraform

This course is no longer available.

Explore other courses

Previously at

Spryker Systems
ING
Deutsche Bank
Renault Group
Autohero Deutschland

Course overview

Deploy, Train & Scale GenAI on AWS EKS: Master Production-Grade AI

Unlock the power of AI with hands-on expertise in Kubernetes, Amazon EKS, Terraform, vLLM, JARK Satck, RayServe & BioNeMo.


Why This Course?


The ability to build, deploy, and scale AI applications efficiently is no longer a luxuryโ€”itโ€™s a necessity.


AI engineers, DevOps professionals, and software developers who can bridge the gap between ML and scalable infrastructure are in high demand.


This course will equip you with end-to-end, production-grade AI deployment skills using Amazon EKS, Kubernetes, Terraform, JARK Stack (Jypyter, Argo Workflows, Ray, Kubernetes), GPU, vLLM, Grafana, Prometheus and BioNeMo, and moreโ€”the same technologies used by leading AI-driven organizations.


You'll learn not just theory, but practical, real-world workflows for LLM deployment, distributed training, GPU optimization, and MLOps automationโ€”ensuring you're ready to build and scale AI solutions at any organization.


๐ŸŽฏ Course Milestones:


You'll begin by mastering the fundamentals of Generative AI and ML workflows, and understanding how Kubernetes enables large-scale model deployment, focusing on Amazon EKS for enterprise-ready solutions.


Next, youโ€™ll provision GPU-enabled EKS clusters, configure auto-scaling using Karpenter and Terraform, and optimize networking and IAM roles, ensuring a robust and automated infrastructure for AI workloads.


With the foundation set, you'll dive into the JARK stack (JupyterHub, Argo Workflows, Ray, and Kubernetes) for seamless model experimentation, workflow automation, and scaling distributed workloads.


From experimentation, youโ€™ll transition into scalable inference, leveraging RayServe and vLLM to deploy high-performance LLMs, optimize GPU memory, and integrate Open WebUI for real-time interactions.


To complete the ML and GenAI pipeline, you'll tackle distributed model training with BioNeMo, handling large datasets, multi-GPU orchestration, and performance tuning for efficient large-scale learning.


By the final stage, youโ€™ll be ready for your graduation projectโ€”building a real-world GenAI-powered chat application, integrating the entire ML workflow from training to inference while implementing CI/CD and monitoring strategies.


Through hands-on labs and project-based learning, you'll gain the practical expertise to deploy enterprise AI solutions, advance your career, and step into leadership roles in cloud-native AI engineering.


๐ŸŒปWhat Youโ€™ll Gain


This course transforms your career by giving you the hands-on experience and expertise needed to operate at the intersection of AI, ML, DevOps, and cloud-native architectures.


๐Ÿš€ By the end of this course, you will:


โœ… Deploy AI Models at Scale โ€“ Learn Kubernetes and Amazon EKS to run AI workloads on GPU-powered clusters with high availability and efficiency.


โœ… Master AI Inference with RayServe โ€“ Scale your LLM deployments seamlessly with RayServe and vLLM, ensuring fast and cost-effective inferencing.


โœ… Optimize AI Workloads โ€“ Use BioNeMo and distributed training techniques to maximize GPU utilization and optimize ML model training at scale.


โœ… Hands-on Real-World Projects โ€“ Work on a full-fledged GenAI chat application, gaining practical expertise that applies directly to industry use cases.


โœ… Career Acceleration โ€“ Position yourself as an AI/ML platform engineer, MLOps expert, or cloud-native AI specialist, opening doors to high-paying roles in top-tier companies.


๐ŸŽ“ Why Learn from This Course?


Most AI courses focus on model training but neglect the critical aspects of deploying, scaling, and operationalizing AI systems. This course bridges the gap.


You wonโ€™t just learn how to train AI modelsโ€”youโ€™ll learn how to deploy, optimize, and maintain them at scale in real-world environments.


โœ”๏ธ Designed for professionals who want to apply DevOps & Kubernetes to ML


โœ”๏ธ Project-based learning ensures practical expertise


โœ”๏ธ Covers the entire AI lifecycle from training to production deployment


โœ”๏ธ Instructor-led deep dives on critical AI deployment patterns


This is the only course that provides a complete skill set in AI, ML, DevOps, and cloud-native architecturesโ€”ensuring you stay ahead of the curve in one of the most sought-after fields in tech today.

DevOps, AI & ML, Software Engineers, and Entry-Level Engineers, this course is for you!

01

DevOps professionals looking to apply their cloud, Kubernetes, and automation skills to AI/ML workloads on AWS.

02

AI/ML engineers who want to learn Kubernetes, cloud automation, and scalable model deployment using DevOps best practices.

03

Software Engineers Expanding into AI, ML & MLOps: Software developers who want to gain expertise in AI/ML model training, deployment, and ML

Mastering AI Engineering Using Cloud-Native Technologies and DevOps Best Practices

Setting Up Strong Foundations

Mastering the fundamentals of Generative AI and ML workflows, such as model training and model inference, while understanding the technologies that enable large-scale model deployment, focusing on Amazon Elastic Kubernetes Service for enterprise readiness.solutions.

Provision and configure GPU-enabled AWS EKS clusters using Terraform, with security, IAM, and networking best practices.

Give students an idea of how they can expect to grow throughout your course. Include specificity and precise results so students can benchmark exactly what theyโ€™ll learn.

Master Kubernetes basics: pods, deployments, autoscaling, resource limits, and troubleshooting logs for ML workloads.

Give students an idea of how they can expect to grow throughout your course. Include specificity and precise results so students can benchmark exactly what theyโ€™ll learn.

Deploy the JARK stack with JupyterHub, Argo Workflows, and Ray to streamline collaborative ML experiments.

Give students an idea of how they can expect to grow throughout your course. Include specificity and precise results so students can benchmark exactly what theyโ€™ll learn.

Launch inference endpoints using Docker, RayServe, vLLM, and HuggingFace libraries on AWS EKS for AI chat apps.

Give students an idea of how they can expect to grow throughout your course. Include specificity and precise results so students can benchmark exactly what theyโ€™ll learn.

Analyze GPU memory usage and optimize model inference performance using real-time metrics and port-forwarding.

Execute distributed training with BioNeMo and Kubeflow, leveraging multi-GPU setups and FSx for Lustre storage.

Learn Kubernetes and AWS EKS security and operation excellence best practices.

Build a complete GenAI chat app from experimentation to training and inference with hands-on labs and milestones.

Utilize the standard DevOps and Cloud Native tools to build AI/ML platforms, including Terraform, Kubernetes, Grafana, Karpenter, and Helm.

Whatโ€™s included

Aymen Segni

Live sessions

Learn directly from Aymen Segni in a real-time, interactive format.

Lifetime access

Go back to course content and recordings whenever you need to.

Community of peers

Stay accountable and share insights with like-minded professionals.

Certificate of completion

Share your new skills with your employer or on LinkedIn.

Maven Guarantee

This course is backed by the Maven Guarantee. Students are eligible for a full refund up until the halfway point of the course.

Course syllabus

8 live sessions โ€ข 7 lessons โ€ข 7 projects

Week 1

Sep 1โ€”Sep 7

    ๐Ÿ“… Week 1: April 2 โ€“ April 6 | Class I โ€“ Foundations & EKS Deployment

    No module content yet

    Module I.1 โ€“Introduction to GenIA, ML and Kubernetes

    • ๐Ÿ“„

      ๐Ÿ“– Lesson: Understanding AI Workloads in Kubernetes

    • Apr

      3

      ๐Ÿ“Œ Live Event I.1: April 3 โ€“ Foundations of GenAI, ML & Kubernetes

      Thu 4/34:00 PMโ€”5:15 PM (UTC)
    • โœ๏ธ

      ๐Ÿ›  Project: Diagram Your AI Infrastructure on AWS EKS

      Submit by Jul 19

    ๐Ÿ“Œ Module I.2 โ€“ Deploying a Production-Ready AWS EKS Cluster

    • ๐Ÿ“„

      Deploy an EKS cluster with GPU support and validate Kubernetes fundamentals.

    • โœ๏ธ

      Deploy an EKS cluster with GPU support and validate Kubernetes fundamentals.

      Submit by Jul 14
    • Apr

      4

      ๐Ÿ“Œ Live Event: April 4 โ€“ Hands-On Lab: Deploy AWS EKS with GPU Support

      Fri 4/410:00 AMโ€”11:30 AM (UTC)

Week 2

Sep 8โ€”Sep 14

    ๐Ÿ“… Week 2: April 7 โ€“ April 13 | Class II โ€“ Model Experimentation & Inference

    No module content yet

    Module II.1 โ€“ GenAI and LLM Models Experiments with the JRAK Stack

    • ๐Ÿ“„

      ๐Ÿ“– Lesson: Setting Up JRAK (JupyterHub, Ray, Anyscale, Kubernetes)

    • โœ๏ธ

      ๐Ÿ›  Project: Train & Benchmark an LLM Using JRAK

      Submit by Jul 14
    • Apr

      8

      ๐Ÿ“Œ Live Event: April 8 โ€“ Hands-On Lab: LLM Experimentation on AWS EKS

      Tue 4/84:00 PMโ€”5:30 PM (UTC)

    Module II.1 โ€“ Model Inference Using RayServe and vLLM

    • ๐Ÿ“„

      ๐Ÿ“– Lesson: Deploying Scalable Model Inference Endpoints

    • โœ๏ธ

      ๐Ÿ›  Project: Deploy an Inference Endpoint with RayServe & vLLM

      Submit by Jul 14
    • Apr

      10

      ๐Ÿ“Œ Live Event: April 10 โ€“ Hands-On Lab: Deploying an LLM Inference Endpoint

      Thu 4/104:00 PMโ€”5:30 PM (UTC)

    Module II.3 โ€“ Distributed Model Training with BioNeMo

    • ๐Ÿ“„

      ๐Ÿ“– Lesson: Training Large-Scale AI Models Using NVIDIA BioNeMo

    • โœ๏ธ

      ๐Ÿ›  Project: Train a Multi-Billion Parameter Model Using BioNeMo

      Submit by Jul 14
    • Apr

      12

      ๐Ÿ“Œ Live Event: April 12 โ€“ Advanced AI Training with BioNeMo

      Sat 4/121:00 AMโ€”2:30 AM (UTC)

Week 3

Sep 15

    Week 3: April 14 โ€“ April 16 | Class III โ€“ Graduation Project

    • ๐Ÿ“„

      ๐Ÿ“– Lesson: Architecting an End-to-End AI Application

    • โœ๏ธ

      ๐Ÿ›  Project: Develop & Deploy a GenAI Chatbot on AWS EKS

      Submit by Jul 19
    • Apr

      16

      ๐Ÿ“Œ Live Event: April 16 โ€“ Graduation Project Showcase & Q&A

      Wed 4/164:00 PMโ€”5:30 PM (UTC)

    Graduation Project โ€“ Building a Real-World GenAI Application

    • ๐Ÿ“„

      Graduation Project

    • โœ๏ธ

      Building a Real-World GenAI Application

      Submit by Jul 14

    Apr

    16

    Graduation Project: Develop a complete GenAI chat application

    Wed 4/169:00 PMโ€”10:15 PM (UTC)

Bonus

    ๐Ÿ”ฅ Bonus Module: Deploying DeepSeek R1 on AWS EKS

    • Feb

      18

      ๐Ÿ”ฅ Live Event Bonus Module - Apr. 18: Deploying DeepSeek R1 on AWS EKS

      Tue 2/185:00 PMโ€”6:30 PM (UTC)

    No module content yet

Meet your instructor

Aymen Segni

Aymen Segni

Cloud & DevOps Expert specializing in modern infrastructure with Kubernetes

Aymen Segni is a Cloud & DevOps Expert, with deep expertise in Cloud Infrastructure, SRE, Data Platforms, Kubernetes, and Cloud-Native architectures. With a background in software engineering and systems architecture,


Aymen has helped startups and enterprises design, deploy, and scale their infrastructure and applications on platforms like AWS and Kubernetes.

His career spans Cloud Platform engineering, DevOps, and AI/ML operations, making him uniquely positioned to bridge the gap between AI model development and scalable, production-grade deployments. Aymen has worked with cloud providers, FineTech industry, software vendors, e-commerce, and consulting, leading teams to optimize technologies to make the business run better.


As a mentor and educator, Aymen has trained engineers in Kubernetes, cloud automation, and AI/ML deployment strategies, helping professionals transition into AI-driven DevOps roles. This course distills years of hands-on experience into a practical, project-based learning journey, ensuring students gain real-world skills for high-impact AI engineering careers.

Learning is better with cohorts

Learning is better with cohorts

Active hands-on learning

This course builds on live workshops and hands-on projects

Interactive and project-based

Youโ€™ll be interacting with other learners through breakout rooms and project teams

Learn with a cohort of peers

Join a community of like-minded people who want to learn and grow alongside you

Frequently Asked Questions