Introduction
Introduction
()
1. Understanding Generative AI, LLMs, and Foundation Models
Understanding generative AI
()
Understanding large language models
()
Foundation models: The core of generative AI
()
2. GPUs and Scalable AI Compute Systems
How generative AI solves enterprise challenges
()
Why GPUs power generative AI
()
Inside the GPU: The AI factory
()
The power of parallel processing
()
CPU vs GPU: The right tool for the job
()
GPU server systems scaling beyond single GPUs
()
3. AI Software, Virtualization, and Data Centre Foundations
GPUs: The power grid for AI
()
Virtual GPU sharing: power maximizing efficiency
()
GPU virtualization intelligent fleet management for AI
()
Machine learning and deep learning frameworks
()
The deep learning software stack
()
Where does AI live
()
The three pillars of AI data centres
()
4. AI Data Centre Architecture and Networking
Managing and monitoring an AI data centre
()
The modern data centre platform
()
Multi GPU systems powering telecom AI
()
The four networks of an AI data centre
()
Networking for AI workloads
()
InfiniBand: The supersonic rail system for AI
()
The silent crisis of AI infrastructure
()
Storage file systems for AI data centres
()
5. Performance and Cloud AI Infrastructure
Storage performance for AI: The championship pit crew
()
Energy efficiency in AI data centres
()
Why GPU cooling architecture matters
()
Reference architectures and AI in the cloud
()
AI in the cloud
()
6. Cloud AI Deployment and Telecom Use Cases
Conclusion
()
Deploying AI in the cloud: A strategic blueprint
()