Tag: Azure
-
Azure OpenAI Deployment Types, Standard vs PTU vs Batch (Azure Gen AI Series, Part 6)
Standard, provisioned throughput units, and Batch are the three ways to bill an Azure OpenAI deployment. How to pick with a utilization break-even, size PTUs, and use spillover so you are not throttled or overpaying.
-
Azure OpenAI vs Foundry vs Azure Machine Learning, and Which One Your Project Belongs In (Azure Gen AI Series, Part 5)
Azure OpenAI, Microsoft Foundry, and Azure Machine Learning look interchangeable and are not. A practical guide to which resource your project belongs in, and how to move between them without a rebuild.
-
Azure Foundry Model Catalog and Models-as-a-Service (Azure Gen AI Series, Part 4)
The Foundry catalog splits into models sold by Azure and partner models, and every one deploys as serverless Models-as-a-Service or as managed compute. Here is how to pick, with the permission and rate-limit traps that bite first.
-
Azure OpenAI in Foundry Models, and How to Choose One (Azure Gen AI Series, Part 3)
Azure OpenAI in Foundry Models is the set of OpenAI models you deploy inside a Foundry resource. Here is which models you get in 2026, how to pick one, and how the call goes out with the v1 API.
-
Azure AI Foundry Platform and Foundry Projects Explained (Azure Gen AI Series, Part 2)
Microsoft renamed Azure AI Foundry and rebuilt it around one resource, projects, and a single SDK. Here is what the platform is, how projects work, and your first API call against it.
-
Azure Generative AI Stack, End to End (Azure Gen AI Series, Part 1)
The whole Azure generative AI stack in one map: Azure AI Foundry, the model catalog, the three deployment types that decide your bill, retrieval, agents, and the compute underneath. Where to start and what to skip.
-
How to Land Your First Cloud Job: Resume, Certs and Interviews (Cloud for Beginners, Part 18)
A practical fresher playbook for getting your first cloud job in 2026: which certification to pick, how to build a resume project, and how to answer the interview questions that actually decide the offer.
-
How to Build Cloud Skills for Free: Free Tiers and Labs (Cloud for Beginners, Part 17)
Learn real cloud skills for free in 2026 with the AWS, Azure and Google Cloud free tiers and free labs, and set up the budget alerts that keep you from a surprise bill.
-
Public, Private, Hybrid and Multicloud Explained (Cloud for Beginners, Part 16)
Public, private, hybrid and multicloud explained in plain English for freshers, with a real egress cost example and the interview question that always comes up.
-
Cloud Reliability, Backups and Disaster Recovery Explained (Cloud for Beginners, Part 15)
Reliability is a number, not a feeling. A plain-English guide to redundancy, backups, RTO and RPO, and the four disaster recovery strategies on AWS, Azure and Google Cloud.
-
How Cloud Billing Works (and How Bills Explode) (Cloud for Beginners, Part 13)
Cloud bills explode from quiet meters, not the obvious server. Here is how cloud billing really works, the charges that ambush freshers, and a real bill broken down line by line.
-
Serverless Explained: What It Is and When to Use It (Cloud for Beginners, Part 12)
Serverless explained for freshers: what it means, how Functions as a Service work, cold starts, a real cost breakdown, and when not to use it.
Architect’s Toolkit
PJ’s Tools
- Infra 360 Hub – All Series
- VCF 9 Interactive Walkthroughs
- VCF Design Cheatsheet
- VCF Upgrade Planner
- VCF 9 Series Hub
- VCF Deployment Hub
- AI Stack Hub
- AI Infra Sizing & Cost Calculator
- LLM & RAG Cost Calculator
- DrJhaGPT – Ask Pranay
VMware Cloud Foundation
- VCF Documentation
- VCF 9 Planning & Preparation Workbook
- VCF Bill of Materials (BoM)
- VMware Compatibility Guide
- VMware Interoperability Matrix
- VMware Configuration Maximums
- VMware Ports & Protocols
- VMware Hands-on Labs
- RVTools Download
Nutanix
AI & Cloud-Native Platform
- NVIDIA Build (Model Catalog)
- NVIDIA AI Enterprise Reference Architecture
- NVIDIA NIM Performance Benchmarking
- NVIDIA NGC Catalog
- NeMo Microservices Helm Chart
- Helm Charts Repository
- Hugging Face Models
Architecture & Design
About the Author

Dr Pranay Jha
Dr. Pranay Jha is a Cloud and AI Consultant with 18+ years of experience in hybrid cloud, virtualization, and enterprise infrastructure transformation. He specializes in VMware technologies, multi-cloud strategy, and Generative AI solutions. He holds a PhD in Computer Applications with research focused on Cloud and AI, has published multiple research papers, and has been a VMware vExpert since 2016 and a VMUG Community Leader.
You May Have Missed


DrJha