Shaamil KarimJune 22, 2025Run vLLM on Runpod Serverless: Deploy Open Source LLMs in MinutesLearn when to use open source vs. closed source LLMs, and how to deploy models like Llama 3 or Qwen3 with vLLM on Runpod Serverless for high-throughput inference.AI WorkloadsAll
Shaamil KarimAugust 22, 2024Deploy Google Gemma 7B with vLLM on Runpod ServerlessDeploy Google’s Gemma 7B model using vLLM on Runpod Serverless in just minutes. Learn how to optimize for speed, scalability, and cost-effective AI inference.AI WorkloadsAll
Shaamil KarimAugust 20, 2024Run Llama 3.1 with vLLM on Runpod ServerlessDiscover how to deploy Meta's Llama 3.1 using Runpod's new vLLM worker. This guide walks you through model setup, performance benefits, and step-by-step.AI InfrastructureAll
Shaamil KarimAugust 13, 2024Run Flux Image Generator in ComfyUI on Runpod (Step-by-Step Guide)Learn how to deploy and run Black Forest Labs' Flux 1 Dev model using ComfyUI on Runpod. This step-by-step guide walks through setting up your GPU pod.AI WorkloadsAll
Shaamil KarimAugust 13, 2024How to Run the FLUX Image Generator with ComfyUI on RunpodStep-by-step guide for deploying FLUX with ComfyUI on Runpod. Perfect for creators looking to generate high-quality AI images with ease.Learn AIAll
Shaamil KarimAugust 8, 2024Run the Flux Image Generator on Runpod (Full Setup Guide)This guide walks you through deploying the Flux image generator on a GPU using Runpod. Learn how to clone the repo, configure your environment, and start.AI WorkloadsAll
Shaamil KarimAugust 2, 2024Run SAM 2 on a Cloud GPU with Runpod (Step-by-Step Guide)Learn how to deploy Meta's Segment Anything Model 2 (SAM 2) on a Runpod GPU using Jupyter Lab. This guide walks through installing dependencies.AI WorkloadsAll