Brendan McKeagOctober 14, 2024How to Code Stable Diffusion Directly in Python on RunpodSkip the front ends, learn how to use Jupyter Notebook on Runpod to run Stable Diffusion directly in Python. Great for devs who want full control.AI WorkloadsAll
Brendan McKeagOctober 1, 2024Why LLMs Can't Spell 'Strawberry' And Other Odd Use CasesLarge language models can write poetry and solve logic puzzles, but fail at tasks like counting letters or doing math. Here's why, and what it tells us.Learn AIAll
Brendan McKeagSeptember 25, 2024Run GGUF Quantized Models Easily with KoboldCPP on RunpodLower VRAM usage and improve inference speed using GGUF quantized models in KoboldCPP with just a few environment variables.AI WorkloadsAll
Brendan McKeagSeptember 25, 2024How to Work with GGUF Quantizations in KoboldCPPGGUF quantizations make large language models faster and more efficient. This guide walks you through using KoboldCPP to load, run, and manage quantized.Learn AIAll
Brendan McKeagSeptember 20, 2024Introducing Better Forge: Spin Up Stable Diffusion Pods FasterBetter Forge is a new Runpod template that lets you launch Stable Diffusion pods in less time and with less hassle. Here's how it improves your workflow.AI InfrastructureAll
Brendan McKeagSeptember 18, 2024Run Very Large LLMs Securely with Runpod ServerlessDeploy large language models like LLaMA or Mixtral on Runpod Serverless with strong privacy controls and no infrastructure headaches. Here’s how.AI InfrastructureAll
Brendan McKeagSeptember 13, 2024Evaluate Multiple LLMs Simultaneously Using Ollama on RunpodUse Ollama to compare multiple LLMs side-by-side on a single GPU pod, perfect for fast, realistic model evaluation with shared prompts.AI WorkloadsAll