Brendan McKeagAugust 15, 2024Supercharge Your LLMs with SGLang: Boost Performance and CustomizationDiscover how to boost your LLM inference performance and customize responses using SGLang, an innovative framework for structured LLM workflows.AI WorkloadsAll
Brendan McKeagJuly 25, 2024Mastering Serverless Scaling on Runpod: Optimize Performance and Reduce CostsLearn how to optimize your serverless GPU deployment on Runpod to balance latency, performance, and cost. From active and flex workers to Flashboot and.AI InfrastructureAll
Brendan McKeagJune 6, 2024Run Larger LLMs on Runpod Serverless Than Ever Before – Llama-3 70B (and beyond!)Runpod Serverless now supports multi-GPU workers, enabling full-precision deployment of large models like Llama-3 70B. With optimized VLLM support.Product UpdatesAll
Brendan McKeagMay 28, 2024Introducing Serverless CPU: High-Performance VMs Without GPUsOur new Serverless CPU offering lets you launch high-performance containers without GPUs, perfect for lighter workloads, dev tasks, and automation.Product UpdatesAll
Brendan McKeagMay 28, 2024Announcing Runpod's New Serverless CPU FeatureRunpod introduces Serverless CPU: high-performance VM containers with customizable CPU options, ideal for cost-effective and versatile workloads not.Product UpdatesAll
Brendan McKeagApril 15, 2024Configurable Endpoints for Deploying Large Language ModelsDeploy any Hugging Face large language model using Runpod's configurable templates. Customize your endpoint with ease and launch scalable LLM deployments.Product UpdatesAll
Brendan McKeagMarch 27, 2024Generate Images with Stable Diffusion on RunpodLearn how to set up a Runpod project, launch a Stable Diffusion endpoint, and generate images from text using a simple Python script and the Runpod CLI.AI WorkloadsAll