Brendan McKeagAugust 1, 2025Deep Cogito Releases Suite of LLMs Trained with Iterative Policy ImprovementDeploy DeepCogito's Cogito v2 models on Runpod to experience frontier-level reasoning at lower inference costs, choose from 70B to 671B parameter variants.AI InfrastructureAll
Brendan McKeagJuly 25, 2025How to Run MoonshotAI’s Kimi-K2-Instruct on Runpod Instant ClusterRun MoonshotAI's Kimi-K2-Instruct on Runpod Clusters using H200 SXM GPUs and a 2TB shared network volume for seamless multi-node training. This guide.AI WorkloadsAll
Brendan McKeagJuly 25, 2025Comparing the 5090 to the 4090 and B200: How Does It Stack Up?Benchmark Qwen2.5-Coder-7B-Instruct across NVIDIA's B200, RTX 5090, and 4090 to identify optimal GPUs for LLM inference, compare token throughput, cost.All
Brendan McKeagJuly 18, 2025Iterative Refinement Chains with Small Language ModelsAs prompt complexity increases, large language models (LLMs) hit a "cognitive wall," suffering performance drops due to task interference and.AI WorkloadsAll
Brendan McKeagJuly 14, 2025Run Kimi-K2 on Runpod in a Single PodMoonshot AI's Kimi-K2-Instruct is a trillion-parameter, mixture-of-experts open-source LLM optimized for autonomous agentic tasks, with 32 billion active.AI WorkloadsAll
Brendan McKeagJune 27, 2025The Dos and Don’ts of VACE: What It Does Well, What It Doesn’tVACE introduces a powerful all-in-one framework for AI video generation and editing, combining text-to-video, reference-based creation, and precise.AI WorkloadsAll
Brendan McKeagJune 20, 2025Deep Dive Into Creating and Listing on the Runpod HubA deep technical dive into how the Runpod Hub streamlines serverless AI deployment with a GitHub-native, release-triggered model. Learn how hub.json and.Product UpdatesAll