Decision library

Workload placement research, grouped by the decision that triggered it.

RunPlacement pages are built for quick scanning: short answers, rough math, tradeoffs, sources, and a decision rule. Start with the confusion closest to the workload.

Decision pages

Start with the durable decision library.

Admin-approved drafts can still publish here, but the main problem pages live at the decision index.

GPU pricingH100 Quote Checklist: What to Ask Before Choosing GPU CloudCommercial investigation

An H100 quote is worth comparing only after the provider exposes the GPU shape, minimum rental window, storage, data transfer, capacity model, retry risk, and support terms.

AWS bill shockAWS NAT Gateway Bill Shock: What to Check FirstProblem diagnosis

NAT Gateway bill shock usually means private subnet traffic is taking an expensive path. Start by finding which workload, route table, availability zone, or transfer pattern created the processed-data spike.

GPU pricingGPU Cloud Idle Cost: How to Price Wasted Accelerator TimeCost estimation

GPU cloud idle cost is the gap between paid accelerator time and useful workload progress. It matters most for training retries, batch queues, and inference fleets with low baseline utilization.

Cloud migrationCloud Egress and Exit Cost: What to Price Before MovingMigration planning

Cloud egress is only one part of exit cost. A serious migration estimate also prices data export, recurring transfer, storage retrieval, rewrites, testing, downtime, rollback, and new operations.

Cloud migrationBare Metal vs Cloud Break-Even: When Dedicated Servers WinCommercial comparison

Bare metal can win when a workload is steady, portable, highly utilized, and operationally owned. Cloud usually wins when flexibility, managed services, or variable demand matter more than unit cost.

GPU pricingRunPod vs Lambda GPU Cloud: How to Compare the FitProvider comparison

RunPod vs Lambda is less about one universal winner and more about workload fit. Compare GPU availability, storage behavior, operational model, support needs, and total job cost for your actual workload.

Start with GPU pricingH100 quotes, hidden GPU fees, provider comparisons, and utilization. Start with AWS bill shockNAT, data transfer, CloudWatch, S3, and AWS-vs-smaller-cloud decisions. Start with cloud migrationAWS exit, bare metal, portability, and when staying put is smarter. Use the resource libraryChecklists and worksheets designed to be shared and reused.

AI inference cost

AI inference cost

API, managed inference, self-hosted GPU, batch, realtime, and hybrid serving decisions.

AWS bill shock

AWS bill shock

Start here when the bill jumped and the expensive line item is not obvious.

Provider comparisons

Provider comparisons

When the hard part is choosing between provider categories, not reading another pricing page.

Capacity decisions

Capacity decisions

Commitment, reservation, on-demand, and utilization tradeoffs.

Cost breakdowns

Cost breakdowns

Line-item checklists for finding what the hourly rate leaves out.

Cloud migration

Cloud migration

AWS exit and workload portability decisions.

Resources

Linkable assets

Checklists and worksheets built for practical sharing, not promotional posting.

GPU pricingGPU Cloud Quote ChecklistChecklist / 7 sections / source-linked

A practical checklist and visual worksheet for comparing GPU cloud quotes beyond the advertised hourly rate.

AWS bill shockAWS Bill Shock Triage ChecklistChecklist / 7 sections / source-linked

A first-pass checklist and visual triage flow for finding the AWS line items that usually make a bill jump.

AWS bill shockAWS Bill Shock Evidence ChecklistResearch checklist / 4 sections / source-linked

A source-backed checklist for collecting AWS Cost Explorer, NAT Gateway, transfer, CloudWatch, storage, and routing evidence before changing architecture.

Cloud migrationCloud Exit Cost ChecklistChecklist / 7 sections / source-linked

A checklist and payback worksheet for pricing the real cost of leaving AWS, GCP, or Azure before migration starts.

Cloud migrationCloud Exit Assumptions IndexResearch index / 4 sections / source-linked

A source-backed index of the assumptions to collect before estimating cloud exit payback, partial migration, or workload re-placement.

Workload placementWorkload Placement WorksheetChecklist / 7 sections / source-linked

A practical worksheet and decision map for deciding where a workload should run before provider choice hardens.

Workload placementWorkload Placement Assumptions IndexResearch index / 4 sections / source-linked

A source-backed index of the assumptions to collect before choosing cloud, GPU cloud, bare metal, managed platform, or hybrid placement.

AI inference costAI Inference Cost ChecklistChecklist / 8 sections / source-linked

A practical checklist for estimating AI inference cost across APIs, managed inference, self-hosted GPUs, batch jobs, realtime endpoints, and hybrid routing.

AI inference costAI Inference Cost Assumptions IndexResearch index / 4 sections / source-linked

A source-backed index of the workload assumptions to collect before estimating API, managed inference, batch, GPU cloud, or self-hosted GPU cost.

AI inference costProvider Pricing Page Field AuditResearch audit / 4 sections / source-linked

A provider-neutral audit of the fields to verify on official pricing and deployment pages before comparing AI inference serving options.

AI inference costRealtime vs Batch Inference Cost Research GuideResearch guide / 7 sections / source-linked

A source-backed guide to deciding when realtime, asynchronous, batch, or hybrid inference changes effective AI serving cost.