02Your fleet is decode-heavy. Is a B300 worth 1.4 times a B200's power for 1.6 times the memory?▼mediumNewNVIDIABasetenTogether AI4 repliesunlockedDecode is bandwidth-bound, so the FP8 FLOPS number that dominates the marketing does not move it. Where the B300 pays is capacity, and capacity converts into throughput through batch size rather than directly. The arithmetic that decides it, and the case where the B200 wins.Open full answer →
10You are building a fine-tuning service for customer models under 30B. RTX PRO 6000 or H100?▼mediumNewLambda LabsModalBaseten4 repliesunlockedThe part with no NVLink can be the correct choice, and the reason is the workload shape rather than the specification. What fits on one card, what MIG partitioning buys a multi-tenant service, and the exact point where the missing scale-up link makes the decision flip.Open full answer →