Apr 30, 2026Running AI Workloads Across NVIDIA, AMD, and Multiple Clouds Without RefactoringAdam King
Apr 21, 2026Fine-Tuning a 32B Legal LLM That Outperformed a Frontier Model at 4× Lower Serving CostAdam King
Apr 14, 2026LLM Inference GPU Sizing: How to Choose the Right GPU for Your Model and TrafficAdam King
Apr 7, 2026Why AI Infrastructure Software Is Harder Than Hardware — Lessons from Building Aurora and FlexAIAdam King
Apr 3, 2026Heterogeneous AI Compute: Why Mixed NVIDIA, AMD, and TPU Clusters Are Harder Than They LookAdam King
Mar 20, 2026From Supercomputers to Serverless: How FlexAI is Solving the GPU Infrastructure ChallengeAdam King