Testing AI-Dependent Code: Mocking LLMs, Evaluating Outputs, and Avoiding Flaky TestsOct 5, 2026·6 min read·68
Deploying AI Workflows on Cloud Run: Latency, Cold Starts, and ScalingThe team set min_instances = 3 to eliminate cold starts on the AI workflow service. Reasonable thinking — LLM calls already add 5–15 seconds of latency, a 4-second cold start on the first request afteOct 8, 2026·6 min read·1
Async AI Pipelines in FastAPI: Streaming, Queuing, and Long-Running RequestsSep 27, 2026·5 min read·5
GitHub Actions I Set Up on Every ProjectMost engineers using GitHub Actions daily know how to write a basic workflow — checkout, install dependencies, run tests, deploy. That covers 80% of what they need. The other 20% is a set of features Sep 21, 2026·6 min read·48
When Terraform Gets Messy: Refactoring, Imports, and State SurgeryS01E06 of Terraform for Application EngineersSep 19, 2026·6 min read·20
Terraform CI/CD: Plan on PR, Apply on MergeS01E05 of Terraform for Application EngineersSep 15, 2026·5 min read·17
Provisioning GCP Services with Terraform: Cloud Run, Pub/Sub, and Cloud SQLS01E04 of Terraform for Application EngineersSep 13, 2026·6 min read·8
Managing State, Secrets, and Environments in TerraformS01E03 of Terraform for Application EngineersSep 11, 2026·6 min read·13
Structuring OpenAI and Gemini Calls in FastAPI: Prompts, Fallbacks, and RetriesThe fallback was supposed to be temporary. When the OpenAI API was unavailable, the service would return the last cached result for that user rather than failing the request. Clean degraded experienceSep 8, 2026·7 min read·13