Cost Per Pull Request: A Unit Metric Worth Tracking
Total AI spend tells you nothing actionable. Cost per merged pull request ties inference to delivered work and exposes exactly where the money goes.
ReadPractical writing for developers building with large language models — how they work, how to pick one, and how to keep the bill predictable.
Total AI spend tells you nothing actionable. Cost per merged pull request ties inference to delivered work and exposes exactly where the money goes.
ReadMost prompts carry substantial waste. Nine techniques that reduce spend while usually improving output quality, ordered by how much they actually save.
ReadA field guide to the errors an LLM API actually returns: what each status means, which ones are worth retrying, and how to reproduce the failure in one curl command.
ReadA 1.6T MoE that activates 49B per token, ships under MIT, and costs $0.435 per million input tokens. Where V4 Pro is strong, where it is not, and Pro versus Flash.
ReadEmbeddings turn text into vectors so you can search by meaning instead of keywords. Here is how they work, what similarity really measures, and the failure modes to expect.
ReadAn agent that works four times in five is not 80 percent useful. Here is how to measure agent reliability in a way that predicts production behaviour.
ReadThree ways to make a general model do your specific job, with very different costs. Here is a decision rule based on what each one can and cannot actually change.
ReadMost AI budget forecasts are a headcount multiplied by a hopeful number. Here is a model that decomposes spend into drivers you can actually measure and control.
ReadFree tier numbers rot within weeks, and several vendors have stopped publishing them entirely. Here is a method for comparing them that survives the churn.
ReadA diff shows what changed, never why. Where generation genuinely helps, a working prepare-commit-msg hook, and the Conventional Commits rules a model gets wrong.
ReadMost generated docs restate the function signature in English. What to generate instead, which formats tools can actually consume, and how to stop docs drifting from code.
ReadText-to-SQL fails two ways: destructive queries and quietly wrong answers. Read-only roles and timeouts handle the first. The second needs a different kind of guardrail.
ReadShowing 325–336 of 404 articles