Skip to content
Prompt Smith AI
📰News🏠Home✍️Blog📋Templates📖Guides🧪Lab🎯Practice🎬Videos📚Reading✅Checklist
Login

Blog

Insights on AI engineering, strategy, and production systems

Engineering

How We Cut Inference Costs by 60% with Prompt Caching

A deep dive into how prompt caching transformed our AI infrastructure costs and improved response times across our production pipeline.

Maya Chen·Mar 15, 2026·8 min read
Strategy

The Case for Smaller Models in Production

Why we moved critical workloads from GPT-4 class models to fine-tuned smaller models — and how it improved both cost and reliability.

James Okafor·Mar 10, 2026·6 min read
Engineering

Building Reliable AI Agents: Lessons from 6 Months in Production

Hard-won lessons from running autonomous AI agents in production — from error handling to graceful degradation.

Sarah Kim·Mar 5, 2026·12 min read
Tutorials

RAG vs Fine-Tuning: A Practical Decision Framework

A practical guide to choosing between RAG and fine-tuning for your AI application, based on real-world trade-offs.

Alex Rivera·Feb 28, 2026·10 min read
Strategy

The Hidden Cost of AI Technical Debt

AI systems accumulate technical debt faster than traditional software. Here's how to identify it, measure it, and pay it down.

Marcus Thompson·Feb 12, 2026·9 min read