← Software Services

Technical deep dives

Technical deep dives

Architecture, code, and trade-offs for engineers and technical leads. Each one pairs with a short, plain-English post.

All blogsTechnical deep dives (15)
Hub-and-spoke of five GenAI patterns around a shared stack
Deep Dive·Oct 5, 2026·7 min read

Five GenAI Patterns and Their Technical Blueprints

Reference blueprints for the five projects that ship fastest: support, document Q&A, drafting, search, and reporting.

Read →
Histogram of outputs by confidence with two thresholds creating full review, light review and auto-approve zones
Deep Dive·Sep 29, 2026·7 min read

Designing Human-in-the-Loop Workflows for GenAI

Where to place review, how to route by risk and confidence, and how to turn corrections into improvements.

Read →
Pyramid of measurement levels: offline evaluation, online metrics and business outcome experiments
Deep Dive·Sep 22, 2026·7 min read

Measuring GenAI Quality: Offline Evals, Online Metrics, and A/B Tests

A measurement stack for GenAI: golden sets, model judges, production signals, and controlled rollouts.

Read →
Cumulative cost lines for buy, platform and custom build crossing over time
Deep Dive·Sep 15, 2026·7 min read

Build vs Buy for GenAI: An Engineering Decision Framework

How to decide between off-the-shelf tools, configurable platforms, and custom builds, including the hidden costs of each.

Read →
Data flow with trust boundaries showing the model provider and logs as risk points
Deep Dive·Sep 8, 2026·7 min read

Secure GenAI Architecture: Protecting Data End to End

Threats, controls, and patterns for GenAI systems: redaction, access control, prompt injection, and audit.

Read →
Bubble chart of candidate use cases by value and feasibility with bubble size showing risk
Deep Dive·Sep 1, 2026·7 min read

Scoring GenAI Use Cases With a Weighted Model

A transparent scoring method for choosing between candidate projects, with a small script you can reuse.

Read →
Radar chart comparing readiness scores across six areas with the required level for limited release
Deep Dive·Aug 25, 2026·7 min read

From Pilot to Production: A GenAI Readiness Checklist

The engineering gaps between a convincing demo and a system people rely on, with a checklist to close them.

Read →
Stacked bars for three adoption scenarios split into build, inference, infrastructure, operations and review
Deep Dive·Aug 18, 2026·7 min read

A Total Cost of Ownership Model for GenAI Systems

A worked cost model: build, run, and operate, with the levers that move each line.

Read →
Quadrant placing prompting, RAG, fine-tuning and their combination by how fast facts change and how specific behavior must be
Deep Dive·Aug 11, 2026·7 min read

RAG vs Fine-Tuning: A Technical Decision Guide

When retrieval is enough, when fine-tuning pays off, and how to test the choice instead of arguing about it.

Read →
Two bar charts of next-token probabilities at low and high temperature
Deep Dive·Aug 4, 2026·7 min read

How LLMs Work for Practitioners: Tokens, Context, Sampling, and Cost

The mechanics that explain price, latency, and quirks, in the depth an engineering lead needs.

Read →
Heatmap of evaluation pass rate by category over five versions
Deep Dive·Jul 28, 2026·7 min read

Our GenAI Delivery Framework: Eval-Driven Development in 30 Days

How we structure a month-long build around an evaluation set, so quality is measured from week one.

Read →
Two ranked lists from vector and keyword search merged by reciprocal rank fusion
Deep Dive·Jul 14, 2026·7 min read

Production Document Q&A: Chunking, Hybrid Search, Citations, and Evals

The engineering choices that decide whether a document assistant is trusted: how you split, search, cite, and test.

Read →
Circular control loop of observe, plan, act and check with stop conditions
Deep Dive·Jun 30, 2026·7 min read

Designing Reliable AI Agents: Tool Use, Control Loops, and Failure Modes

An agent is a loop with tools and limits. Here is how to build one that behaves predictably in production.

Read →
Swimlane of a support copilot from ticket creation through retrieval, drafting, a safety gate and agent review
Deep Dive·Jun 16, 2026·7 min read

Engineering a Support Copilot: Retrieval, Drafting, and Review

A technical walk through ticket intake, hybrid retrieval, grounded drafting, confidence gating, and the feedback loop.

Read →
Four stacked platform layers with notes on what each provides
Deep Dive·Jun 2, 2026·7 min read

A GenAI Reference Architecture and Delivery Roadmap

The platform pieces you need before the second GenAI project, and the order to build them in.

Read →