Review architecture and failure paths
Generating the code is the cheap part now. Deciding whether it can be trusted in production is still yours.
Inspect production-style AI code for architecture, tool safety, error handling and reliability. Build a repeatable review process with hands-on labs, checklists and rubrics.
With Nachiketh MurthyFounder, Manifold AI Learning · Systems ship. Demos don’t.
Review architecture and failure paths
Inspect tool calls and safeguards
Document actionable engineering feedback
The code runs. Tests pass. The demo looks clean. And then production traffic, real users, and live cost meters expose every assumption that nobody reviewed.
The API responds locally, but long-running LLM calls, missing timeouts, and synchronous execution can collapse under real traffic — latency spikes, exhausted workers, timeout cascades.
Direct tool access may look impressive in demos, but production systems need policy checks, audit logs, approval boundaries, and retry safety — otherwise the agent is a liability surface.
Without evals, metadata controls, document freshness checks, and access filtering, a RAG system can silently return the wrong context — and nobody notices until a customer does.
In most courses, you start with a blank file and build a demo. In this accelerator, you enter like a reviewer.
You look at existing AI system patterns. You inspect the code path. You mark risks. You explain why they matter. You decide what must be fixed before production.
Nine concrete review surfaces. Each card answers one question: what exactly will you inspect in the code path?
Inspect: endpoint contracts, LLM call boundaries, sync vs async paths, timeout handling, background-job hand-offs, response shapes.
Inspect: user_id / session_id scoping, global-state leaks, Redis TTLs, persistence boundaries, crash recovery, multi-user isolation.
Inspect: chunking strategy, metadata filtering, reranking, document freshness, retrieval evals, tenant access control.
Inspect: tool allowlists, gateway patterns, policy checks, approval boundaries, audit logs, unsafe direct-call paths.
Inspect: retry wrappers, duplicate side-effects, idempotency keys, partial-failure recovery, dedupe boundaries, queue safety.
Inspect: trace IDs, structured logs, latency capture per stage, token / cost attribution, failure classification, debuggability gaps.
Inspect: prompt-injection surface, secret handling, data exfiltration paths, PII boundaries, permission scoping, audit trails.
Inspect: environment config, Docker hygiene, health checks, queue / worker separation, rate limits, scaling assumptions.
Inspect: how to convert a vague worry into a specific, defensible review comment — with risk, fix, and reasoning in plain language.
Each module is a focused review workshop. You don't watch concepts — you inspect code patterns and leave with a review artefact you can reuse. This is — Architecture & Judgement. Modules have shipped weekly since 10 June 2026 and continue to roll out.
Goal: Learn how to review AI systems beyond "it works." Set the bar: open the code, walk the path, mark the risks, write the review.
Goal: Inspect FastAPI endpoints, LLM call boundaries, timeouts, async patterns, and background job needs. Identify exactly where the request path fails under load.
Goal: Inspect user_id / session_id handling, global state risks, Redis TTL, persistence, and crash recovery. Spot the leaks before users do.
Goal: Inspect chunking, metadata, retrieval strategy, reranking, stale documents, access control, and eval gaps. Catch silent retrieval drift before it ships.
Goal: Inspect tool access, gateway patterns, policy checks, retry behavior, audit logs, and unsafe execution paths. Identify every ungoverned-call surface.
Goal: Inspect duplicate calls, unsafe retries, partial failures, queue boundaries, and recovery logic. Find the path where one timeout becomes two charges.
Goal: Inspect trace IDs, structured logs, token / cost capture, latency visibility, and failure classification. Catch the systems nobody can debug after the fact.
Goal: Convert findings into clear code review comments, architecture feedback, and interview-ready explanations. Move from vague concern to specific engineering finding.
You will not just watch concepts. You will inspect broken or incomplete production-style patterns and learn how to review them — risk by risk, comment by comment.
Find long-running LLM calls inside request-response paths. Walk the failure modes — gateway timeouts, worker exhaustion, cascading retries — and write the review.
Find unsafe state handling and missing user / session isolation. Track exactly how one user's context can land in another user's reply, and where the boundary should have been.
Find where retries can create duplicate LLM / tool execution. Identify the missing idempotency key, the right layer to enforce it, and what to recommend in the review.
Find missing evals, weak metadata control, and stale document risks. Show how the system would silently return the wrong context — and what evidence the team should have shipped.
Find missing policy checks, audit logs, and approval boundaries. Map the unsafe execution surface and write the gateway-pattern recommendation the codebase needs.
Find missing trace IDs, token capture, latency logging, and cost tracking. Identify what cannot be debugged today and the first three pieces of instrumentation to add.
Find hardcoded config, missing health checks, and weak runtime assumptions. Surface every deployment assumption that breaks the moment this leaves localhost.
The difference between a worry and a finding is specificity — named surface, named failure, named fix. Three examples of the shift you will practise until it is automatic.
“This API may not scale.”
The endpoint performs a long-running LLM call inside the synchronous request path. Under concurrent traffic, this can increase latency, exhaust workers, and create timeout failures. This should move behind a job queue or async execution boundary with status polling.
“Memory handling is risky.”
Conversation state is not isolated by user_id and session_id. In a multi-user environment, this can cause data leakage or context contamination. State access should be scoped per user / session with TTL and persistence boundaries.
“Tool calling needs governance.”
The agent can invoke tools directly without policy validation, approval rules, or audit logs. Production tool execution should go through a gateway that enforces permissions, validates inputs, records execution, and handles failure safely.
Concrete review artefacts you carry into your own codebase, your own pull requests, and your own interview rounds.
The grading sheet — every surface, scored by severity and readiness.
The master pre-ship list across reliability, governance, observability, ops.
Request flow, LLM call boundaries, timeouts, async patterns, background hand-offs.
A structured write-up format for chunking, retrieval, evals, freshness, and access.
Per-tool risk pass — gateways, policies, audit trails, side-effect safety.
Duplicate calls, idempotency keys, partial-failure recovery, queue dedupe.
Template that captures missing trace IDs, cost capture, latency visibility.
Config, Docker, health checks, queues, rate limits, scaling assumptions.
Ready-to-adapt phrasings for risk, fix, and reasoning — no PR-blocking tone wars.
How to narrate findings clearly in interviews and engineering discussions.
Built for engineers, architects, data and cloud professionals who already ship, and now want the judgement to say what releases and what gets held — not absolute beginners or shortcut-seekers.
Software engineers shipping AI-powered features
Backend, DevOps, and MLOps engineers moving into GenAI
Data and ML engineers reviewing production AI codebases
AI/GenAI engineers tightening their architecture judgment
Working professionals preparing for AI/GenAI interviews
Bootcamp and accelerator learners wanting deeper review skills
Modules have shipped weekly since 10 June 2026 and continue to roll out — labs, checklists, and walkthroughs included. Early-access learners lock in today's price and receive every future update at no extra cost. Early-access pricing will increase as more modules go live.
🔒 Premium accelerator · No refunds once access is provisioned. Please review the course scope before enrolling.
Two sides of the same moment. AI Architect System Design is where you decide the shape before anyone writes code. AI Production Readiness Review is where you judge what actually got written. Both sit at — Architecture & Judgement. Production-Style RAG, the Interview Playbook and NCP-AAI Prep pick up on either side of it.
You already have the build skill. What is worth owning next is the judgement call — what ships, what gets held back, and the reasoning you can put in writing before users, cost meters and on-call pages make the argument for you.
You already bring real engineering experience. These live programs add the production layer on top of it — without asking you to start over.
Eight live weeks. One production-style Agentic AI system you build end to end — orchestration, governed tools & MCP, production RAG, async execution, evaluation, security, deployment — and every decision something you can defend. Nothing else required first: Python and LangChain foundation bonuses included free.
Secure Your Seat →Once you can ship the system, the harder question is which system to build. Discovery, scoping, an architecture you can defend, evaluation, delivery and adoption — twelve weeks of live case labs. Reserved for Diamond Members; not sold separately.
Explore Diamond →
I am Nachiketh. I help experienced technology professionals move beyond AI demos and build production-ready AI systems.
Systems ship. Demos don’t.Meet Nachiketh Murthy →