Suspense Boundary Strategy
Place loading boundaries so the page fills in usefully rather than all at once.
What it adds
A decision about where the interface is allowed to wait, so slow regions never hold up the parts that are ready.
What your agent is told to do
5
What your agent is told to do
5-
1
Map each screen into the regions that can be shown independently — the shell and navigation, the primary content, then secondary panels and counts — and let each one arrive on its own rather than waiting for the slowest.
-
2
Draw a boundary wherever a slow region would otherwise delay a fast one, and only there. A boundary per component fragments the page into a dozen separately twitching rectangles.
-
3
Start every request a region needs as early as the parameters are known, so the requests overlap instead of queueing behind whichever component happens to render first.
-
4
Give each boundary a placeholder with the same dimensions as its eventual content, so regions resolving in an unpredictable order do not push each other around the page.
-
5
This entry decides where the boundaries go and what the placement is trying to achieve; the appearance of the pending, empty, and refreshing states inside them belongs to Async Boundary Component. Do not define a second set of skeletons here.
Edge cases it handles
7
Edge cases it handles
7- Content already on screen must stay on screen during a refetch. Falling back to a placeholder for data the user is currently reading is a regression, not a loading state.
- Server-rendered markup and the client's first render must agree about which regions were pending, or the page will visibly reshuffle the moment it becomes interactive.
- Requests must not chain. A region that fetches a record, renders, then fetches that record's related items turns two round trips into a sequence, and the boundary hides the delay rather than removing it.
- Every boundary needs a matching failure containment from Error Boundary Pattern. A region that can be pending can also fail, and without a fallback the failure escapes to the nearest ancestor and takes healthy regions with it.
- Streaming a page in fragments changes when the document title, the metadata, and the response status can be decided. Settle those before the parts that stream, or a not-found record will be delivered inside a successful page.
- Keyboard focus and scroll restoration must account for content that arrives after the initial paint, or a user who tabs immediately will land somewhere that then moves.
- Boundaries around content above the fold should be few. The user is judging the page by how quickly the top of it settles, not by how many pieces it was split into.
Definition of done
9
Definition of done
9- Each screen resolves in named stages, with the shell and primary content never waiting on secondary regions.
- The number of boundaries is deliberate and documented per screen, not one per component.
- Requests for a screen's regions are issued in parallel, with no region waiting on another's response to begin.
- Refetching data leaves the currently visible content in place.
- Server and client agree on the pending regions, with no visible reshuffle on hydration.
- Every boundary is paired with failure containment from Error Boundary Pattern.
- Placeholders match the dimensions of their content, and regions resolving out of order cause no layout shift.
- The feature matches the existing design system.
- No existing functionality is broken.
Related features
Prompt Versioning
Prompt Versioning
Tie every AI output to the exact prompt version that produced it.
What it does
Immutable, numbered versions of each prompt, with the run configuration recorded and every output stamped with the version used.
How it works
- 1 Make every publish create a new immutable version rather than overwriting the previous text. Editing history in place destroys the only record of what produced last month's outputs.
- 2 Capture the whole run configuration with each version, not just the wording: which model tier and parameters were used, which tools were available, and the expected output shape. A prompt that behaves differently under different settings is not one prompt.
- 3 Stamp every generated output with the version identifier that produced it, and keep that stamp with the record so an output found later can be traced back to its exact instructions.
Copy the prompt
No account needed
Add this feature to my app:
https://addthisfeature.com/x/prompt-versioning
Retrieval Debugger
Retrieval Debugger
Show exactly which sources, chunks, and scores produced a given AI answer.
What it does
A per-answer inspector showing the query as issued, the filters applied, the candidate chunks with their scores, and what reached the model.
How it works
- 1 Capture for each answer the query as it was issued, the filters applied, the candidates returned with their scores, and which of those actually made it into the request after the context ceiling was applied.
- 2 Show results after permission filtering, with a count of how many candidates were excluded and why. Displaying the pre-filter set turns the debugger into a way to read content the viewer cannot open.
- 3 Present each scoring stage separately — keyword, semantic, and any reranking — because a chunk that ends up first overall may have been rescued by one stage after being buried by another, and a single blended number hides that.
Copy the prompt
No account needed
Add this feature to my app:
https://addthisfeature.com/x/retrieval-debugger
AI Cost Budgets
AI Cost Budgets
Cap what AI features are allowed to spend before the bill arrives.
What it does
Monetary spending limits on AI work, scoped by workspace, feature, and time period, enforced before a run starts.
How it works
- 1 Find every place the app calls a model and route all of them through one accounting point that records estimated and actual spend against a scope. A budget that only covers the chat feature is not a budget.
- 2 Estimate the cost of a run from the size of its input before dispatching it, and refuse anything that would exceed the remaining budget on its own.
- 3 Reserve the estimate against the budget when the run starts, then reconcile to the real usage figures when it finishes, releasing whatever was over-reserved.
Copy the prompt
No account needed
Add this feature to my app:
https://addthisfeature.com/x/ai-cost-budgets
How it works
-
1
Copy the link
Grab the Markdown instruction URL for this feature.
-
2
Give it to your AI
Paste it into Claude Code, Cursor, v0, Lovable — whatever you build with.
-
3
It inspects, then implements
Your agent reads your existing app first, then adds the feature to fit it.
Works with your stack
These instructions are written to adapt. They tell the agent to detect your framework, match your existing design system, and reuse what you already have — rather than assuming a particular stack.
Need it tighter than that? Customize the feature and tell it exactly what you're running.