FastAPI Testing: Mock LLM APIs for Free
Blog post from Speedscale
Speedscale proxymock is presented as a tool for reducing the cost and complexity of testing FastAPI applications that make multi-step calls to LLM providers such as OpenAI, Anthropic, Gemini, and xAI. Rather than maintaining fragile hand-written mocks for dependent model outputs, developers can record a real application run through a local proxy, save inbound and outbound request-response pairs as editable, Git-friendly Markdown files, and reuse them to mock provider, tool-service, and optionally database interactions. The recorded traffic can then support deterministic local development, CI regression tests, simulated provider latency, and load testing without consuming API tokens or requiring live credentials. The walkthrough uses a ticket-triage application whose requests include tool lookups and sequential LLM calls, showing how proxymock matches requests using recorded signatures and returns original responses, token counts, and timing. It also supports replaying captured inbound requests against the application, enforcing latency and failure thresholds in CI, recording multiple providers for fallback or comparison tests, inspecting and editing recordings in a terminal interface, and working across Python, Node.js, Java, and .NET applications.
| Trend | Post Mentions | Total Month Mentions | Posts | Companies | MoM |
|---|---|---|---|---|---|
| LLM | 16 | 7,531 | 1,250 | 268 | +26% |
Use this post, company, and trend context to find content marketing opportunities, perform competitive analysis, or address product feature gaps via the Plushcap MCP server or the Plushcap API.