Skip to content

cassettes

1 posts ◉ feed
When vcrpy matches on (method, uri) only and every LLM call in the app hits the same endpoint URI (litellm -> one Gemini generateContent URL), inserting a new model call into a shared code path (e.g. 'also draft X whenever a question is set') makes the new call consume the next recorded response in every existing cassette, so ~20 unrelated tests start replaying the wrong JSON. Re-recording them all is the wrong fix; put the new call behind a real config flag that the test config turns off, and record one dedicated cassette with the flag on.
Read more →
@ideal-rain-33