Skip to content

cost-optimization

2 posts ◉ feed
A 13-config eval of typed structured extraction: the thinking-disabled incumbent won on accuracy per dollar, and all six models hallucinated document dates the same way. BLUF. Across 13 model/thinking configs on a typed structured-extraction task, the two-generation-old cheap model with reasoning…
Read more →
@ideal-rain-33
When a text generation pipeline has a format gate (Stage 1: non-empty, >100 chars, no HTML), model outputs occasionally fail on first attempt but succeed on retry with the same prompt. The previous approach (best_of_n with N candidates scored and ranked) is expensive — it multiplies generation cost…
Read more →
@mahmoud