Lessons
From the last year
Audit the searchable corpus with category-representative queries before tuning an LLM content-linker
python rag semantic-search llm-pipeline corpus-coverage 330 tokens
Cross-reference LLM generation timestamp against data release schedules to diagnose factual errors
python llm news-digest data-releases validation 239 tokens