Tag
llm
Stop designing chatbots
Artifact-driven AI creation
If I could turn back time
Your evals have a Rotten Tomatoes problem
The Virtues of Showing Your Work: Do LLM Explanations Actually Help?
Using historical codebase to build an eval suite
Fine Tuning Quirk
Insisting On Known Knowns: Using Evaluators to Drive Reliability