<?xml version="1.0" encoding="UTF-8"?><rss version="2.0"><channel><title>Engineering — Ode</title><description>Engineering notes, architecture decisions, and build logs from the team at Ode.</description><link>https://www.ode.com/engineering</link><language>en</language><item><title>How we build evals for document extraction</title><link>https://www.ode.com/engineering/evals-for-document-extraction</link><guid isPermaLink="true">https://www.ode.com/engineering/evals-for-document-extraction</guid><description>Learn how to evaluate LLM-powered PDF extraction using exhaustive ground truth, synthetic documents, partial labels, and output validation.</description><pubDate>Thu, 03 Sep 2026 00:00:00 GMT</pubDate></item><item><title>Stop designing chatbots</title><link>https://www.ode.com/engineering/stop-designing-chatbots</link><guid isPermaLink="true">https://www.ode.com/engineering/stop-designing-chatbots</guid><description>On defaulting to chat as the interface for AI products — and why it&apos;s holding you back.</description><pubDate>Mon, 30 Mar 2026 00:00:00 GMT</pubDate></item><item><title>Artifact-driven AI creation</title><link>https://www.ode.com/engineering/artifact-driven-ai-creation</link><guid isPermaLink="true">https://www.ode.com/engineering/artifact-driven-ai-creation</guid><description>How artifact-based thinking makes AI-assisted creation precise and cumulative.</description><pubDate>Fri, 27 Mar 2026 00:00:00 GMT</pubDate></item><item><title>If I could turn back time</title><link>https://www.ode.com/engineering/if-i-could-turn-back-time</link><guid isPermaLink="true">https://www.ode.com/engineering/if-i-could-turn-back-time</guid><description>Have you ever exclaimed after a few hours engaged in some fruitless activity, &quot;Well, that&apos;s X hours of my life I&apos;ll never get back&quot;?</description><pubDate>Thu, 19 Feb 2026 00:00:00 GMT</pubDate></item><item><title>Your evals have a Rotten Tomatoes problem</title><link>https://www.ode.com/engineering/your-evals-have-a-rotten-tomatoes-problem</link><guid isPermaLink="true">https://www.ode.com/engineering/your-evals-have-a-rotten-tomatoes-problem</guid><description>A single eval score tells you that something regressed, not what regressed. Scoring outputs along specific dimensions — as close to each individual LLM call as you can get — turns an unhelpful number into a diagnosis.</description><pubDate>Thu, 12 Feb 2026 21:30:09 GMT</pubDate></item><item><title>The Virtues of Showing Your Work: Do LLM Explanations Actually Help?</title><link>https://www.ode.com/engineering/the-virtues-of-showing-your-work-do-llm-explanations-actually-help</link><guid isPermaLink="true">https://www.ode.com/engineering/the-virtues-of-showing-your-work-do-llm-explanations-actually-help</guid><description>If you&apos;ve spent any time prompting large language models, you&apos;ve probably heard of Chain of Thought (CoT) reasoning—the technique of asking an LLM to...</description><pubDate>Thu, 05 Feb 2026 00:00:00 GMT</pubDate></item><item><title>Do what I mean</title><link>https://www.ode.com/engineering/do-what-i-mean</link><guid isPermaLink="true">https://www.ode.com/engineering/do-what-i-mean</guid><description>Tests are crucial to reliable code. I’ve found I do a better job writing tests if I think of them as specifications and write them before I write code.</description><pubDate>Wed, 05 Feb 2025 00:00:00 GMT</pubDate></item><item><title>Using historical codebase to build an eval suite</title><link>https://www.ode.com/engineering/using-historical-codebase-to-build-an-eval-suite</link><guid isPermaLink="true">https://www.ode.com/engineering/using-historical-codebase-to-build-an-eval-suite</guid><description>How thousands of pull requests let Ode automate annotating a ground truth set for a new product.</description><pubDate>Mon, 11 Nov 2024 00:00:00 GMT</pubDate></item><item><title>Fine Tuning Quirk</title><link>https://www.ode.com/engineering/fine-tuning-quirk</link><guid isPermaLink="true">https://www.ode.com/engineering/fine-tuning-quirk</guid><description>I was working on content moderation for Change.org. Our goal was to automatically identify violative content to flag for reviewers.</description><pubDate>Tue, 27 Aug 2024 00:00:00 GMT</pubDate></item><item><title>Insisting On Known Knowns: Using Evaluators to Drive Reliability</title><link>https://www.ode.com/engineering/insisting-on-known-knowns-using-evaluators-to-drive-reliability</link><guid isPermaLink="true">https://www.ode.com/engineering/insisting-on-known-knowns-using-evaluators-to-drive-reliability</guid><description>Learn how evaluators can enhance the accuracy and reliability of LLM-driven GenAI projects through rigorous correctness measurement techniques.</description><pubDate>Thu, 08 Aug 2024 00:00:00 GMT</pubDate></item><item><title>Customizing Cursor&apos;s code generation and personality</title><link>https://www.ode.com/engineering/customizing-cursors-code-generation-and-personality</link><guid isPermaLink="true">https://www.ode.com/engineering/customizing-cursors-code-generation-and-personality</guid><description>One of our rules at Ode is “overuse AI”. This is a bit tongue-in-cheek, but what we mean by it is that even within our own work, we occasionally suspend...</description><pubDate>Sun, 29 Oct 2023 00:00:00 GMT</pubDate></item></channel></rss>