<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Subject: Possible incorrect answer key — Evaluating AI Agents assessment (Generative AI Engineer) in Certifications</title>
    <link>https://community.databricks.com/t5/certifications/subject-possible-incorrect-answer-key-evaluating-ai-agents/m-p/168493#M4918</link>
    <description>&lt;P&gt;Hello all,&lt;/P&gt;&lt;P&gt;I'd like to report what appears to be an incorrectly keyed question in the assessment for the &lt;STRONG&gt;Evaluating AI Agents&lt;/STRONG&gt; course.&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;Question:&lt;/STRONG&gt;&lt;BR /&gt;"According to the lecture on evaluating AI agents, which of the following is described as a key reason why evaluation must be treated as a continuous process rather than a one-time validation step?"&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;My answer (marked incorrect, score 0):&lt;/STRONG&gt;&lt;BR /&gt;"Production agents encounter diverse user queries, usage patterns change over time, and new failure modes emerge that were not anticipated during development."&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;The answer the platform marks correct:&lt;/STRONG&gt;&lt;BR /&gt;"Unity Catalog enforces version limits on registered models, requiring re-evaluation before each new version is promoted."&lt;/P&gt;&lt;P&gt;I believe the keyed answer is incorrect for two reasons.&lt;/P&gt;&lt;P&gt;First, Unity Catalog does not enforce a version limit on registered models that triggers mandatory re-evaluation. Even if such a limit existed, a registry constraint would be an operational detail rather than a reason evaluation must be continuous.&lt;/P&gt;&lt;P&gt;Second, the keyed answer contradicts other questions within the same assessment. The question on offline versus online evaluation is keyed to: &lt;EM&gt;"Offline evaluation datasets may not fully represent real user behavior, can become stale as usage patterns evolve, and cannot capture issues that only emerge at scale."&lt;/EM&gt; The question on the offline/online feedback loop is keyed to a sequence built on analysing production traces for failures and edge cases not anticipated during development. Both reflect the same reasoning as the answer marked incorrect here.&lt;/P&gt;&lt;P&gt;I've attached a screenshot showing the question, my selection, and the feedback.&lt;/P&gt;&lt;P&gt;Could someone from the course team confirm whether this key needs updating? And has anyone else encountered the same result on this question?&lt;/P&gt;</description>
    <pubDate>Mon, 14 Sep 2026 01:14:24 GMT</pubDate>
    <dc:creator>amittian</dc:creator>
    <dc:date>2026-09-14T01:14:24Z</dc:date>
    <item>
      <title>Subject: Possible incorrect answer key — Evaluating AI Agents assessment (Generative AI Engineer)</title>
      <link>https://community.databricks.com/t5/certifications/subject-possible-incorrect-answer-key-evaluating-ai-agents/m-p/168493#M4918</link>
      <description>&lt;P&gt;Hello all,&lt;/P&gt;&lt;P&gt;I'd like to report what appears to be an incorrectly keyed question in the assessment for the &lt;STRONG&gt;Evaluating AI Agents&lt;/STRONG&gt; course.&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;Question:&lt;/STRONG&gt;&lt;BR /&gt;"According to the lecture on evaluating AI agents, which of the following is described as a key reason why evaluation must be treated as a continuous process rather than a one-time validation step?"&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;My answer (marked incorrect, score 0):&lt;/STRONG&gt;&lt;BR /&gt;"Production agents encounter diverse user queries, usage patterns change over time, and new failure modes emerge that were not anticipated during development."&lt;/P&gt;&lt;P&gt;&lt;STRONG&gt;The answer the platform marks correct:&lt;/STRONG&gt;&lt;BR /&gt;"Unity Catalog enforces version limits on registered models, requiring re-evaluation before each new version is promoted."&lt;/P&gt;&lt;P&gt;I believe the keyed answer is incorrect for two reasons.&lt;/P&gt;&lt;P&gt;First, Unity Catalog does not enforce a version limit on registered models that triggers mandatory re-evaluation. Even if such a limit existed, a registry constraint would be an operational detail rather than a reason evaluation must be continuous.&lt;/P&gt;&lt;P&gt;Second, the keyed answer contradicts other questions within the same assessment. The question on offline versus online evaluation is keyed to: &lt;EM&gt;"Offline evaluation datasets may not fully represent real user behavior, can become stale as usage patterns evolve, and cannot capture issues that only emerge at scale."&lt;/EM&gt; The question on the offline/online feedback loop is keyed to a sequence built on analysing production traces for failures and edge cases not anticipated during development. Both reflect the same reasoning as the answer marked incorrect here.&lt;/P&gt;&lt;P&gt;I've attached a screenshot showing the question, my selection, and the feedback.&lt;/P&gt;&lt;P&gt;Could someone from the course team confirm whether this key needs updating? And has anyone else encountered the same result on this question?&lt;/P&gt;</description>
      <pubDate>Mon, 14 Sep 2026 01:14:24 GMT</pubDate>
      <guid>https://community.databricks.com/t5/certifications/subject-possible-incorrect-answer-key-evaluating-ai-agents/m-p/168493#M4918</guid>
      <dc:creator>amittian</dc:creator>
      <dc:date>2026-09-14T01:14:24Z</dc:date>
    </item>
    <item>
      <title>Re: Subject: Possible incorrect answer key — Evaluating AI Agents assessment (Generative AI Engineer</title>
      <link>https://community.databricks.com/t5/certifications/subject-possible-incorrect-answer-key-evaluating-ai-agents/m-p/168530#M4920</link>
      <description>&lt;P&gt;Hi &lt;a href="https://community.databricks.com/t5/user/viewprofilepage/user-id/250199"&gt;@amittian&lt;/a&gt;, Good catch. Your answer looks more aligned with the idea of continuous evaluation, especially since user behavior can change and new issues can appear in production.&lt;/P&gt;&lt;P&gt;The Unity Catalog option seems unrelated to the question. I would suggest the course team review the answer key and update it if needed. Thanks for sharing this with the community.&lt;/P&gt;</description>
      <pubDate>Mon, 14 Sep 2026 13:29:18 GMT</pubDate>
      <guid>https://community.databricks.com/t5/certifications/subject-possible-incorrect-answer-key-evaluating-ai-agents/m-p/168530#M4920</guid>
      <dc:creator>Brahmareddy</dc:creator>
      <dc:date>2026-09-14T13:29:18Z</dc:date>
    </item>
  </channel>
</rss>

