<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>LLM Application on 中文AI术语词典</title><link>https://terms-en.ai-term-hub.com/zh/tags/llm-application/</link><description>Recent content in LLM Application on 中文AI术语词典</description><generator>Hugo</generator><language>zh-cn</language><lastBuildDate>Sat, 18 Jul 2026 11:44:45 +0000</lastBuildDate><atom:link href="https://terms-en.ai-term-hub.com/zh/tags/llm-application/index.xml" rel="self" type="application/rss+xml"/><item><title>大模型作为裁判</title><link>https://terms-en.ai-term-hub.com/zh/terms/llm_as_a_judge/</link><pubDate>Sat, 18 Jul 2026 11:23:29 +0000</pubDate><guid>https://terms-en.ai-term-hub.com/zh/terms/llm_as_a_judge/</guid><description>&lt;h2 id="definition">Definition&lt;/h2>
&lt;p>“大模型作为裁判”（LLM-as-a-Judge）是一种评估范式，其中大语言模型充当其他模型输出质量的自动化评估者。这种方法旨在减少对人工标注员或严格规则匹配的依赖，通过提示工程让LLM根据特定标准（如相关性、安全性、创造性等）对生成内容进行打分或排序，从而提高评估效率和一致性。&lt;/p>
&lt;h3 id="summary">Summary&lt;/h3>
&lt;p>一种通过使用另一个大语言模型根据标准对响应进行评分或排名，从而评估大语言模型输出的方法。&lt;/p>
&lt;h2 id="key-concepts">Key Concepts&lt;/h2>
&lt;ul>
&lt;li>自动化评估&lt;/li>
&lt;li>提示工程&lt;/li>
&lt;li>模型对齐&lt;/li>
&lt;li>质量指标&lt;/li>
&lt;/ul>
&lt;h2 id="use-cases">Use Cases&lt;/h2>
&lt;ul>
&lt;li>RLHF模型的基准测试&lt;/li>
&lt;li>创意写作评估&lt;/li>
&lt;li>安全性和偏见检测&lt;/li>
&lt;/ul>
&lt;h2 id="related-terms">Related Terms&lt;/h2>
&lt;ul>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/rlhf-%E5%9F%BA%E4%BA%8E%E4%BA%BA%E7%B1%BB%E5%8F%8D%E9%A6%88%E7%9A%84%E5%BC%BA%E5%8C%96%E5%AD%A6%E4%B9%A0/">rlhf (基于人类反馈的强化学习)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/evaluation_metrics-%E8%AF%84%E4%BC%B0%E6%8C%87%E6%A0%87/">evaluation_metrics (评估指标)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/prompt_engineering-%E6%8F%90%E7%A4%BA%E5%B7%A5%E7%A8%8B/">prompt_engineering (提示工程)&lt;/a>&lt;/li>
&lt;li>&lt;a href="https://terms-en.ai-term-hub.com/en/terms/model_alignment-%E6%A8%A1%E5%9E%8B%E5%AF%B9%E9%BD%90/">model_alignment (模型对齐)&lt;/a>&lt;/li>
&lt;/ul></description></item></channel></rss>