> For the complete documentation index, see [llms.txt](https://book.likun.ai/llms.txt). Markdown versions of documentation pages are available by appending `.md` to page URLs; this page is available as [Markdown](https://book.likun.ai/01-evaluation.md).

# 第一章：先学会衡量——评测体系的建立

> 把评测从"上线前测试"，重新放回"需求定义"的位置

***

## [1. 评测即需求，评测即产品](/01-evaluation/01-ping-ce-ji-xu-qiu.md)

* 从确定性操作到概率型供给
* 评测即需求：解构 Eval 的三件事
* 为什么是现在？——监督信号的范式转移
* 评测怎么做？——组织、流程与要素
* 结语：Sample / Eval 是 AI PM 的新语言

***

## 即将更新

* 2. Dataset：怎么构建一份"活的"评测集
* 3. Rubric：怎么把"什么算好"写成可执行的判分标准
* 4. Metric：从 Accuracy 到业务一致性，怎么选指标
* 5. Judge / Evaluator：人工、专家、LLM-as-a-Judge 的边界
* 6. 评测与产品迭代：让评测跟着产品一起活
* 7. Agent 时代：从 eval answer 到 eval trace


---

# Agent Instructions
This documentation is published with GitBook. GitBook is the documentation platform designed so that both humans and AI agents can read, navigate, and reason over technical content effectively. Learn more at gitbook.com.

## Querying This Documentation
If you need additional information that is not directly available in this page, you can query the documentation dynamically by asking a question.

Perform an HTTP GET request on the current page URL with the `ask` query parameter, and the optional `goal` query parameter:

```
GET https://book.likun.ai/01-evaluation.md?ask=<question>&goal=<endgoal>
```

`ask` is the immediate question: it should be specific, self-contained, and written in natural language.
`goal` is optional and describes the broader end goal you are ultimately trying to accomplish on behalf of the user. GitBook uses it to tailor the answer towards what is most useful for that goal.

The response will contain a direct answer to the question and relevant excerpts and sources from the documentation.

Use this mechanism when the answer is not explicitly present in the current page, you need clarification or additional context, or you want to retrieve related documentation sections.
