テキスト評価 (Evaluate Text Result)
Evaluate Text Result ノードを使用すると、Griptape の評価エンジン (Eval Engine) を用いて、特定の評価基準に基づいてテキスト出力を評価できます。このノードは、AI が生成したコンテンツの検証、事実としての正確性のチェック、または出力品質の査定に役立ちます。
入力 (Inputs)
- Examples(プロパティ): プリセットの例から選択するか、独自の評価設定を作成します
- 選択肢:
- Choose a preset.. (プリセットを選択)
- Paraphrase (言い換え)
- Factual (事実性)
- Analogy (類推)
- 選択肢:
- Input(入力ピン / プロパティ): 評価対象となる入力テキスト
- 複数行テキスト入力に対応
- Expected Output(入力ピン / プロパティ): 期待される正解または参照テキスト
- 単一行テキスト入力
- Actual Output(入力ピン / プロパティ): 実際に生成され評価される出力テキスト
- 単一行テキスト入力
- Criteria(入力ピン / プロパティ): 適用する評価基準
- 複数行テキスト入力に対応
- 例: "Does the output accurately paraphrase the input without losing meaning?(出力は意味を損なうことなく入力を正確に言い換えていますか?)"
出力 (Outputs)
- Score(出力): 評価スコアを表す 0 から 1 までの浮動小数点数値
- 1.0 は完全な一致または基準達成を示します
- 0.0 は完全な不一致を示します
- Reason(出力): 評価結果の詳細な理由・根拠
- そのスコアが付与された理由のフィードバックを提供
- 検出された齟齬や問題点を説明
使用例 (Example Usage)
言い換えの評価 (Paraphrase Evaluation)
Input: "The quick brown fox jumps over the lazy dog."
Expected Output: "A swift brown fox leaps above a sleeping dog."
Actual Output: "A fast fox jumps over a dog that's not awake."
Criteria: "Does the output accurately paraphrase the input without losing meaning?"
事実性の評価 (Factual Evaluation)
Input: "The capital of France is Paris."
Expected Output: "Paris is the capital city of France."
Actual Output: "France's capital is Paris."
Criteria: "Is the output factually correct based on the input?"
類推の評価 (Analogy Evaluation)
Input: "A bird is to sky as a fish is to ______."
Expected Output: "water"
Actual Output: "concrete"
Criteria: "Does the output correctly complete the analogy?"
補足事項 (Notes)
- このノードは評価の実行に Griptape の Eval Engine を使用します
- 評価は指定された評価基準(Criteria)に基づいて実施されます
- スコアは 0 から 1 の範囲に正規化されます
- 理由は評価に関する詳細な定性フィードバックを提供します
- プリセットのサンプルを利用することも、完全にカスタムな評価基準を定義することも可能です