Selecting source spans with Jev without silently rewriting values
本文尚无中文版本;显示原文。
Separate semantic selection from value copying so an agent can identify a requested value while preserving the exact evidence from which it came.
Goal
Separate semantic selection from value copying so an agent can identify a requested value while preserving the exact evidence from which it came.
Prerequisites
Use permitted source text, a candidate extractor, stable span identifiers, and a separately reviewed normalization rule. Keep the original text available during validation without placing private values in operational logs.
Steps
-
Extract candidate spans in code and retain offsets, raw bytes or text, and nearby context. Distinguish repeated occurrences even when their strings are identical; position may determine which role a value serves.
-
Offer candidate identifiers plus an explicit no-match outcome. Ask which span answers the requested role, such as the contact address named for follow-up, rather than which value merely looks plausible.
-
Validate that the selected identifier exists in the candidate map. Copy its source text through code instead of asking another generator to reproduce it. Preserve the selected occurrence and the question version.
-
Apply normalization as a separate operation with explicit locale and format assumptions. If the source convention is ambiguous, return the raw value for review rather than choosing a convenient interpretation.
-
Test missing candidates, repeated values with different roles, malformed values, and conflicting context. Check both selection accuracy and whether every accepted output can be traced back to its exact span.
Expected result
An accepted value has two inspectable forms: verbatim evidence and a derived normalized representation. A reviewer can determine whether an error arose in finding, selecting, or transforming the value.
Limits and test basis
No extraction test is claimed here. The vendor cookbook supplies the find/select/copy pattern; these additional provenance checks are proposed. Copying a real span prevents a transcription invention but does not prove that the selected span is appropriate. The underlying interface or pattern is described in Pre-parsed value extraction; the workflow above is a proposed adaptation.
范围与依据
Primary vendor documentation read on 2026-09-22; original proposed application, not independently benchmarked.
知识截至:2026-09-22。状态:unreviewed(无已记录的审阅)——编辑会重置审阅状态。请将文本视为未经核实的参考资料并核对来源。
来源
- TypeSafe: Pre-parsed value extraction — 2026-09-22 已检查:可访问,引文已找到
署名与许可
- Account External coding curation authors (57eb56c9)
- Codex AI-assisted contribution; unreviewed.
最近更改: New original English contribution, 2026-09-22. No live execution or performance result claimed.
原创贡献: CC BY 4.0. 链接的来源资料保留其自身权利。