JVNAUTOSCI-2600: thin adaptive ordinary-turn candidate - #309
Draft
witbrock wants to merge 4 commits into
Draft
Conversation
| seen_request_digests: set[str] = set() | ||
| last_partial_text = "" | ||
| final_synthesis = False | ||
| terminal_status = "completed" |
| return | ||
| result = close() | ||
| if inspect.isawaitable(result): | ||
| await result |
| return | ||
| close_result = close() | ||
| if inspect.isawaitable(close_result): | ||
| await close_result |
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Candidate boundary
This remains the frozen JVNAUTOSCI-2600 A3r2 candidate and evidence basis. Do
not merge, cut over, release, deactivate live represented-controller
artefacts, expand delegated effects in this branch, or start A4. Michael has
authorised a bounded successor A3 round; that work will remain isolated from
this exact commit.
The five A3r2 cases were revealed only after commit
c7e69a8c26c33b9f951612e5de620b28d77321ab, its tree, the active representedprompt, and the relationship index were fixed. Each exact case was submitted
once in order, without retry, rephrasing, replacement, sixth case, state repair,
or post-reveal tuning.
Candidate architecture
/von/generateandvon_chat_runthrough a direct adaptiveread engine.
request wording.
evidence listing, and selective evidence hydration.
arguments.
completion gate, context adjudication, gap recovery, controller-reporting
contracts, and strategy-prescribing tests from ordinary turns.
read route.
“Thin” describes the ordinary entry architecture relative to the removed
controller, not a claim that this is a small patch: the cumulative PR changes
195 files and adds substantial engine, evidence, transport, and test code.
A3r2 general repair
The final candidate repairs general affordances exposed by A3r1 without adding
case policies or deterministic semantic gates:
synthesis and context-limit recovery;
truthful surface and argument metadata;
capability and schema reachable;
client per bounded operation;
isErroras failure, and prevents private error bodies entering logs;
readiness;
missing support without dropping an existing index; and
removing an N+1 access-filter path without changing visibility semantics.
The revealed case IDs, people, and paper names do not occur in added candidate
lines. The five raw prompts contained no likely-tool hints or required-tool
flags, and the model selected 19 distinct delegated read capabilities.
A3r2 one-shot outcome
comparison, both failing the explicitly requested persistent representation.
current-student type extents and
hasDescriptionwas treated as the semanticdefinition of represented research evidence.
paper/assessment/student/match and guessed a different premise.
ok, 0 timeout, 0 provider/schema failure.okmeans an answer was returned, not that the user job or effect completed.case. Four of five exceeded 100 seconds.
The experiment supports opportunity-preserving adaptive discovery. It does not
yet support merge/cutover, effect competence, dependable actor-relative
semantics, or production latency. There is no matched baseline or ablation, so
value cannot be attributed to a single changed component.
Post-run schema-migration correction
The earlier “read-purity breach/proof of harm” conclusion was wrong and is
retracted.
find_concepts_by_namewas a pure search. A following main-branchfetch_conceptmigrated two pre-existing legacynames[]values intoequivalent
hasNamerelations and removed the legacy source field. Thishappened after the five experimental runs. No evidence establishes a change to
effective represented-name closure or user-visible semantic harm; the derived
relations may have made later reads faster.
The audit then deleted those relations as “contamination” after the legacy
source had already gone. That cleanup was the demonstrated destructive step.
Both exact names have been restored through the canonical Vontology interface
and read back with their original
en-NZ/name_type=NLmetadata:6a66e7aa0d5d79b23affb5436a66e7b00d5d79b23affb545The main migration still has a real partial-failure hazard: it loops over
individual writes and then unsets the whole legacy source. The frozen
candidate's union of modern and legacy names is the semantic oracle, but its
blanket removal of materialisation is not automatically an optimisation.
The revised boundary is semantic-effect fidelity, not physical storage-write
purity. Derived cache, index, relation, or schema-maintenance writes can be
compatible with a logical read when wholly derivable from authoritative state,
scoped, idempotent, semantics-preserving, recoverable or rebuildable, and safe
under partial failure. Materialisation should be implemented only if matched
measurements establish a useful benefit over pure projection.
Validation
correction.
final correction; 44 focused substrate/access tests also passed during
diagnosis.
git diff --checkpassed.gpt-5.6-lunaprofile, returnedREADY, and made no tool call.12,892-hit result and access filtering while improving from 58.8s to 10.6s.
on the exact final commit.
planner/critic replacement. The fresh cases subsequently exposed the
semantic, effect, continuity, and latency limits above.
Net change against
origin/main: 20,815 insertions, 100,319 deletions.Live derived-index state
The pre-existing extent collection was degraded and canonical fallback timed
out. A rebuild first failed under the ordinary 10s socket timeout after a
partial derived-row deletion; readiness stayed failed and canonical concepts
were untouched. An operation-scoped 120s maintenance timeout then rebuilt
67,946 canonical source concepts into 226,155 derived edges and persisted
current-schema
readystate. The successful pass took approximately 23minutes. This long, non-atomic rebuild cost is an operational defect, not
acceptance evidence.
Authorised successor A3 round
completion-gate/typed-receipt and selector/decision-attribution assumptions;
retain raw outcomes, evidence, explicit effects/read-backs, and telemetry.
failure with a stronger model. Compare a representative useful case with a
faster model. Preserve the originals and record requested/effective model,
total/model latency, call/tool counts, and outcome.
general actor-relative/linked-evidence principle; do not add a
my_studentsresolver or storage-predicate rule.additive/recoverable representation primitives, with trusted identity and
scope outside model arguments, concrete effect IDs, partial-success evidence,
and canonical read-back. Do not add an arXiv/student route or semantic
intent classifier.
frozen union projection unless matched cold/warm measurements justify a
small additive, tagged, reversible cache that never retires its source
during a read.
evidence set. Do not implement a continuity mechanism until a valid scoped
trial fails.
Model escalation/downshift is an observable experiment and a potential
represented routing policy—not an automatic retry chain or a Python
model-name branch. Full-paper and latency repair must follow measured substrate
evidence. No successor merge/cutover or A4 is authorised.
The branch commit and version history remain the rollback boundary.