fix: improve video agent research, finalization, and diagnostics - #30
Merged
Conversation
|
The latest updates on your projects. Learn more about Vercel for GitHub.
|
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
Video research can lose useful evidence or return an overly broad partial label when extraction, validation, or synthesis does not complete. This change improves provider configuration, source selection, transcript grounding, and run diagnostics while keeping the classification, research, and finalization phases separate.
Validation: 115 focused tests and 17 Durable Object integration tests passed on this exact working tree, along with TypeScript, generated OpenAPI checks, and git diff checks. The move onto the latest main preserved all changed files byte-for-byte. Earlier component and live-provider checks informed the implementation; their captures and reports remain local and are excluded from this PR.
Known limitation: a fetched transcript still becomes accepted evidence only after analysis succeeds. A repair canceled at the research deadline can therefore still return insufficient evidence. These diagnostics make future rejections inspectable; they do not recover missing diagnostics from historical runs or establish that all inference reliability issues are resolved. The latest classification, storyboard-selection, and diagnostic changes have not been deployed to production. Production must have FIREWORKS_API_KEY configured, and the processor image must be rebuilt for the storyboard changes.