Conversation
Deploying with
|
| Status | Name | Latest Commit | Updated (UTC) |
|---|---|---|---|
| ✅ Deployment successful! View logs |
stackone-talks | 3267ecf | Sep 22 2026, 08:49 PM |
This file contains hidden or bidirectional Unicode text that may be interpreted or compiled differently than what appears below. To review, open the file in an editor that reveals hidden Unicode characters.
Learn more about bidirectional Unicode characters
Sign up for free
to join this conversation on GitHub.
Already have an account?
Sign in to comment
Add this suggestion to a batch that can be applied as a single commit.This suggestion is invalid because no changes were made to the code.Suggestions cannot be applied while the pull request is closed.Suggestions cannot be applied while viewing a subset of changes.Only one suggestion per line can be applied in a batch.Add this suggestion to a batch that can be applied as a single commit.Applying suggestions on deleted lines is not supported.You must change the existing code in this line in order to create a valid suggestion.Outdated suggestions cannot be applied.This suggestion has been applied or marked resolved.Suggestions cannot be applied from pending reviews.Suggestions cannot be applied on multi-line comments.Suggestions cannot be applied while the pull request is queued to merge.Suggestion cannot be applied right now. Please check back later.
The 16-slide San José deck now includes the completed Jev comparison, healthy classifier results and a clearer Jira → Notion → Slack attack. The example also exposes Jev through the actual Vercel AI SDK search loop, with concise implementation and model-training guides. Follow-up to #51.
Changes
--search-provider jev, metadata-onlysearch, and a six-tool staff lookup fixture. StackOne’s hosted search remains the default. Jev validates every candidate response and records resolved model, requests, tokens and elapsed time.Validation
Measurement limits
The 93.8 result is published, while Jev and BM25 were reproduced. Ranking quality is not task accuracy or latency. The selected complex attacks largely bypassed the classifier; the separately collected bars do not establish a reliable improvement. The Playground hybrid cascade and CLI combined policy differ, and neither has a matched combined Sonnet/Qwen result here. The live reviewer checkpoint still needs verification.
Keep this draft until the new Jev CLI rehearsal is recorded. Complete a timed stage rehearsal and review the exported PDF before upload. No merge or deployment is included.