Twenty-one demo folders make this a pattern catalog
The root is deliberately thin: a short README points to 21 top-level demo directories, each built around a latency-sensitive judgment. The names reveal the range, including commit-sentry, jev-shell-guard, modstream, turbo-rerank, and agent-assist. There is no shared root package that turns them into one platform. You pick an example, enter its directory, and treat it as a separate application.
That shape is useful when your question is, "Where would a direct decision model fit?" It is less useful when you need a maintained SDK abstraction or a production service. The examples mix web apps, developer tools, and a macOS project. Dependencies and operating assumptions vary with each folder. The repository's value is concrete product patterns, especially cases where probabilities arrive quickly enough to alter an interface while a person is still acting.
Agent Assist asks 9 judgments in one hosted request
The measured subproject is a support console for 8 concurrent scripted chats. After each customer message, its server sends one Jev request containing 9 independent questions. Those cover a 13-way macro choice, a 10-option intent, 2 four-level scores, and 5 yes-or-no judgments. The response drives churn and escalation badges, queue ordering, and a suggested reply without asking the model to compose customer-facing text.
That boundary is the best decision in the demo. Twelve macro bodies remain client-side, and ordinary TypeScript determines refund eligibility from plan and tenure. Jev selects a macro from summaries, then code auto-fills it only when confidence reaches 0.6. A none choice sends the agent back to freehand writing. The model judges ambiguous language, while policy, display order, thresholds, and final wording stay visible in source.
What happened when we ran it
Our sandbox ran commit c469e5b from agent-assist/ because the repository has no root application. The checkout contained 728 files, about 59,458 lines of source, and used 136.4 MB. npm installed 141 packages in 22 seconds and occupied another 180 MB. The build succeeded in 13 seconds on 3 CPUs and 8 GB of RAM inside an unprivileged Node 22 container with no secrets.
Vitest completed in 7 seconds: 14 tests passed and 0 failed. npm audit found 0 known vulnerabilities at every reported severity. Our scan found 1 CI workflow, no Dockerfile, and no tests directory. The test files live beside the Agent Assist library source, so the missing directory does not mean tests are absent. The workflow is scoped to say/, however, and does not run the 14 Agent Assist tests.
Fourteen tests check policy wiring, not Jev's judgment quality
The passing suite covers useful seams. It checks the 0.6 macro gate, none handling, probability ordering, latency percentiles, refund rules, queue priority, all 9 question IDs, 12 macro choices plus none, the 6-message window, and the 8-chat, 40-message fixture. These tests guard application logic around the model and would catch several easy integration mistakes.
They do not call Jev or score its answers against labelled support data. The README describes hand-checked outcomes on the seeded scripts, including a duplicate charge, an account lockout, a cancellation threat, and a data-erasure request. That is enough for a demo walkthrough. A contact center still needs its own confusion cases, policy language, languages, and false-escalation costs before trusting the queue ordering or allowing the 1.5-second hands-free auto-send path.
Live mode sends 6 messages and customer attributes off the box
Starting the real path requires TYPESAFE_API_KEY. The key stays in the Node proxy and does not enter the browser bundle, which is the right local boundary. Each request still sends the last 6 conversation messages, the customer's plan, tenure in months, prior-ticket count, and 12 macro summaries to TypeSafe. Names are stripped from the customer object, but message bodies can contain personal or regulated data.
The server permits at most 8 Jev calls in flight and configures retries for rate limits and server errors. Without a key, /api/judge returns 503. MOCK=1 keeps interface work offline with a visible label and 5-millisecond canned delay. That mode is useful for CSS and state transitions; its keyword answers say nothing about model quality or live latency. A real evaluation needs approved sample data and a provider agreement that fits the data involved.
The claimed 93 ms median was outside our secret-free run
The Agent Assist README reports a 93 ms median panel refresh and 223 ms at the 95th percentile across 40 messages on a Linux VM. It also labels its 4-second language-model baseline as simulated. Those disclosures are better than presenting the animation as a benchmark, but neither number came from our sandbox. We measured installation, build, tests, audit, and repository signals only.
This distinction changes the buying decision. The demo supports the argument that one request can answer several typed questions and update an interface. It does not establish performance from your region, under your rate limits, with your conversation sizes. Run the same 8-chat control against the account and network you plan to use, then add real cases. Compare end-to-end paint time and answer quality, not the artificial 4-second countdown.
Seven open pull requests have not produced a release or license
GitHub showed 398 stars and 7 open items on October 5, 2026, all of them pull requests. The repository was last pushed on September 21, while three testing and documentation pull requests opened September 25. That combination shows outside work after the last main-branch update. It also leaves the proposed adversarial policy probes and cross-demo pattern guide unmerged.
There is no tagged GitHub release, and GitHub detects no license. The only workflow builds the say/ macOS app when its paths change. For evaluation code, those gaps are manageable. For copied production code, they are decision points. Agent Assist earned its place as a readable, passing example in our run; it still needs independent live results, an explicit license, and a CI path before it can serve as more than a well-made reference.

