Three live Jev requests on synthetic architecture examples, 2026-10-01. These demonstrate integration, not general recommendation accuracy.
The numbers below are uncalibrated model choice probabilities. A value of 1 does not mean guaranteed correctness. Statistical confidence intervals are unavailable: no independent labeled evaluation set was supplied.
Preserve eight application instances and the existing database connection limit. Prefer changing the application connection-pool setting if it satisfies the declared workload.
Selected: smaller-pools. Model: jev-1.13.0. Calls: 1.
{
"fewer-instances": 0,
"smaller-pools": 1
}Original proposal: FAIL (introduced). Both eligible alternatives pass the declared scenarios. No changes were applied.
Satisfy the declared workload while retaining the existing two-vCPU host size. Prefer adding an instance over changing host size.
Selected: more-instances. Model: jev-1.13.0. Calls: 1.
{
"larger-host": 0,
"more-instances": 1
}Original proposal: FAIL (introduced). Both eligible alternatives pass the declared scenarios. No changes were applied.
Cover the declared 21-second lag, choosing the smallest offered passing retention allocation. Do not assume memory requirements are verified.
Selected: retain-240. Model: typesafe/jev-1.13-20260917. Calls: 1.
{
"retain-240": 1,
"retain-480": 0
}Original proposal: FAIL (introduced). Both eligible alternatives pass the declared scenarios. No changes were applied.
CPU inference with outbound socket connections blocked completed with a 180-second limit after a 90-second startup timeout.
Laya selected fewer-instances (0.8544), versus smaller-pools (0.1456). This is a passing modeled alternative but a poor match for the objective to preserve eight instances. Integration succeeded; this example does not demonstrate good goal alignment.
The browser and general English checkpoints received identical objective, candidate descriptions, and compact state. Both ran locally on CPU.
| Checkpoint | Smaller pools | Fewer instances |
|---|---|---|
| Browser | 14.56% | 85.44% (selected) |
| General English | 39.41% | 60.59% (selected) |
Both selections conflict with the objective to preserve eight instances. A greater probability for the preferred choice did not improve the selected answer. One case does not establish general model quality.