Skip to content

Stabilize per-item local inference - #242

Merged
djdefi merged 1 commit into
mainfrom
fix/llama-server-stability
Aug 9, 2026
Merged

djdefi merged 1 commit into
mainfrom
fix/llama-server-stability

Conversation

@djdefi

@djdefi djdefi commented Aug 9, 2026

Copy link
Copy Markdown
Owner

Runs llama-server with one inference slot to avoid the observed concurrent-request crash, falls back to complete source sentences after configured-server errors, and refuses to deploy if llama-server dies or summary error placeholders reach the HTML. All 81 tests and 275 assertions pass.

Co-authored-by: Copilot <223556219+Copilot@users.noreply.github.com>

Copilot-Session: 17590816-0a13-4b19-990e-8aae72213acc
Comment thread test/render_test.rb
true
end

def test_generate_grounded_facts_uses_source_fallback_after_local_server_error
@djdefi
djdefi merged commit efb5ce8 into main Aug 9, 2026
5 checks passed
@djdefi
djdefi deleted the fix/llama-server-stability branch August 9, 2026 00:17
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

2 participants