af9da3c912e5bf8a2799f8b1263df3a842a07f8b
Align the provider's HTTP client timeout with the orchestrator per-stage timeout (1_800_000ms). A slow local reasoning model with a large prompt and a multi-thousand-token reasoning budget can exceed 10min on a single call; the old 600s HTTP timeout failed the inference before the stage timeout applied, exhausting retries on timeouts alone.
fix: workflow inference chain — surface llama-server errors, user-turn-last, bundle-relative prompts
Description
No description provided
Languages
Kotlin
88.4%
Go
11.4%
Python
0.2%