LLM Reddit Apr 9, 2026 1 min read LocalLLaMA, Qwen 3.5 chat template bug가 prefix-cache reuse를 조용히 무너뜨린다고 지적 r/LocalLLaMA의 debugging post는 Qwen 3.5의 chat template 문제가 tool-heavy turn 뒤 prefix-cache reuse를 깨뜨려 대량의 불필요한 recomputation을 만들 수 있다고 주장한다. #qwen-3.5#prefix-caching#chat-template 21