paperless-ai needs a model that holds strict JSON without being asked, because it does not send response_format. llama3.1:8b breaks the shape and qwen2.5:72b cannot answer inside its client timeout. 14B clears both. Co-Authored-By: Claude Opus 5 (1M context) <[email protected]>