e97e8b1d33
The first server deployment downloaded only 1101 bytes (a redirect/error page) and the entrypoint promoted it to ggml-large-v3.bin, after which every restart skipped the download and the server ran without a valid model. Now both fresh downloads and existing files are validated (minimum 50 MB + the ggml magic bytes); invalid files are logged, deleted, and re-downloaded, and a failed download dumps the first 400 bytes of what was actually received for diagnosis before exiting. Validated locally: a poisoned model file is detected, removed, and a valid one re-downloaded; inference served correctly afterwards.