r/LocalLLaMA · · 1 min read

Does anyone else suddenly experience unusually fast model loading times?

Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.

In the last week or so, in LM Studio, i've had models load into memory very fast for some reason, even when i'm loading them from HDD.

I know that if you just had a certain model in memory, ejected it, and then try to reload it right away, it often loads almost instantly, because, i assume, the system doesn't actually clear it out of RAM for a while, but this is not it. I have it happen with models i'm running for the very first time.

Like, just now, i finished downloading qwen 3.8 27b, ~17gb file, on HDD. I click to load it, and it took, i don't know, maybe 10-20 seconds at most? It used to take at least a minute+ for a model this size. What gives?

And it feels random too, doesn't happen every time. I tried to google if maybe there were some loading times related commits to llama.cpp recently, but don't see anything like that.

Am i the only one it happens to?

submitted by /u/iz-Moff
[link] [comments]

Discussion (0)

Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.

Sign in →

No comments yet. Sign in and be the first to say something.

More from r/LocalLLaMA