Appreciation Post - thomsonreuters/Thomson-1.0-Small
Mirrored from r/LocalLLaMA for archival readability. Support the source by reading on the original site.
With the lack of support from Qwen regarding the smaller 9B and 35B MOE models. Like myself, not everyone is looking for an agentic coding model, I particularly use it for RAG and reviewing and require high reasoning across different documents & came across this Finetune:
thomsonreuters/Thomson-1.0-Small · Hugging Face
It's not MTP, but i have a 9070xt and it still runs fairly well (~25 T/s) but the quality is there.
[link] [comments]
Discussion (0)
Sign in to join the discussion. Free account, 30 seconds — email code or GitHub.
Sign in →No comments yet. Sign in and be the first to say something.