跳到正文
r/LocalLLaMA· /u/AdFickle8681·· 3 小时前AI 评分26

如何判断一个社区微调模型是否可信?

How do you decide whether to trust a community fine-tune?

AI 导读

一位研究者正在调研人们如何挑选和审查 Hugging Face 上的微调与合并模型,提出四个问题:在哪里发现想试的模型、使用前会检查什么(benchmark、model card、评价还是自己的测试提示词)、是否遇到过微调后表现反而不如基座模型的情况(如异常拒答、推理能力丢失、奇怪输出),以及如果存在与基座模型快速对比的功能是否会使用、需要展示什么。

正文

I'm researching how people choose and vet fine-tunes and merges from Hugging Face. I'm not selling anything. I just want to understand what people actually do.
1. Where do you find the models you try?
2. What do you check before you start using one (benchmarks, model card, reviews, your own test prompts)?
3. Has a fine-tune ever behaved worse than its base model? For example, odd refusals, lost reasoning, strange outputs, or things it should not say. What happened?
4. If a quick side-by-side check of a download against its base model existed, would you use it? What would it need to show?
Short answers are great, and stories are even better. Thanks!

submitted by /u/AdFickle8681
[link] [留言]

来源:r/LocalLLaMA · reddit.com