▲ 513 ▼ What could possibly go wrong (lemmy.world) submitted 2 years ago by simplejack@lemmy.world to c/politicalmemes@lemmy.world 73 comments fedilink hide all child comments
[–] cybervseas@lemmy.world 56 points 2 years ago (3 children) It's open source. Apparently folks have already made mods of it that add CCP-sensitive info back in. Disclaimer: I have yet to see this for myself. permalink fedilink source hideshow 6 child comments replies: [–] Even_Adder@lemmy.dbzer0.com 53 points 2 years ago (2 children) The answer I got out of DeepSeek-R1-Distill-Llama-8B-abliterate.i1-Q4_K_S permalink fedilink source parent hideshow 4 child comments replies: [–] taiyang@lemmy.world 28 points 2 years ago So a real answer, basically. Too bad your average person isn't going to bother with that. Still nice it's open source. permalink fedilink source parent [–] felixwhynot@lemmy.world 11 points 2 years ago (1 child) Seems like the model you mentioned is more like a fine tuned Llama? Specifically, these are fine-tuned versions of Qwen and Llama, on a dataset of 800k samples generated by DeepSeek R1. https://github.com/Emericen/deepseek-r1-distilled permalink fedilink source parent hideshow 2 child comments replies: [–] Even_Adder@lemmy.dbzer0.com 8 points 2 years ago* (2 children) Yeah, it's distilled from deepseek and abliterated. The non-abliterated ones give you the same responses as Deepseek R1. permalink fedilink source parent hideshow 4 child comments replies: [–] obre@lemmy.world 7 points 2 years ago permalink fedilink source parent [+] obre@lemmy.world 1 point 2 years ago [deleted] permalink fedilink source parent [–] perviouslyiner@lemmy.world 4 points 2 years ago* (last edited 2 years ago) just running it locally, apparently. The output of this model is being filtered by another AI, but only on the public-hosted copy. permalink fedilink source parent [+] stebo02@lemmy.dbzer0.com 2 points 2 years ago* (last edited 1 year ago) [deleted] permalink fedilink source parent
[–] Even_Adder@lemmy.dbzer0.com 53 points 2 years ago (2 children) The answer I got out of DeepSeek-R1-Distill-Llama-8B-abliterate.i1-Q4_K_S permalink fedilink source parent hideshow 4 child comments replies: [–] taiyang@lemmy.world 28 points 2 years ago So a real answer, basically. Too bad your average person isn't going to bother with that. Still nice it's open source. permalink fedilink source parent [–] felixwhynot@lemmy.world 11 points 2 years ago (1 child) Seems like the model you mentioned is more like a fine tuned Llama? Specifically, these are fine-tuned versions of Qwen and Llama, on a dataset of 800k samples generated by DeepSeek R1. https://github.com/Emericen/deepseek-r1-distilled permalink fedilink source parent hideshow 2 child comments replies: [–] Even_Adder@lemmy.dbzer0.com 8 points 2 years ago* (2 children) Yeah, it's distilled from deepseek and abliterated. The non-abliterated ones give you the same responses as Deepseek R1. permalink fedilink source parent hideshow 4 child comments replies: [–] obre@lemmy.world 7 points 2 years ago permalink fedilink source parent [+] obre@lemmy.world 1 point 2 years ago [deleted] permalink fedilink source parent
[–] taiyang@lemmy.world 28 points 2 years ago So a real answer, basically. Too bad your average person isn't going to bother with that. Still nice it's open source. permalink fedilink source parent
[–] felixwhynot@lemmy.world 11 points 2 years ago (1 child) Seems like the model you mentioned is more like a fine tuned Llama? Specifically, these are fine-tuned versions of Qwen and Llama, on a dataset of 800k samples generated by DeepSeek R1. https://github.com/Emericen/deepseek-r1-distilled permalink fedilink source parent hideshow 2 child comments replies: [–] Even_Adder@lemmy.dbzer0.com 8 points 2 years ago* (2 children) Yeah, it's distilled from deepseek and abliterated. The non-abliterated ones give you the same responses as Deepseek R1. permalink fedilink source parent hideshow 4 child comments replies: [–] obre@lemmy.world 7 points 2 years ago permalink fedilink source parent [+] obre@lemmy.world 1 point 2 years ago [deleted] permalink fedilink source parent
[–] Even_Adder@lemmy.dbzer0.com 8 points 2 years ago* (2 children) Yeah, it's distilled from deepseek and abliterated. The non-abliterated ones give you the same responses as Deepseek R1. permalink fedilink source parent hideshow 4 child comments replies: [–] obre@lemmy.world 7 points 2 years ago permalink fedilink source parent [+] obre@lemmy.world 1 point 2 years ago [deleted] permalink fedilink source parent
[–] perviouslyiner@lemmy.world 4 points 2 years ago* (last edited 2 years ago) just running it locally, apparently. The output of this model is being filtered by another AI, but only on the public-hosted copy. permalink fedilink source parent
[+] stebo02@lemmy.dbzer0.com 2 points 2 years ago* (last edited 1 year ago) [deleted] permalink fedilink source parent