-
LoRA Fine-tuning Efficiently Undoes Safety Training in Llama 2-Chat 70B
Paper • 2310.20624 • Published • 13 -
Unleashing the Power of Pre-trained Language Models for Offline Reinforcement Learning
Paper • 2310.20587 • Published • 18 -
BadLlama: cheaply removing safety fine-tuning from Llama 2-Chat 13B
Paper • 2311.00117 • Published -
VideoFusion: Decomposed Diffusion Models for High-Quality Video Generation
Paper • 2303.08320 • Published • 3
Vikarti Anatra
vikarti-anatra
·
AI & ML interests
None yet
Recent Activity
liked
a model
2 days ago
mradermacher/Qwen3-30B-A6B-16-Extreme-GGUF
liked
a model
3 days ago
Gryphe/Pantheon-Proto-RP-1.8-30B-A3B
liked
a model
6 days ago
unsloth/Qwen3-30B-A3B-GGUF
Organizations
None yet
Collections
2
Various quants of Faro-Yi-9B-DPO
models
3
datasets
0
None public yet