Skip to content
Sonic AI
A small model fine-tuned with Reinforcement Learning from Human Feedback (RLHF) was preferred by ..., Sonic AI