
We are proud to present StableVicuna, the first large-scale open source chatbot trained via reinforced learning from human feedback (RLHF). StableVicuna is a further instruction fine-tuned and RLHF-trained version of Vicuna v0 13b, which is ...
No discussion yet. Be the first to share your thoughts!