jackaduma/Vicuna-LoRA-RLHF-PyTorch
A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Vicuna architecture. Basically ChatGPT but with Vicuna
Join the conversation
Reviews · Questions · Posts
Share what you know about Vicuna-LoRA-RLHF-PyTorch — write a review from your real experience, ask an implementation question, or publish a post about how you use it.
Share your experience
Write or update your review
Explain what worked, what broke down, and what another team should know before adopting Vicuna-LoRA-RLHF-PyTorch.
Project Q&A
Questions and answers
Browse implementation threads tied directly to jackaduma/Vicuna-LoRA-RLHF-PyTorch. Each question links through to the full answer page.
Be the first to ask how teams run Vicuna-LoRA-RLHF-PyTorch in production. Every question you post becomes a durable, searchable answer page other developers can find.
Ask the first questionRelated posts
Posts tagged with the same topics
These posts come from the same topic surface as this repo, so readers can move from project evaluation into practical writeups and migration notes without leaving context.
Share how your team uses Vicuna-LoRA-RLHF-PyTorch — a migration note, an architecture writeup, or a comparison. Your post reaches everyone browsing these same topics.
Write the first post