5 ms·
This works extremely well, is there a source for this to run locally?
by cglan 4y ago
This works extremely well, is there a source for this to run locally?
- gigel82 4y agohttps://github.com/antimatter15/alpaca.cpp https://github.com/antimatter15/alpaca.cpp
- cglan 4y agodoes this have the trained variant with the clean dataset used in this demo?
- sockeye 4y agoThanks! - Training code is https://github.com/tatsu-lab/stanford_alpaca https://github.com/tatsu-lab/stanford_alpaca - Params were mostly default (as in stanford_alpaca README), except for: per_device_train_batch_size=1, per_device_eval_batch_size=1 - Fine-tuning dataset was based on https://github.com/tloen/alpaca-lora/raw/81eb72f707b0505a03b7c3ff76f1a5e200a7f51d/alpaca_data_cleaned.json https://github.com/tloen/alpaca-lora/raw/81eb72f707b0505a03b... with minor improvements; I'm going to publish my version soon - The training itself took about 3 hours on 8x Nvidia A100 80GB