mirror of
https://github.com/SakanaAI/doc-to-lora.git
synced 2026-07-23 17:01:04 +02:00
* faster slice * add facts + ctx_qa * new configs * new scripts * intx_sft.py to train.py * add kaggle for downloading facts * max_new_tokens cli for eval * generate negative_nq * scripts + configs * default vals * max_val_samples_per_ds=500 * more efficient layer-to-layer ctx encoder * use_per_ctx_average_loss * faster processing * small exp distill * scripts * more robust watcher * per-module l1_norm avg * per-ctx average loss * clear_gpu |
||
|---|---|---|
| .. | ||
| small_exp | ||
| context_numbers_10.yaml | ||
| context_numbers_10_self_gen.yaml | ||
| fw_qa_v2_level_0.yaml | ||
| fw_qa_v2_level_1.yaml | ||
| fw_qa_v2_level_2.yaml | ||
| fw_qa_v2_level_3.yaml | ||
| pwc_tiny.yaml | ||
| qa_short_ctx.yaml | ||
| qa_short_ctx_compact.yaml | ||
| qa_short_ctx_self_gen.yaml | ||
| qa_short_ctx_self_gen_no_fw_qa.yaml | ||
| qa_short_ctx_self_gen_squad.yaml | ||
| squad.yaml | ||
| squad_compact.yaml | ||
| squad_fw_qa_v2_level_3.yaml | ||
| squad_hotpot_qa.yaml | ||