doc-to-lora/configs/small_exp
Rujikorn Charakorn 891c0bd256
Refactor_and_improve_data (#7)
* faster slice

* add facts + ctx_qa

* new configs

* new scripts

* intx_sft.py to train.py

* add kaggle for downloading facts

* max_new_tokens cli for eval

* generate negative_nq

* scripts + configs

* default vals

* max_val_samples_per_ds=500

* more efficient layer-to-layer ctx encoder

* use_per_ctx_average_loss

* faster processing

* small exp distill

* scripts

* more robust watcher

* per-module l1_norm avg

* per-ctx average loss

* clear_gpu
2025-08-04 20:52:37 +09:00
..
qa_short_ctx_self_gen_lv1_closed_qa_1.yaml Refactor_and_improve_data (#7) 2025-08-04 20:52:37 +09:00
qa_short_ctx_self_gen_lv1_closed_qa_1_and_lv3.yaml Refactor_and_improve_data (#7) 2025-08-04 20:52:37 +09:00
qa_short_ctx_self_gen_lv1_closed_qa_1_l2l.yaml Refactor_and_improve_data (#7) 2025-08-04 20:52:37 +09:00
qa_short_ctx_self_gen_lv1_closed_qa_1_next_token_pred.yaml Refactor_and_improve_data (#7) 2025-08-04 20:52:37 +09:00