GPT2-chitchat icon indicating copy to clipboard operation
GPT2-chitchat copied to clipboard

更换更大的gpt2模型进行训练 如:gpt2_large

Open htthYjh opened this issue 2 years ago • 1 comments

楼主或者其他哥们有没有使用更大的gpt2模型进行训练,求分享经验!!!!! 我使用如下配置config.json进行训练,但是测试结果比较差:求指教 { "_name_or_path": "model/epoch41", "activation_function": "gelu_new", "architectures": [ "GPT2LMHeadModel" ], "attn_pdrop": 0.1, "bos_token_id": 50256, "embd_pdrop": 0.1, "eos_token_id": 50256, "gradient_checkpointing": false, "id2label": { "0": "LABEL_0" }, "initializer_range": 0.02, "label2id": { "LABEL_0": 0 }, "layer_norm_epsilon": 1e-05, "model_type": "gpt2", "n_ctx": 1024, "n_embd": 1280, "n_head": 20, "n_inner": null, "n_layer": 48, "n_positions": 1024, "output_past": true, "resid_pdrop": 0.1, "summary_activation": null, "summary_first_dropout": 0.1, "summary_proj_to_labels": true, "summary_type": "cls_index", "summary_use_proj": true, "transformers_version": "4.18.0", "use_cache": true, "task_specific_params": { "text-generation": { "max_length": 1000 } }, "vocab_size": 50005 }

htthYjh avatar May 08 '22 03:05 htthYjh

你好,如果不想他自己下载,应该在哪下载model/epoch41这个文件呢

1zhangtianqing avatar May 17 '23 09:05 1zhangtianqing