← 开源
openai

finetune-transformer-lm

Code and model for the paper "Improving Language Understanding by Generative Pre-Training"

Model DevelopmentFine-tuningArchitecturePython
在 GitHub 打开
增长势头
+024 小时新增 Star0.0%
2.31k
Star
525
Fork
+2
本周
3
贡献者
创建于 2018-06-11 · 更新于 2026-10-03 · 今日第 8613 名
主要开发者
README

Status: Archive (code is provided as-is, no updates expected)

finetune-transformer-lm

Code and model for the paper "Improving Language Understanding by Generative Pre-Training"

Currently this code implements the ROCStories Cloze Test result reported in the paper by running: python train.py --dataset rocstories --desc rocstories --submit --analysis --data_dir [path to data here]

Note: The code is currently non-deterministic due to various GPU ops. The median accuracy of 10 runs with this codebase (using default hyperparameters) is 85.8% - slightly lower than the reported single run of 86.5% from the paper.

The ROCStories dataset can be downloaded from the associated website.