项目说明:

百度2021年语言与智能技术竞赛机器阅读理解Pytorch版baseline
比赛链接:https://aistudio.baidu.com/aistudio/competition/detail/66?isFromLuge=true

官方的baseline版本是基于paddlepaddle框架的,我把它改写成了Pytorch框架,其中大部分代码沿用的是官方提供的代码,只是有一些框架部分进行了修改,另外增加了早停策略/对抗训练等优化措施,习惯用Pytorch版本的可以基于此进行优化.

环境

python=3.6
torch=1.7
transformers=4.5.0

训练示例

训练

python run.py
--max_len=256
--model_name_or_path=下载的预训练模型路径
--per_gpu_train_batch_size=7
--per_gpu_eval_batch_size=40
--learning_rate=1e-5
--linear_learning_rate=1e-4
--num_train_epochs=100
--output_dir="./output"
--weight_decay=0.01
--early_stop=2

预测

python predict.py
--max_len=400
--model_name_or_path=下载的预训练模型路径
--per_gpu_eval_batch_size=120
--output_dir="./output"
--fine_tunning_model=微调后的模型路径

实验结果

用的baseline模型是base版MacBERT(具体请看https://github.com/ymcui/MacBERT)

后续优化策略

数据清洗，据官方工作人员讲解到，训练集的准确率只能确保92%以上
更多的数据
更细粒度的数据增强
模型结构的优化

百度2021年语言与智能技术竞赛机器阅读理解Pytorch版baseline

Related tags

Overview

项目说明:

环境

训练示例

实验结果

后续优化策略

Owner

周俊贤

Keras-retinanet - Keras implementation of RetinaNet object detection.

NVIDIA Deep Learning Examples for Tensor Cores

NU-Wave: A Diffusion Probabilistic Model for Neural Audio Upsampling @ INTERSPEECH 2021 Accepted

Source code and notebooks to reproduce experiments and benchmarks on Bias Faces in the Wild (BFW).

Multi-query Video Retreival

This is the repository for Learning to Generate Piano Music With Sustain Pedals

Start-to-finish tutorial for interactive music co-creation in PyTorch and Tensorflow.js

Code to reproduce the experiments in the paper "Transformer Based Multi-Source Domain Adaptation" (EMNLP 2020)

Gif-caption - A straightforward GIF Captioner written in Python

The repository contain code for building compiler using puthon.

Universal Probability Distributions with Optimal Transport and Convex Optimization

Hypernetwork-Ensemble Learning of Segmentation Probability for Medical Image Segmentation with Ambiguous Labels

This repository contains code from the paper "TTS-GAN: A Transformer-based Time-Series Generative Adversarial Network"

An addon uses SMPL's poses and global translation to drive cartoon character in Blender.

OBG-FCN - implementation of 'Object Boundary Guided Semantic Segmentation'

AbelNN: Deep Learning Python module from scratch

Submission to Twitter's algorithmic bias bounty challenge

Implementation of the HMAX model of vision in PyTorch

A tiny, friendly, strong baseline code for Person-reID (based on pytorch).

Official code repository for the work: "The Implicit Values of A Good Hand Shake: Handheld Multi-Frame Neural Depth Refinement"