Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Shilun Li

Playing 2048 With Reinforcement Learning

Oct 20, 2021

Shilun Li, Veronica Peng

Figure 1 for Playing 2048 With Reinforcement Learning

Figure 2 for Playing 2048 With Reinforcement Learning

Figure 3 for Playing 2048 With Reinforcement Learning

Abstract:The game of 2048 is a highly addictive game. It is easy to learn the game, but hard to master as the created game revealed that only about 1% games out of hundreds million ever played have been won. In this paper, we would like to explore reinforcement learning techniques to win 2048. The approaches we have took include deep Q-learning and beam search, with beam search reaching 2048 28.5 of time.

Via

Access Paper or Ask Questions

Distributionally Robust Classifiers in Sentiment Analysis

Oct 20, 2021

Shilun Li, Renee Li, Carina Zhang

Figure 1 for Distributionally Robust Classifiers in Sentiment Analysis

Figure 2 for Distributionally Robust Classifiers in Sentiment Analysis

Figure 3 for Distributionally Robust Classifiers in Sentiment Analysis

Figure 4 for Distributionally Robust Classifiers in Sentiment Analysis

Abstract:In this paper, we propose sentiment classification models based on BERT integrated with DRO (Distributionally Robust Classifiers) to improve model performance on datasets with distributional shifts. We added 2-Layer Bi-LSTM, projection layer (onto simplex or Lp ball), and linear layer on top of BERT to achieve distributionally robustness. We considered one form of distributional shift (from IMDb dataset to Rotten Tomatoes dataset). We have confirmed through experiments that our DRO model does improve performance on our test set with distributional shift from the training set.

Via

Access Paper or Ask Questions

Ensemble ALBERT on SQuAD 2.0

Oct 19, 2021

Shilun Li, Renee Li, Veronica Peng

Figure 1 for Ensemble ALBERT on SQuAD 2.0

Figure 2 for Ensemble ALBERT on SQuAD 2.0

Figure 3 for Ensemble ALBERT on SQuAD 2.0

Abstract:Machine question answering is an essential yet challenging task in natural language processing. Recently, Pre-trained Contextual Embeddings (PCE) models like Bidirectional Encoder Representations from Transformers (BERT) and A Lite BERT (ALBERT) have attracted lots of attention due to their great performance in a wide range of NLP tasks. In our Paper, we utilized the fine-tuned ALBERT models and implemented combinations of additional layers (e.g. attention layer, RNN layer) on top of them to improve model performance on Stanford Question Answering Dataset (SQuAD 2.0). We implemented four different models with different layers on top of ALBERT-base model, and two other models based on ALBERT-xlarge and ALBERT-xxlarge. We compared their performance to our baseline model ALBERT-base-v2 + ALBERT-SQuAD-out with details. Our best-performing individual model is ALBERT-xxlarge + ALBERT-SQuAD-out, which achieved an F1 score of 88.435 on the dev set. Furthermore, we have implemented three different ensemble algorithms to boost overall performance. By passing in several best-performing models' results into our weighted voting ensemble algorithm, our final result ranks first on the Stanford CS224N Test PCE SQuAD Leaderboard with F1 = 90.123.

Via

Access Paper or Ask Questions

Trajectory Prediction using Generative Adversarial Network in Multi-Class Scenarios

Oct 18, 2021

Shilun Li, Tracy Cai, Jiayi Li

Figure 1 for Trajectory Prediction using Generative Adversarial Network in Multi-Class Scenarios

Figure 2 for Trajectory Prediction using Generative Adversarial Network in Multi-Class Scenarios

Figure 3 for Trajectory Prediction using Generative Adversarial Network in Multi-Class Scenarios

Figure 4 for Trajectory Prediction using Generative Adversarial Network in Multi-Class Scenarios

Abstract:Predicting traffic agents' trajectories is an important task for auto-piloting. Most previous work on trajectory prediction only considers a single class of road agents. We use a sequence-to-sequence model to predict future paths from observed paths and we incorporate class information into the model by concatenating extracted label representations with traditional location inputs. We experiment with both LSTM and transformer encoders and we use generative adversarial network as introduced in Social GAN to learn the multi-modal behavior of traffic agents. We train our model on Stanford Drone dataset which includes 6 classes of road agents and evaluate the impact of different model components on the prediction performance in multi-class scenes.

Via

Access Paper or Ask Questions