Pipeline mind map
Standard pipeline template for a text competition. Not yet specialised by an editor.
Data processing
- Clean and tokenize text
- Check label balance
Model design
- TF-IDF + linear baseline
- Pretrained transformer
Training
- Folds by label
- Small learning rate, few epochs
Post-processing
- Calibrate / threshold for the metric
- Ensemble seeds
Metric
F-Score (Micro).
Public baselines
Most-voted public notebooks, refreshed daily from the Kaggle API (last 2026-10-10).
- NLP Getting Started Tutorial ▲5052
- NLP with Disaster Tweets - EDA, Cleaning and BERT ▲4521
- Knowledge Graph & NLP Tutorial-(BERT,spaCy,NLTK) ▲2899
- Basic EDA,Cleaning and GloVe ▲2837
- KerasNLP starter notebook Disaster Tweets ▲2500
- NLP 📝 GloVe, BERT, TF-IDF, LSTM... 📝 Explained ▲1615
- NLP - EDA, Bag of Words, TF IDF, GloVe, BERT ▲1471
- Natural Language Processing (NLP) for Beginners ▲1430
- Disaster NLP: Keras BERT using TFHub ▲1051
- Useful Python libraries for Data Science ▲793
- AutoML Getting Started Notebook ▲722
- A Real Disaster - Leaked Label ▲633
Discussion 0
No discussion yet. Ask the first question or post a team-up for this competition.
Sign in with PocketPlay
Sign-in opens soon. Until then the forum is read-only.
Same account as pocketplay.win (Google or username). Posting needs a Google-linked account.
Markdown: **bold**, *italic*, `code`, ``` code blocks, - lists, > quotes, [text](https://link). No HTML.