Pipeline mind map
Standard pipeline template for a tabular competition. Not yet specialised by an editor.
Data processing
- Read the CSVs, find id and target
- Encode categorical columns
- Handle missing values
- Scale / transform skewed numbers
Model design
- Logistic/linear baseline
- Gradient boosting (LightGBM, XGBoost, CatBoost)
Training
- K-fold cross-validation
- Early stopping
- Save out-of-fold predictions
Post-processing
- Match the sample_submission format
- Blend models by CV
- Choose finals by CV
Metric
Categorization Accuracy. Accuracy: share of rows predicted exactly right.
Public baselines
Most-voted public notebooks, refreshed daily from the Kaggle API (last 2026-10-10).
- Spaceship Titanic with TFDF ▲8028
- 🚀 Spaceship Titanic: A complete guide 🏆 ▲2047
- 🚀Spaceship Titanic -📊EDA + 27 different models📈 ▲1073
- 🛸 Space Titanic| EDA|Advanced Feature Engineering ▲700
- Titanic Dataset ▲662
- Space Titanic 🚀 機械学習 ▲406
- 🌌🚀Spaceship Titanic-Top 6%|For Beginners 🤓🚢 ▲401
- fastai-lesson-6a-random-forests ▲387
- Spaceship Titanic Competition End To End Project ▲366
- Best models Spaceship_Titanic ▲338
- XGBoost | Wrangling with Hyperparameters | Guide ▲279
- 🚀[Pycaret] Visualization + Optimization (0.81) ▲234
Discussion 0
No discussion yet. Ask the first question or post a team-up for this competition.
Sign in with PocketPlay
Sign-in opens soon. Until then the forum is read-only.
Same account as pocketplay.win (Google or username). Posting needs a Google-linked account.
Markdown: **bold**, *italic*, `code`, ``` code blocks, - lists, > quotes, [text](https://link). No HTML.