Episode

How Data Scientists Use Gradient Boosting for Tabular Data

Podcast
The Data Science Podcast with Fexingo: Analytics, Machine Learning, and Data-Driven Conversations
Published
Jul 11, 2026
Duration seconds
551
Processing state
not_requested
Canonical source
https://audio.fexingo.com/business/the-data-science-podcast/episode-0104.mp3
Audio
https://audio.fexingo.com/business/the-data-science-podcast/episode-0104.mp3
JSON
/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-gradient-boosting-for-tabular-data
Markdown
/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-gradient-boosting-for-tabular-data.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-gradient-boosting-for-tabular-data/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-gradient-boosting-for-tabular-data.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

A deep dive into the enduring power of gradient boosting machines (GBMs) for structured, tabular data—the bread and butter of most real-world data science. Lucas and Luna explore why gradient boosting consistently wins Kaggle competitions and beats deep learning on many business problems. They break down the core mechanics: sequential tree-building, learning rate, and regularization. The episode focuses on a case study from a mid-size e-commerce company that used XGBoost to reduce customer churn prediction error by 18% year-over-year. They also discuss modern variants like LightGBM and CatBoost, and when to choose each. Practical guidance on hyperparameter tuning and common pitfalls (overfitting, categorical encoding) grounds the conversation in daily data-science work. Listeners will walk away understanding why gradient boosting remains a must-have in any data scientist's toolkit, especially for data with mixed data types and missing values. #GradientBoosting #XGBoost #LightGBM #CatBoost #TabularData #MachineLearning #DataScience #Kaggle #HyperparameterTuning #ChurnPrediction #EnsembleMethods #DecisionTrees #Regularization #FeatureEngineering #BusinessAnalytics #Technology #FexingoBusiness #BusinessPodcast Keep every episode free: buymeacoffee.com/fexingo