Episode

How Data Scientists Use Reinforcement Learning for Dynamic Pricing

Podcast
The Data Science Podcast with Fexingo: Analytics, Machine Learning, and Data-Driven Conversations
Published
Jun 15, 2026
Duration seconds
548
Processing state
not_requested
Canonical source
https://audio.fexingo.com/business/the-data-science-podcast/episode-0052.mp3
Audio
https://audio.fexingo.com/business/the-data-science-podcast/episode-0052.mp3
JSON
/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-reinforcement-learning-for-dynamic-pricing
Markdown
/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-reinforcement-learning-for-dynamic-pricing.md

Actions

  • POST https://stenobird.com/v1/public/podcasts/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/episodes/how-data-scientists-use-reinforcement-learning-for-dynamic-pricing/transcription-requests
    Idempotently request low-priority transcript generation for this episode.
  • GET https://stenobird.com/podcast/the-data-science-podcast-with-fexingo-analytics-machine-learning-and-data-driven-conversations-7871831/how-data-scientists-use-reinforcement-learning-for-dynamic-pricing.md
    Read the agent-friendly Markdown representation of this episode resource.

Summary

In this episode of The Data Science Podcast, Lucas and Luna explore how reinforcement learning (RL) is transforming dynamic pricing strategies. Using the example of a major ride-hailing company, they break down how RL algorithms learn to set prices in real time by balancing exploration (testing new price points) and exploitation (using known optimal prices). Lucas explains the core RL concepts of state, action, reward, and the epsilon-greedy algorithm. Luna digs into the practical trade-offs: how often should a model explore versus exploit, and why the reward function must account for long-term customer retention, not just immediate revenue. The conversation also touches on how RL differs from A/B testing in dynamic pricing, the role of simulation environments for training, and ethical considerations around price discrimination. Listeners will walk away with a concrete understanding of RL-based pricing mechanics and a mental model to evaluate pricing algorithms they encounter daily. #ReinforcementLearning #DynamicPricing #DataScience #MachineLearning #RL #PricingStrategy #RideHailing #ExplorationExploitation #EpsilonGreedy #RewardFunction #AIBusiness #Technology #TechPodcast #DataDriven #FexingoBusiness #BusinessPodcast #DataSciencePodcast #Fexingo Keep every episode free: buymeacoffee.com/fexingo