{"podcast":{"title":"Kubernetes Podcast from Google","slug":"kubernetes-podcast-from-google","podcast_index_feed_id":860803,"rss_url":"https://rss.libsyn.com/shows/419861/destinations/3486674.xml","website_url":"https://kubernetespodcast.com","image_url":"https://static.libsyn.com/p/assets/a/3/a/a/a3aa4f08236059f5e55e3c100dce7605/NewKPodRoboWhite-20240924-1zwoxpn1sl.png","author":"Kubernetes Podcast from Google","episode_count":265,"summary":"A biweekly podcast focused on what's happening in the Kubernetes community hosted by Abdel Sghiouar and Kaslin Fields. We cover Kubernetes, cloud-native applications, and other developments in the ecosystem. Abdel and Kaslin on Twitter at @KubernetesPod or by email at kubernetespodcast@google.com.","last_synced_at":null,"page_url":"https://stenobird.com/podcast/kubernetes-podcast-from-google"},"episode":{"title":"Kubernetes at LinkedIn, with Ahmet Alp Balkan and Ronak Nathani","slug":"kubernetes-at-linkedin-with-ahmet-alp-balkan-and-ronak-nathani","published_at":"2025-03-25T16:06:00+00:00","page_url":"https://stenobird.com/podcast/kubernetes-podcast-from-google/kubernetes-at-linkedin-with-ahmet-alp-balkan-and-ronak-nathani","show_page_url":"https://stenobird.com/podcast/kubernetes-podcast-from-google","url":"https://e780d51f-f115-44a6-8252-aed9216bb521.libsyn.com/kubernetes-at-linkedin-with-ahmet-alp-balkan-and-ronak-nathani","audio_url":"https://traffic.libsyn.com/secure/e780d51f-f115-44a6-8252-aed9216bb521/KPod249.mp3?dest-id=3486674","summary":"LinkedIn engineers reveal how they manage massive Kubernetes clusters on bare metal, focusing on handling latency-sensitive workloads and hardware fragmentation. The discussion explores the complexities of building an opinionated platform that abstracts infrastructure for developers while maintaining high performance.","meta_description":"Learn how LinkedIn runs Kubernetes at scale on bare metal, managing hardware skews, stateful workloads, and custom controllers for high-performance apps.","key_points":["Main idea: LinkedIn uses a custom, opinionated platform on bare metal to manage hardware generations and prevent fragmentation","Practical takeaway: Use node profiles and specific pools to isolate latency-sensitive applications from older hardware generations","Failure mode: Large-scale clusters can cause controller informers to fail to sync, potentially leading to catastrophic state loss like label clearing","Main idea: Kubernetes can be adapted for stateful workloads if you control the full stack from bare metal to disk attachment","Practical takeaway: Abstracting Kubernetes complexity via custom APIs helps developers focus on resources like CPU and memory rather than low-level networking"],"chapters":[{"start_ms":60000,"title":"Kubernetes Ecosystem News","summary":"Updates on CubeFS graduation to CNCF and Canonical's announcement regarding 12-year Kubernetes Long Term Support."},{"start_ms":250000,"title":"Running Kubernetes on Bare Metal","summary":"An exploration of LinkedIn's transition from legacy containerization to a large-scale Kubernetes implementation on bare metal."},{"start_ms":625000,"title":"Networking and Flat Data Center Architectures","summary":"Discussion on bypassing standard CNI plugins like Flannel in favor of a flat, routable data center network."},{"start_ms":1000000,"title":"Managing Hardware Skews and Fragmentation","summary":"How to handle different CPU generations and the challenges of preventing resource fragmentation in multi-tenant pools."},{"start_ms":1370000,"title":"The Power of Custom Controllers","summary":"The necessity and reality of developing custom Kubernetes controllers and CRDs to extend platform functionality."},{"start_ms":1935000,"title":"Abstracting Complexity for Developers","summary":"Providing an opinionated interface that allows engineers to specify compute resources without managing low-level Kubernetes primitives."},{"start_ms":2325000,"title":"Lessons from Large-Scale Failures","summary":"A deep dive into how controller informer timeouts in massive clusters can lead to unintended side effects like mass label deletion."}],"topics":["Kubernetes","Bare Metal","Infrastructure at Scale","Cloud Native","Custom Controllers","LinkedIn Engineering","Site Reliability Engineering","Container Orchestration"],"duration_seconds":2522,"processing_state":"processed","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/kubernetes-podcast-from-google/episodes/kubernetes-at-linkedin-with-ahmet-alp-balkan-and-ronak-nathani/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/kubernetes-podcast-from-google/kubernetes-at-linkedin-with-ahmet-alp-balkan-and-ronak-nathani.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}