{"podcast":{"title":"UpNext AI","slug":"upnext-ai-7846034","podcast_index_feed_id":7846034,"rss_url":"https://feeds.transistor.fm/upnext-ai","website_url":"https://www.upnext.fm","image_url":"https://img.transistorcdn.com/X_pKGwRFM_7u9_tArrtTE5GdY29L3xdGHcVeceTkhfE/rs:fill:0:0:1/w:1400/h:1400/q:60/mb:500000/aHR0cHM6Ly9pbWct/dXBsb2FkLXByb2R1/Y3Rpb24udHJhbnNp/c3Rvci5mbS85YWE0/MDRlYTJkZTY4ZDhk/N2YzZjZmYzg0ZmQ2/Mzg1NC5wbmc.jpg","author":"Matthew McMaster","episode_count":60,"summary":"Daily AI news and research, distilled. UpNext AI breaks down the most important developments in artificial intelligence—from major industry moves to cutting-edge papers.","last_synced_at":"2026-07-23T22:18:34.587279+00:00","page_url":"https://stenobird.com/podcast/upnext-ai-7846034"},"episode":{"title":"Kimi K3, AI Travel’s Unicorn Moment, and the Reliability Problem in AI Benchmarks | UpNext AI – July 17, 2026","slug":"kimi-k3-ai-travel-s-unicorn-moment-and-the-reliability-problem-in-ai-benchmarks-upnext-ai-july-17-2026","published_at":"2026-07-17T11:31:11+00:00","page_url":"https://stenobird.com/podcast/upnext-ai-7846034/kimi-k3-ai-travel-s-unicorn-moment-and-the-reliability-problem-in-ai-benchmarks-upnext-ai-july-17-2026","show_page_url":"https://stenobird.com/podcast/upnext-ai-7846034","url":"https://share.transistor.fm/s/f5db7b47","audio_url":"https://media.transistor.fm/f5db7b47/878dd21d.mp3","summary":"A quick end-of-week catch-up on the AI stories that matter most. Today: Moonshot AI’s new Kimi K3 model makes a big open-model play on size, price, and coding performance; AI-powered travel startup Fora hits unicorn status with a fresh round; and a new research paper questions whether a popular benchmark scoring method can really be trusted. Covered in this episode: - Kimi K3 launches as Moonshot AI’s most capable model to date, with 2.8 trillion parameters and an open-weight release promised by July 27 - AI-powered travel agency Fora raises a $60 million Series D at a $1 billion valuation - New arXiv research asks whether item response theory is reliable for ranking models and interpreting AI benchmarks - Google renames NotebookLM to Gemini Notebook and adds code execution for data analysis - Thinking Machines Lab releases Inkling, its first open-weights model - Netflix says around 300 titles on its platform used generative AI, mostly in post-production - The EU orders Google to share search data and open up AI on Android under the Digital Markets Act Source links: - https://simonwillison.net/2026/Jul/16/kimi-k3/#atom-everything - https://techcrunch.com/2026/07/16/ai-powered-travel-agency-fora-hits-unicorn-status-raises-60m/ - https://arxiv.org/abs/2607.15190v1 - https://techcrunch.com/2026/07/16/google-continues-its-renaming-streak-by-turning-notebooklm-to-gemini-notebook/ - https://simonwillison.net/2026/Jul/16/inkling/#atom-everything - https://www.theverge.com/streaming/966633/netflix-ai-titles-q2-2026-earnings - https://arstechnica.com/gadgets/2026/07/its-official-eu-will-force-google-to-share-search-data-and-open-up-ai-on-android/","meta_description":"A quick end-of-week catch-up on the AI stories that matter most. Today: Moonshot AI’s new Kimi K3 model makes a big open-model play on size, price, and co…","key_points":[],"chapters":[],"topics":[],"duration_seconds":429,"processing_state":"not_requested","actions":[{"name":"request_transcript","method":"POST","url":"https://stenobird.com/v1/public/podcasts/upnext-ai-7846034/episodes/kimi-k3-ai-travel-s-unicorn-moment-and-the-reliability-problem-in-ai-benchmarks-upnext-ai-july-17-2026/transcription-requests","description":"Idempotently request low-priority transcript generation for this episode."},{"name":"read_markdown","method":"GET","url":"https://stenobird.com/podcast/upnext-ai-7846034/kimi-k3-ai-travel-s-unicorn-moment-and-the-reliability-problem-in-ai-benchmarks-upnext-ai-july-17-2026.md","description":"Read the agent-friendly Markdown representation of this episode resource."}]}}