<?xml version="1.0" encoding="UTF-8"?><rss xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:atom="http://www.w3.org/2005/Atom" version="2.0" xmlns:itunes="http://www.itunes.com/dtds/podcast-1.0.dtd" xmlns:psc="http://podlove.org/simple-chapters" xmlns:podcast="https://podcastindex.org/namespace/1.0"><channel><title><![CDATA[This Week in Open Lakehouse]]></title><description><![CDATA[<p>This Week in Open Lakehouse is a weekly data engineering podcast discussing the latest in open source news, including query engines, data catalogs, open table formats, orchestrators, and streaming technologies. Produced by the Databricks DevRel team. <a href="https://openlakehouse.io/" rel="noopener noreferrer nofollow" target="_blank">https://openlakehouse.io/</a></p>]]></description><link>openlakehouse.io</link><generator>Riverside.fm (https://riverside.com)</generator><lastBuildDate>Fri, 02 Oct 2026 03:19:25 GMT</lastBuildDate><atom:link href="https://api.riverside.com/hosting/aeXB70YD.rss" rel="self" type="application/rss+xml"/><author><![CDATA[Lisa Cao and Scott Haines]]></author><pubDate>Fri, 18 Sep 2026 21:41:51 GMT</pubDate><copyright><![CDATA[2026 Lisa Cao and Scott Haines]]></copyright><language><![CDATA[en]]></language><ttl>60</ttl><category><![CDATA[Technology]]></category><category><![CDATA[Tech News]]></category><itunes:author>Lisa Cao and Scott Haines</itunes:author><itunes:summary>&lt;p&gt;This Week in Open Lakehouse is a weekly data engineering podcast discussing the latest in open source news, including query engines, data catalogs, open table formats, orchestrators, and streaming technologies. Produced by the Databricks DevRel team. &lt;a href=&quot;https://openlakehouse.io/&quot; rel=&quot;noopener noreferrer nofollow&quot; target=&quot;_blank&quot;&gt;https://openlakehouse.io/&lt;/a&gt;&lt;/p&gt;</itunes:summary><itunes:type>episodic</itunes:type><itunes:owner><itunes:name>Lisa Cao and Scott Haines</itunes:name><itunes:email>carly.akerly@databricks.com</itunes:email></itunes:owner><itunes:explicit>no</itunes:explicit><itunes:category text="Technology"/><itunes:category text="News"><itunes:category text="Tech News"/></itunes:category><itunes:image href="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/logos/e0e4749a-7ff6-4101-9540-5e11b292b86c.jpeg"/><item><title><![CDATA[XML Functions, Polars 2.0, delta-rs, and Agent Thrashing Policies ]]></title><description><![CDATA[<p>In this episode, Lisa and Scott explore recent updates in ecosystem technologies including Apache Iceberg, Apache Flink, Polars, delta-rs, highlighting their frustrations in agent workflows and a potential oversight on the importance of XML. </p><ul><li>Iceberg 1.12: REST catalogs, C++ support, equality deletes, and interoperability</li><li>Polaris catalog governance, lineage, pagination, and deletion policies</li><li>Flink updates, including State Fun deprecation and native XML functions</li><li>Polars 2.0, streaming execution, lazy frames, and memory efficiency</li><li>Delta Rust releases, partition handling, correctness fixes, and migration</li><li>Parquet performance, FastLanes benchmarks, and Arrow’s Intel macOS support</li><li>AI agent orchestration with Omnigent, Kubernetes sandboxes, Spark pipelines, and thrash detection</li></ul>]]></description><guid isPermaLink="false">2b036625-518b-4af8-9c98-374e3ec77aa6</guid><dc:creator><![CDATA[Lisa Cao and Scott Haines]]></dc:creator><pubDate>Tue, 29 Sep 2026 18:00:30 GMT</pubDate><enclosure url="https://api.riverside.com/hosting-analytics/media/3d40256a3a24d53e4276d39c5b4e36e87f61dd17aba8519b929246a1ea2a07a3/eyJlcGlzb2RlSWQiOiIyYjAzNjYyNS01MThiLTRhZjgtOWM5OC0zNzRlM2VjNzdhYTYiLCJwb2RjYXN0SWQiOiIzNGM1ZTgxYS03ODY2LTQxZDQtYTgwMy1iYjhiNmRkMWMyMjYiLCJhY2NvdW50SWQiOiI2N2EwYWU5OTU0OGViN2E4MGYyZDU3NDQiLCJwYXRoIjoibWVkaWEvY2xpcHMvNmFiYmZjYmY1NWJkYzBiMDRiZWZiZTVlL29wZW4tbGFrZWhvdXNlLS1haS0yMDI2LTktMjlfXzE4LTAtMzkubXAzIn0=.mp3" length="96464605" type="audio/mpeg"/><podcast:transcript url="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/episodes/2b036625-518b-4af8-9c98-374e3ec77aa6/transcripts.txt" type="text/plain"/><itunes:summary>&lt;p&gt;In this episode, Lisa and Scott explore recent updates in ecosystem technologies including Apache Iceberg, Apache Flink, Polars, delta-rs, highlighting their frustrations in agent workflows and a potential oversight on the importance of XML. &lt;/p&gt;&lt;ul&gt;&lt;li&gt;Iceberg 1.12: REST catalogs, C++ support, equality deletes, and interoperability&lt;/li&gt;&lt;li&gt;Polaris catalog governance, lineage, pagination, and deletion policies&lt;/li&gt;&lt;li&gt;Flink updates, including State Fun deprecation and native XML functions&lt;/li&gt;&lt;li&gt;Polars 2.0, streaming execution, lazy frames, and memory efficiency&lt;/li&gt;&lt;li&gt;Delta Rust releases, partition handling, correctness fixes, and migration&lt;/li&gt;&lt;li&gt;Parquet performance, FastLanes benchmarks, and Arrow’s Intel macOS support&lt;/li&gt;&lt;li&gt;AI agent orchestration with Omnigent, Kubernetes sandboxes, Spark pipelines, and thrash detection&lt;/li&gt;&lt;/ul&gt;</itunes:summary><itunes:explicit>no</itunes:explicit><itunes:duration>01:06:59</itunes:duration><itunes:image href="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/logos/e0e4749a-7ff6-4101-9540-5e11b292b86c.jpeg"/><itunes:title>XML Functions, Polars 2.0, delta-rs, and Agent Thrashing Policies </itunes:title><itunes:episodeType>full</itunes:episodeType></item><item><title><![CDATA[FileTypes, Language-Agnostic UDFs, Parquet Versioning, and Iceberg Updates]]></title><description><![CDATA[<p>In this episode, Lisa and Scott explore recent updates in data ecosystem technologies including Apache Iceberg, Apache DataFusion, Apache Parquet, and the Lakekeeper Iceberg Rest Catalog, highlighting their implications for data management and interoperability.<br /></p><ul><li>Iceberg Data Fusion integration</li><li>Parquet 2.14 release and versioning</li><li>Nanosecond timestamp support in Parquet</li><li>Iceberg 1.12 release and V4 specification</li><li>Iceberg rest catalog support in Unity Catalog</li><li>Lakekeeper updates and multimodal support<br /></li></ul><p></p>]]></description><guid isPermaLink="false">6b23aea8-db39-43cc-84b5-ac7df5c0594d</guid><dc:creator><![CDATA[Lisa Cao and Scott Haines]]></dc:creator><pubDate>Fri, 18 Sep 2026 23:10:48 GMT</pubDate><enclosure url="https://api.riverside.com/hosting-analytics/media/d79bfa150626cff83f2bbba0a4ce7b1c296ed1d2283c164f0376a9558f1881a6/eyJlcGlzb2RlSWQiOiI2YjIzYWVhOC1kYjM5LTQzY2MtODRiNS1hYzdkZjVjMDU5NGQiLCJwb2RjYXN0SWQiOiIzNGM1ZTgxYS03ODY2LTQxZDQtYTgwMy1iYjhiNmRkMWMyMjYiLCJhY2NvdW50SWQiOiI2N2EwYWU5OTU0OGViN2E4MGYyZDU3NDQiLCJwYXRoIjoibWVkaWEvY2xpcHMvNmFiMjQ0MTgwNGVhNGM0ZGM1MDdiZGE4L29wZW4tbGFrZWhvdXNlLS1haS0yMDI2LTktMjJfXzktMi0xOS5tcDMifQ==.mp3" length="33285477" type="audio/mpeg"/><podcast:transcript url="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/episodes/6b23aea8-db39-43cc-84b5-ac7df5c0594d/transcripts.txt" type="text/plain"/><itunes:summary>&lt;p&gt;In this episode, Lisa and Scott explore recent updates in data ecosystem technologies including Apache Iceberg, Apache DataFusion, Apache Parquet, and the Lakekeeper Iceberg Rest Catalog, highlighting their implications for data management and interoperability.&lt;br /&gt;&lt;/p&gt;&lt;ul&gt;&lt;li&gt;Iceberg Data Fusion integration&lt;/li&gt;&lt;li&gt;Parquet 2.14 release and versioning&lt;/li&gt;&lt;li&gt;Nanosecond timestamp support in Parquet&lt;/li&gt;&lt;li&gt;Iceberg 1.12 release and V4 specification&lt;/li&gt;&lt;li&gt;Iceberg rest catalog support in Unity Catalog&lt;/li&gt;&lt;li&gt;Lakekeeper updates and multimodal support&lt;br /&gt;&lt;/li&gt;&lt;/ul&gt;&lt;p&gt;&lt;/p&gt;</itunes:summary><itunes:explicit>no</itunes:explicit><itunes:duration>00:23:07</itunes:duration><itunes:image href="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/logos/e0e4749a-7ff6-4101-9540-5e11b292b86c.jpeg"/><itunes:title>FileTypes, Language-Agnostic UDFs, Parquet Versioning, and Iceberg Updates</itunes:title><itunes:episodeType>full</itunes:episodeType></item><item><title><![CDATA[Nanosecond Timestamp Precision, Iceberg-Datafusion,Spark Connect Rust Client, and Parquet Types]]></title><description><![CDATA[<p>The biggest throughline this week is format-layer consolidation. Apache Parquet 2.14.0 shipped as a final release, carrying chronological ordering for INT96 timestamps and Variant documentation fixes, while a parallel dev-list debate over versioning semantics and reader behavior for unsupported versions shows the community is actively thinking about long-term compatibility contracts, not just feature additions.<br /></p><p>One layer up, Lance v12 is in a sustained beta sprint, shipping six beta releases in a single week. The work spans file format stabilization (resolving the stable format to 2.2), distributed vector search in Java, in-memory WAL improvements, and IVF index optimization. This is not routine maintenance; it reads as a deliberate push to lock down a production-ready v12 surface before a GA cut.<br /></p><p>At the query engine and integration layer, two parallel votes to move iceberg-rust's DataFusion integration into the Apache DataFusion project are the most structurally interesting governance event of the week. If both votes pass, Iceberg catalog access becomes a first-class DataFusion concern rather than a satellite project, which changes how engine builders think about Iceberg adoption. Separately, the Spark Connect Rust client reached RC2 for its 4.2.0 release, and Spark 4.3.0 RC1 is in vote, keeping the Spark release cadence unusually active.</p><p></p>]]></description><guid isPermaLink="false">dc867db8-d3f9-468f-b326-99a8305650e9</guid><dc:creator><![CDATA[Lisa Cao and Scott Haines]]></dc:creator><pubDate>Fri, 18 Sep 2026 22:00:18 GMT</pubDate><enclosure url="https://api.riverside.com/hosting-analytics/media/99c53fd073e134c81bd32e422e69c3ff8e764a378137b89b92b5550470ddeb58/eyJlcGlzb2RlSWQiOiJkYzg2N2RiOC1kM2Y5LTQ2OGYtYjMyNi05OWE4MzA1NjUwZTkiLCJwb2RjYXN0SWQiOiIzNGM1ZTgxYS03ODY2LTQxZDQtYTgwMy1iYjhiNmRkMWMyMjYiLCJhY2NvdW50SWQiOiI2N2EwYWU5OTU0OGViN2E4MGYyZDU3NDQiLCJwYXRoIjoibWVkaWEvY2xpcHMvNmFhNDc3MWU3MWJkZGVhNjQ3NjRkZTcwL29wZW4tbGFrZWhvdXNlLS1haS0yMDI2LTktMThfXzIyLTAtMjAubXAzIn0=.mp3" length="146261620" type="audio/mpeg"/><podcast:transcript url="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/episodes/dc867db8-d3f9-468f-b326-99a8305650e9/transcripts.txt" type="text/plain"/><itunes:summary>&lt;p&gt;The biggest throughline this week is format-layer consolidation. Apache Parquet 2.14.0 shipped as a final release, carrying chronological ordering for INT96 timestamps and Variant documentation fixes, while a parallel dev-list debate over versioning semantics and reader behavior for unsupported versions shows the community is actively thinking about long-term compatibility contracts, not just feature additions.&lt;br /&gt;&lt;/p&gt;&lt;p&gt;One layer up, Lance v12 is in a sustained beta sprint, shipping six beta releases in a single week. The work spans file format stabilization (resolving the stable format to 2.2), distributed vector search in Java, in-memory WAL improvements, and IVF index optimization. This is not routine maintenance; it reads as a deliberate push to lock down a production-ready v12 surface before a GA cut.&lt;br /&gt;&lt;/p&gt;&lt;p&gt;At the query engine and integration layer, two parallel votes to move iceberg-rust&apos;s DataFusion integration into the Apache DataFusion project are the most structurally interesting governance event of the week. If both votes pass, Iceberg catalog access becomes a first-class DataFusion concern rather than a satellite project, which changes how engine builders think about Iceberg adoption. Separately, the Spark Connect Rust client reached RC2 for its 4.2.0 release, and Spark 4.3.0 RC1 is in vote, keeping the Spark release cadence unusually active.&lt;/p&gt;&lt;p&gt;&lt;/p&gt;</itunes:summary><itunes:explicit>no</itunes:explicit><itunes:duration>01:16:11</itunes:duration><itunes:image href="https://hosting-media.riverside.com/media/podcasts/34c5e81a-7866-41d4-a803-bb8b6dd1c226/logos/e0e4749a-7ff6-4101-9540-5e11b292b86c.jpeg"/><itunes:title>Nanosecond Timestamp Precision, Iceberg-Datafusion,Spark Connect Rust Client, and Parquet Types</itunes:title><itunes:episodeType>full</itunes:episodeType></item></channel></rss>