<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Apache-Spark on Project Wintermute</title><link>https://wintermutecore.com/tags/apache-spark/</link><description>Recent content in Apache-Spark on Project Wintermute</description><generator>Hugo</generator><language>en-us</language><lastBuildDate>Tue, 21 Jul 2026 14:00:07 +0000</lastBuildDate><atom:link href="https://wintermutecore.com/tags/apache-spark/index.xml" rel="self" type="application/rss+xml"/><item><title>Apache Spark Cache And Streaming Metadata Notes</title><link>https://wintermutecore.com/posts/apache-spark-cache-streaming-metadata/</link><pubDate>Tue, 21 Jul 2026 14:00:07 +0000</pubDate><guid>https://wintermutecore.com/posts/apache-spark-cache-streaming-metadata/</guid><description>&lt;p&gt;&lt;a href="https://github.com/apache/spark"&gt;Apache Spark&lt;/a&gt; had a busy week, with 108 commits on &lt;code&gt;master&lt;/code&gt; and several changes that matter to data platform teams more than application authors. The main thread is practical: cache layout, Kafka metadata calls, Real Time Mode checkpoint cost, and catalog metadata that needs to be correct when tools read JSON output.&lt;/p&gt;</description></item><item><title>Apache Spark Python UDF And SQL Operator Notes</title><link>https://wintermutecore.com/posts/apache-spark-python-sql-operator-notes/</link><pubDate>Mon, 06 Jul 2026 14:00:02 +0000</pubDate><guid>https://wintermutecore.com/posts/apache-spark-python-sql-operator-notes/</guid><description>&lt;p&gt;Apache Spark had a busy week on master: 120 commits touched 457 files. The activity in &lt;a href="https://github.com/apache/spark"&gt;apache/spark&lt;/a&gt; is not a release note dump; it is mostly data path cleanup around Python UDF transport, SQL correctness, security, and test determinism that affects teams running Spark as part of an ETL platform.&lt;/p&gt;</description></item></channel></rss>