---
title: "Spark Topic Articles | Data-Driven AI"
canonical: "https://data-driven.com/tag/spark/"
description: "Browse Data-Driven AI articles tagged Spark, ordered from newest to oldest."
---

Topic archive

# Articles tagged _Spark_

Open the matching articles in date order, or return to the full archive to search across every topic.

[Browse this archive](#archive-articles) [Search all articles](/blog/)

Topic match

4 articles shown

Every article shown includes this tag in its published metadata.

Retained archive

## Follow the Spark thread

The newest matching article appears first. Titles, summaries, images, dates, and destinations remain unchanged.

[![Delta Lake and Spark lakehouse meetup artwork from October 2021](/_astro/Deep-dive-into-building-a-Data-Lakehouse-with-Delta-Lake-and-Spark-1.D6ycXdHU_1AvGyq.webp)](/blog/deep-dive-into-building-a-data-lakehouse-with-delta-lake-and-spark/)

Databricks

### [Delta Lake and Spark Lakehouse Webinar](/blog/deep-dive-into-building-a-data-lakehouse-with-delta-lake-and-spark/)

An archived Sydney Databricks Meetup session on building an ingestion pipeline with Delta Lake and Spark, with source details retained.

[![Azure sentiment analysis workflow using Databricks and PySpark](/_astro/Azure-cognitive-services.oSFay03s_2oi8Wc.webp)](/blog/azure-cognitive-services-sentiment-analysis-v3-0-using-databricks-pyspark/)

Databricks

### [Azure Sentiment Analysis with Databricks PySpark](/blog/azure-cognitive-services-sentiment-analysis-v3-0-using-databricks-pyspark/)

Follow a legacy PySpark walkthrough that sends text to Azure's sentiment API, parses the JSON response, and joins scores to source data.

[![Delta Lake reliability and performance artwork](/_astro/Delta-Lake-Reliability-Performance1.Cgmq5G3X_Z1zLuo8.webp)](/blog/databricks-performance-fixing-the-small-file-problem-with-delta-lake/)

Databricks

### [Databricks Small-File Performance with Delta Lake](/blog/databricks-performance-fixing-the-small-file-problem-with-delta-lake/)

A versioned technical walkthrough of partitioning, MERGE, compaction, and validation for a high-volume Delta Lake workload.

[![Data Science for Dummies Tech Talk series artwork](/_astro/data-science-for-dummies.DyQBFWaJ_1Ymhrw.webp)](/blog/data-science-for-dummies-data-engineering-with-titanic-dataset-databricks-python-tech-talk-3-of-9/)

Databricks

### [Titanic Data Engineering with Databricks and Python](/blog/data-science-for-dummies-data-engineering-with-titanic-dataset-databricks-python-tech-talk-3-of-9/)

An archived Tech Talk showing how the Titanic dataset was prepared and engineered in Databricks before model training.

[Return to all articles](/blog/)

## Need a different route through the archive?

Search every article, or ask the team about the topic you are researching.

[Search all articles](/blog/)

[Contact the team](/contact-us/)
