Skip to content
ByDataFix
  • Home
  • Contact
  • Privacy Policy
  • Terms & Conditions
Subscribe
Subscribe
Databricks Troubleshooting

Fix Corrupted Parquet Files in Spark (“Could Not Read Footer”)

One bad file shouldn't take down the whole read. Here's why Parquet files corrupt, how to skip past them to get your data now, and how to find and fix the real culprit.

Read More
July 22, 2026
Databricks Troubleshooting

Fix Delta ConcurrentAppendException (Concurrent Write Conflicts)

Two jobs wrote the same Delta table and one blew up. Here's what optimistic concurrency is really doing, and how to fix conflicts with partition filters, retries, and row-level concurrency.

Read More
July 22, 2026
Databricks Troubleshooting

Fix Over-Partitioned Delta Tables (Slow Queries) in Databricks

Partitioning by the wrong column shatters your table into millions of tiny files and grinds queries to a halt. Here's how to right-size it — and why liquid clustering is usually the better answer.

Read More
July 22, 2026
Databricks Troubleshooting

Why Empty Strings Become Null in a Partitioned Column (Spark)

No error, just wrong data: partition by a column with empty strings and Spark reads them back as null. Here's exactly why — and how to keep the values distinct.

Read More
July 22, 2026
Databricks Troubleshooting

Fix “User Does Not Have Permission SELECT on ANY File” in Databricks

You're reading data by file path on a Table-ACL cluster. The blunt fix is dangerous — here's what the error means and the scoped, safe way to grant file access.

Read More
July 22, 2026
Databricks Troubleshooting

Fix “User Does Not Have USAGE on Database” in Databricks

Granting SELECT isn't enough. Databricks needs USAGE (USE SCHEMA) on the database too — because permissions are a hierarchy. Here's how to grant the right privileges.

Read More
July 22, 2026
Databricks Troubleshooting

Fix Parquet “Incompatible Schema” Errors in Databricks & Spark

"Failed to merge incompatible data types" and "Parquet column cannot be converted" both mean your files disagree on a column's type. Here's how to fix compatible and incompatible cases.

Read More
July 22, 2026
Databricks Troubleshooting

Fix “Unable to Infer Schema for Parquet” in Databricks & Spark

The error almost always means one thing: Spark found no data files at the path. Here's the 30-second fix, the real root causes, and how to stop it happening again.

Read More
July 22, 2026
1 2 3 Next »

Recent Posts

  • Fix Corrupted Parquet Files in Spark (“Could Not Read Footer”)
  • Fix Delta ConcurrentAppendException (Concurrent Write Conflicts)
  • Fix Over-Partitioned Delta Tables (Slow Queries) in Databricks
  • Why Empty Strings Become Null in a Partitioned Column (Spark)
  • Fix “User Does Not Have Permission SELECT on ANY File” in Databricks

Archives

  • July 2026

Categories

  • Data Engineering Basics
  • Databricks Troubleshooting
ByDataFix

All about data engineering

© 2026 ByDataFix. All rights reserved.

Subscribe to ByDataFix

Practical data engineering — PySpark, Azure, Snowflake and more. New posts straight to your inbox. No spam, unsubscribe anytime.