Forum Discussion

jisaac's avatar
jisaac
Helper II
1 year ago
Solved

Spark Environment Properties not applied during runtime from Pipeline Notebook Activity

My database has timestamps before 1900, so I have to use a few spark configuration settings to get around the ancient datetime errors in notebooks. The problem is, setting those properties in the Spa...
  • Anonymous's avatar
    Anonymous
    1 year ago

    Hi jisaac 
    unlike Databricks, Microsoft Fabric doesn't currently support initialization scripts for Spark environments. In Databricks, these scripts let you apply settings automatically when a cluster starts, but Fabric doesn’t offer that feature yet.

    you can manually set the Spark configs at the top of each notebook to ensure the necessary settings (like handling pre-1900 timestamps) are always applied. Alternatively, if you disable High Concurrency mode in pipeline runs, each notebook will start a fresh Spark session and correctly pick up the environment settings.

    Another option is to create a small helper notebook that contains all your spark.conf.set(...) lines and use %run at the top of your main notebooks. This way, you keep things consistent without repeating code everywhere.

    If the issue still persists we recommend you to raise a support ticket.

    You can submit a ticket through the Microsoft Power BI Support Portal:

    How to create a Fabric and Power BI Support ticket - Power BI | Microsoft Learn

    Thank you.