Forum Discussion
On Premises SQL Historical table taking long time to load into Fabric Lakehouse.
- 1 year ago
Hi devaj123,
Using Notebooks with PySpark provides the best performance and flexibility, especially for large tables. If this is a historical table that changes over time, implement watermarking and load only new or updated rows using a timestamp or identity column.
Repartition the data to spread it across multiple parallel tasks, especially helpful for performance when writing large volumes. Save the data in Delta format, which supports efficient queries, updates, and time travel versioning.
Thanks,Prashanth Are
MS Fabric community support
If our resolved your issue, please mark it as "Accept as solution" and click "Yes" if you found it helpful.
May I ask if you have resolved this issue? If so, please mark the helpful reply and accept it as the solution. This will be helpful for other community members who have similar problems to solve it faster.
If we don’t hear back, we’ll go ahead and close this thread. For any further discussions or questions, please start a new thread in the Microsoft Fabric Community Forum we’ll be happy to assist.
Thank you for being part of the Microsoft Fabric Community.