<?xml version="1.0" encoding="UTF-8"?>
<rss xmlns:content="http://purl.org/rss/1.0/modules/content/" xmlns:dc="http://purl.org/dc/elements/1.1/" xmlns:rdf="http://www.w3.org/1999/02/22-rdf-syntax-ns#" xmlns:taxo="http://purl.org/rss/1.0/modules/taxonomy/" version="2.0">
  <channel>
    <title>topic Re: Setting a Spark cluste when running a notebook or running anything else(eg. pipeline, dataflow etc)? in Pipelines</title>
    <link>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4252031#M6003</link>
    <description>&lt;P&gt;Hi&amp;nbsp;Anonymous&lt;/LI-USER&gt;&amp;nbsp;,&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;You can set the spark properties in Workspace setting, you can create pool based on your requirements for spark nodes and attach to your workspace. You can create new environment on which you can specify the executor core, memory, dynamic allocation to be used. using Spark properties tab you can add your environment specific spark properties, Libraries.&lt;/P&gt;&lt;P&gt;please refer screenshot -&amp;nbsp;&lt;/P&gt;&lt;P&gt;&lt;img /&gt;&lt;img /&gt;&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;Thanks,&lt;/P&gt;&lt;P&gt;Srisakthi&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;---------------------------------------------------------------------------&lt;/P&gt;&lt;P&gt;If this answers your question please mark as solution accepted.&lt;/P&gt;</description>
    <pubDate>Tue, 22 Oct 2024 06:45:01 GMT</pubDate>
    <dc:creator>Srisakthi</dc:creator>
    <dc:date>2024-10-22T06:45:01Z</dc:date>
    <item>
      <title>Setting a Spark cluste when running a notebook or running anything else(eg. pipeline, dataflow etc)?</title>
      <link>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4251616#M5999</link>
      <description>&lt;P&gt;Normally in Databricks I can set the cluster that a notebook would run on in the box in the top right of a notebook. However in a Fabric notebook, I can't see it in the top right corner anymore.&lt;/P&gt;&lt;P&gt;How can I assign a cluster to a notebook or any other activity eg. pipeline?&lt;/P&gt;&lt;P&gt;&lt;img /&gt;&lt;img /&gt;&lt;/P&gt;</description>
      <pubDate>Tue, 22 Oct 2024 01:43:23 GMT</pubDate>
      <guid>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4251616#M5999</guid>
      <dc:creator>Anonymous</dc:creator>
      <dc:date>2024-10-22T01:43:23Z</dc:date>
    </item>
    <item>
      <title>Re: Setting a Spark cluste when running a notebook or running anything else(eg. pipeline, dataflow etc)?</title>
      <link>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4251810#M6001</link>
      <description>&lt;P&gt;For Notebook, it's in the Environment dropdown.&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;For general use cases, I think I would just use the default, i.e. Starter pool, as they have shorter startup times.&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;In that case, you don't need to think about cluster. The Notebook will automatically use the default pool (cluster) if we don't specify otherwise.&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;I think only Notebooks (and Spark Job Definitons) use Spark.&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;I'm not aware that Data Pipeline or Dataflow Gen2 use Spark. I think they use another technology which is fully managed (hidden from us).&lt;/P&gt;&lt;P&gt;If you run a Notebook in Data Pipeline, it will use Spark.&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;Here is a link to some information about Apache Spark in Fabric:&lt;/P&gt;&lt;P&gt;&lt;A href="https://learn.microsoft.com/en-us/fabric/data-engineering/spark-compute" target="_blank"&gt;https://learn.microsoft.com/en-us/fabric/data-engineering/spark-compute&lt;/A&gt;&lt;/P&gt;</description>
      <pubDate>Tue, 22 Oct 2024 07:15:31 GMT</pubDate>
      <guid>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4251810#M6001</guid>
      <dc:creator>frithjof_v</dc:creator>
      <dc:date>2024-10-22T07:15:31Z</dc:date>
    </item>
    <item>
      <title>Re: Setting a Spark cluste when running a notebook or running anything else(eg. pipeline, dataflow etc)?</title>
      <link>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4252031#M6003</link>
      <description>&lt;P&gt;Hi&amp;nbsp;Anonymous&lt;/LI-USER&gt;&amp;nbsp;,&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;You can set the spark properties in Workspace setting, you can create pool based on your requirements for spark nodes and attach to your workspace. You can create new environment on which you can specify the executor core, memory, dynamic allocation to be used. using Spark properties tab you can add your environment specific spark properties, Libraries.&lt;/P&gt;&lt;P&gt;please refer screenshot -&amp;nbsp;&lt;/P&gt;&lt;P&gt;&lt;img /&gt;&lt;img /&gt;&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;Thanks,&lt;/P&gt;&lt;P&gt;Srisakthi&lt;/P&gt;&lt;P&gt;&amp;nbsp;&lt;/P&gt;&lt;P&gt;---------------------------------------------------------------------------&lt;/P&gt;&lt;P&gt;If this answers your question please mark as solution accepted.&lt;/P&gt;</description>
      <pubDate>Tue, 22 Oct 2024 06:45:01 GMT</pubDate>
      <guid>https://community.fabric.microsoft.com/t5/Pipelines/Setting-a-Spark-cluste-when-running-a-notebook-or-running/m-p/4252031#M6003</guid>
      <dc:creator>Srisakthi</dc:creator>
      <dc:date>2024-10-22T06:45:01Z</dc:date>
    </item>
  </channel>
</rss>

