This is best Fabric, Power BI, SQL and AI community event. How do we know? The last event sold out! Save €200 with code FABCMTY200.
Register nowThe Fabric community is now in read-only for platform upgrade. Learn more
Hello,
im currently facing a Python clustering problem. I would like to extract from a column (say "comment") from the table "rawdata" the individual comments from customers, take out stop words and then cluster by similar sounding terms. Similar sounding because umlauts like ä,ö... are not always saved/ displayed correctly.
Unfortunately, i cant get a visual to work here. Can someone help me here?
Code I tried to use:
# The following code for creating a data frame and removing duplicate rows is always executed and serves as a preamble to your script:
| ID | Comment |
| 1 | Kündigung rückgängig machen |
| 2 | Invoice |
Join us in Barcelona for FabCon and SQLCon, the Fabric, Power BI, SQL, and AI community event. Save €200 with code FABCMTY200.
| User | Count |
|---|---|
| 23 | |
| 22 | |
| 14 | |
| 14 | |
| 13 |
| User | Count |
|---|---|
| 47 | |
| 39 | |
| 24 | |
| 20 | |
| 20 |