Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
3.67
0.06/0.00
2.00
0.49
0.24
147.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 1.0
Best clustering variants
2 0.99/0.01 2 1.0
Parameters [6, 0.99, 0.01, 2] average weight 1.0 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
5.12
11.08
0.00/0.26
2.00
1.62
0.81
100.00
7.50
0.25
0
0.00
6.36
0.19/0.00
2.00
0.73
0.36
298.00
4.00
0.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
142.00
3.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
88.00
5.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
35.00
3.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
93.00
6.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.26
5.12
11.08
2.00
1.62
0.81
100.00
7.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.81 | 8.25 | 1.00 | 26-09-15 22:46 | 7 | 6 | A post shared by mewin (@vin.spots) View this post on Instagram.
0.81 | 8.25 | 1.00 | 26-09-15 22:46 | 8 | 7 | A post shared by mewin (@vin.spots).
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.19/0.00
0.00
6.36
2.00
0.73
0.36
298.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
142.00
3.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
88.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
35.00
3.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
93.00
6.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
5.74
14.34
0.00/0.10
3.00
1.56
0.52
431.00
4.00
0.67
0
2.96
9.10
1.48/0.26
5.00
1.34
0.27
325.00
5.60
3.44
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.10
5.74
14.34
3.00
1.56
0.52
431.00
4.00
0.67
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.64 | 8.21 | 2.00 | 26-09-15 22:46 | 5 | 3 | Еден од тие примероци е забележен и снимен на скопските улици, во близина на „Сити мол“.
0.56 | 8.13 | 1.00 | 26-09-16 08:50 | 3 | 0 | Еден од само 20 произведени примероци на специјалната серија „Maserati MC20 Icona“ е забел... улиците во Скопје, во близина на „Сити мол“.
0.36 | 8.13 | 2.00 | 26-09-15 22:46 | 4 | 2 | Луксузниот автомобил „Maserati MC20 Icona“ е специјална серија од 2024 година, направена п...изведени се само 20 примероци во целиот свет.
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.48/0.26
2.96
9.10
5.00
1.34
0.27
325.00
5.60
3.44