Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
1.87
0.31/0.00
2.00
0.24
0.12
165.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 3.0
Best clustering variants
2 0.99/0.01 2 3.0
Parameters [5, 0.99, 0.01, 2] average weight 3.0 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 5 clusters
Cluster# (5)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
14.78
16.53
0.00/0.35
2.00
1.80
0.90
374.00
2.00
1.00
0
0.00
4.50
0.27/0.00
2.00
0.57
0.28
188.00
5.50
0.25
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
199.00
4.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
141.00
4.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
52.00
5.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.35
14.78
16.53
2.00
1.80
0.90
374.00
2.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.90 | 13.61 | 1.00 | 26-09-30 15:49 | 1 | 3 | Изминува првиот месец од новата учебна година, а за учениците од основните училишта на под...е не е профункциониран организираниот превоз.
0.90 | 13.61 | 1.00 | 26-09-30 15:49 | 3 | 5 | Проблемот беше актуелизиран на денешната седница на Изминува првиот месец од новата учебна...е не е профункциониран организираниот превоз.
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.27/0.00
0.00
4.50
2.00
0.57
0.28
188.00
5.50
0.25
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
199.00
4.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
141.00
4.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
52.00
5.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
6.44
15.65
1.10/0.35
4.00
1.80
0.45
567.00
3.25
2.19
1
0.00
5.88
0.45/0.00
3.00
0.65
0.22
387.00
5.00
0.67
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.10/0.35
6.44
15.65
4.00
1.80
0.45
567.00
3.25
2.19
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.45/0.00
0.00
5.88
3.00
0.65
0.22
387.00
5.00
0.67