Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
3.02
0.15/0.00
2.00
0.40
0.20
156.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 3.8
Best clustering variants
2 0.99/0.01 2 3.8
Parameters [6, 0.99, 0.01, 2] average weight 3.8 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
18.80
18.35
0.00/0.43
2.00
1.95
0.98
425.00
2.00
1.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
98.00
4.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
128.00
5.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
132.00
6.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
146.00
4.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
15.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.43
18.80
18.35
2.00
1.95
0.98
425.00
2.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.98 | 15.46 | 1.00 | 26-08-31 14:43 | 1 | 3 | Ученици од социјално загрозени семејства од подрачјето на Гостивар ќе бидат израдувани со ...ребни за на училиште во новата учебна година.
0.98 | 15.46 | 1.00 | 26-08-31 14:43 | 3 | 5 | Волонтерите на Ученици од социјално загрозени семејства од подрачјето на Гостивар ќе бидат...ребни за на училиште во новата учебна година.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
98.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
128.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
132.00
6.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
146.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
15.00
5.00
0.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.00 | 0.00 | 1.00 | 26-08-31 14:43 | 5 | 6 | Волонтерите на.
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
8.38
18.97
1.26/0.43
5.00
2.35
0.47
714.00
3.60
2.24
0
0.00
0.46
0.49/0.00
2.00
0.06
0.03
230.00
5.00
1.00
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.26/0.43
8.38
18.97
5.00
2.35
0.47
714.00
3.60
2.24
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.49/0.00
0.00
0.46
2.00
0.06
0.03
230.00
5.00
1.00