Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
3.35
0.10/0.00
2.00
0.45
0.22
147.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 0.5
Best clustering variants
2 0.99/0.01 2 0.5
Parameters [6, 0.99, 0.01, 2] average weight 0.5 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
2.60
9.64
0.00/0.12
2.00
1.34
0.67
122.00
4.50
6.25
1
0.00
6.74
0.15/0.00
2.00
0.81
0.40
236.00
7.50
2.25
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
110.00
5.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
32.00
4.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
148.00
8.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
294.00
10.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.12
2.60
9.64
2.00
1.34
0.67
122.00
4.50
6.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.67 | 7.67 | 1.00 | 26-10-05 11:47 | 2 | 0 | Сомненија за разбојништво и обид за разбојништво.
0.67 | 7.67 | 1.00 | 26-10-05 11:32 | 7 | 3 | Се сомничат за сторено кривично дело разбојништво и разбојништво во обид.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.15/0.00
0.00
6.74
2.00
0.81
0.40
236.00
7.50
2.25
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
110.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
32.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
148.00
8.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
294.00
10.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
3.61
9.69
0.16/0.12
3.00
1.06
0.35
416.00
6.33
10.89
1
0.00
9.94
1.29/0.00
5.00
1.31
0.26
526.00
6.40
3.44
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.16/0.12
3.61
9.69
3.00
1.06
0.35
416.00
6.33
10.89
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.29/0.00
0.00
9.94
5.00
1.31
0.26
526.00
6.40
3.44