Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
3.35
0.11/0.00
2.00
0.44
0.22
160.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 1.5
Best clustering variants
2 0.99/0.01 2 1.5
Parameters [7, 0.99, 0.01, 2] average weight 1.5 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 7 clusters
Cluster# (7)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
7.48
12.37
0.00/0.39
2.00
1.87
0.94
87.00
3.50
6.25
0
0.00
5.09
0.25/0.00
2.00
0.60
0.30
259.00
4.50
6.25
6
0.00
0.00
0.00/0.00
1.00
0.00
0.00
60.00
8.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
72.00
9.00
0.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
48.00
10.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
159.00
11.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
60.00
12.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.39
7.48
12.37
2.00
1.87
0.94
87.00
3.50
6.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.94 | 10.39 | 1.00 | 26-08-12 19:30 | 1 | 1 | Случајот бил пријавен вчера во СВР Битола.
0.94 | 10.39 | 1.00 | 26-08-12 19:30 | 6 | 6 | По Случајот бил пријавен вчера во СВР Битола.
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.25/0.00
0.00
5.09
2.00
0.60
0.30
259.00
4.50
6.25
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
6
0.00/0.00
0.00
0.00
1.00
0.00
0.00
60.00
8.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
72.00
9.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
48.00
10.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
159.00
11.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
60.00
12.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
3.04
7.71
2.39/0.39
7.00
1.29
0.18
486.00
8.14
11.84
0
0.00
5.09
0.25/0.00
2.00
0.60
0.30
259.00
4.50
6.25
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
2.39/0.39
3.04
7.71
7.00
1.29
0.18
486.00
8.14
11.84
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.25/0.00
0.00
5.09
2.00
0.60
0.30
259.00
4.50
6.25