Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
3.75
0.05/0.00
2.00
0.50
0.25
143.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 0.1
Best clustering variants
2 0.99/0.01 2 0.1
Parameters [5, 0.99, 0.01, 2] average weight 0.1 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 5 clusters
Cluster# (5)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.37
5.73
0.00/0.03
2.00
1.15
0.58
34.00
2.00
1.00
2
0.00
3.02
0.40/0.00
2.00
0.30
0.15
588.00
4.50
0.25
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
161.00
3.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
77.00
4.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
45.00
6.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.03
0.37
5.73
2.00
1.15
0.58
34.00
2.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.58 | 3.34 | 1.00 | 26-10-07 18:35 | 1 | 2 | На бул.
0.58 | 3.34 | 1.00 | 26-10-07 18:35 | 3 | 4 | Во несреќата што се На бул.
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.40/0.00
0.00
3.02
2.00
0.30
0.15
588.00
4.50
0.25
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
161.00
3.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
77.00
4.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
45.00
6.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.47
9.23
1.07/0.03
5.00
1.31
0.26
384.00
3.80
2.96
0
0.00
3.48
0.34/0.00
2.00
0.36
0.18
521.00
3.50
0.25
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.07/0.03
0.47
9.23
5.00
1.31
0.26
384.00
3.80
2.96
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.34/0.00
0.00
3.48
2.00
0.36
0.18
521.00
3.50
0.25