Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
0.57
0.48/0.00
2.00
0.07
0.04
184.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 2.7
Best clustering variants
2 0.99/0.01 2 2.7
Parameters [5, 0.99, 0.01, 2] average weight 2.7 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 5 clusters
Cluster# (5)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
13.45
15.63
0.00/0.32
2.00
1.74
0.87
336.00
1.50
0.25
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
64.00
3.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
121.00
4.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
48.00
5.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
30.00
6.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.32
13.45
15.63
2.00
1.74
0.87
336.00
1.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.87 | 12.71 | 1.00 | 26-10-04 07:00 | 1 | 0 | Британија е под притисок да воведе царини за евтиниот увоз на кинески возила, пред да стапат на сила протекционистичките трговски мерки.
0.87 | 12.71 | 1.00 | 26-10-04 07:00 | 2 | 1 | Британската автомобилска индустрија се соочува со „тежок избор“ Британија е под притисок д...т на сила протекционистичките трговски мерки.
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
64.00
3.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
121.00
4.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
48.00
5.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
30.00
6.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
4.28
12.92
1.65/0.32
5.00
1.74
0.35
478.00
3.40
3.44
0
0.00
0.00
0.00/0.00
1.00
0.00
0.00
121.00
4.00
0.00
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.65/0.32
4.28
12.92
5.00
1.74
0.35
478.00
3.40
3.44
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.00
0.00
0.00
1.00
0.00
0.00
121.00
4.00
0.00