Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
0.58
0.48/0.00
2.00
0.07
0.04
195.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 7.3
Best clustering variants
2 0.99/0.01 2 7.3
Parameters [3, 0.99, 0.01, 2] average weight 7.3 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 3 clusters
Cluster# (3)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
22.00
22.63
0.08/0.37
3.00
2.32
0.77
582.00
2.00
0.67
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
149.00
4.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
55.00
5.00
0.00
Sentence-cluster ID (count= 3)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.08/0.37
22.00
22.63
3.00
2.32
0.77
582.00
2.00
0.67
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.92 | 14.57 | 2.00 | 26-09-18 09:00 | 1 | 1 | Кинеските производители на автомобили забрзано бараат постоечки фабрики низ Европа со цел ... правила за локална содржина, пишува Ројтерс.
0.92 | 14.57 | 2.00 | 26-09-18 09:00 | 3 | 3 | Производителите првенствено бараат постоечки погони за Кинеските производители на автомоби... правила за локална содржина, пишува Ројтерс.
0.47 | 6.70 | 1.00 | 26-09-17 13:41 | 2 | 0 | Кинеските производители на автомобили бараат европски фабрики пред да стапат на сила правилата на ЕУ за локално производство.
Sentence-cluster ID (count= 3)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
149.00
4.00
0.00
Sentence-cluster ID (count= 3)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
55.00
5.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
8.14
17.96
0.94/0.37
4.00
2.01
0.50
661.00
3.25
2.19
1
0.00
0.00
0.00/0.00
1.00
0.47
0.47
125.00
2.00
0.00
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.94/0.37
8.14
17.96
4.00
2.01
0.50
661.00
3.25
2.19
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.47
0.47
125.00
2.00
0.00