Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
2.12
0.26/0.00
2.00
0.29
0.14
136.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 3.6
Best clustering variants
2 0.99/0.01 2 3.6
Parameters [4, 0.99, 0.01, 2] average weight 3.6 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 4 clusters
Cluster# (4)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
14.28
16.89
0.00/0.37
2.00
1.84
0.92
378.00
4.00
1.00
2
0.00
4.54
0.26/0.00
2.00
0.58
0.29
178.00
2.50
2.25
0
0.00
6.70
0.16/0.00
2.00
0.79
0.39
258.00
4.50
0.25
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
159.00
6.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.37
14.28
16.89
2.00
1.84
0.92
378.00
4.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.92 | 13.40 | 1.00 | 26-08-31 09:45 | 3 | 2 | Со цел избегнување на метежите пред шалтерите и заштеда на време, од НЛБ Банка ги советува...густовските пензии за клиентите на НЛБ Банка.
0.92 | 13.40 | 1.00 | 26-08-31 09:45 | 5 | 3 | Со цел избегнување на метежите пред шалтерите и заштеда на време, од НЛБ Банка ги советува...големо користење на безготовински плаќања во.
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.26/0.00
0.00
4.54
2.00
0.58
0.29
178.00
2.50
2.25
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.16/0.00
0.00
6.70
2.00
0.79
0.39
258.00
4.50
0.25
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
159.00
6.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00
15.40
1.21/0.00
5.00
1.91
0.38
715.00
3.80
2.96
0
0.00
6.70
0.16/0.00
2.00
0.79
0.39
258.00
4.50
0.25
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.21/0.00
0.00
15.40
5.00
1.91
0.38
715.00
3.80
2.96
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.16/0.00
0.00
6.70
2.00
0.79
0.39
258.00
4.50
0.25