Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00
2.93
0.68/0.00
3.00
0.42
0.14
128.00
1.00
0.00
All clustering variants
Old ranking - len = 3
2 ■ 0.99/0.01 ■ 2 ■ 0.0
Best clustering variants
2 0.99/0.01 2 0.0
Parameters [4, 0.99, 0.01, 2] average weight 0.0 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 4 clusters
Cluster# (4)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00
6.75
0.66/0.00
3.00
0.67
0.22
668.00
3.00
0.67
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
180.00
3.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
112.00
4.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
56.00
2.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.66/0.00
0.00
6.75
3.00
0.67
0.22
668.00
3.00
0.67
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
180.00
3.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
112.00
4.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
56.00
2.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
10.10
9.82
0.00/0.23
2.00
1.02
0.51
497.00
3.00
1.00
0
0.00
7.69
0.75/0.00
4.00
0.90
0.23
519.00
3.00
0.50
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.23
10.10
9.82
2.00
1.02
0.51
497.00
3.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.78 | 11.36 | 1.00 | 26-09-22 17:36 | 2 | 3 | Оваа не е првпат коњи да минуваат на обој булевар, па дури и да трчаат по коловозот ставај... ризик возачите и пешаците, но и самите себе.
0.24 | 11.36 | 1.00 | 26-09-22 17:36 | 4 | 4 | Оваа не е првпат коњи да минуваат на обој булевар, па дури и да трчаат по коловозот ставај...дот (ЈП Комунална Хигиена), АХВ, полицијата….
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.75/0.00
0.00
7.69
4.00
0.90
0.23
519.00
3.00
0.50