Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
1.75
0.33/0.00
2.00
0.22
0.11
206.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 2.1
Best clustering variants
2 0.99/0.01 2 2.1
Parameters [4, 0.99, 0.01, 2] average weight 2.1 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 4 clusters
Cluster# (4)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
8.56
14.13
0.00/0.19
2.00
1.48
0.74
464.00
2.50
0.25
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
91.00
3.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
108.00
4.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
212.00
1.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.19
8.56
14.13
2.00
1.48
0.74
464.00
2.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.74 | 10.46 | 1.00 | 26-09-28 08:40 | 2 | 3 | Највисок раст е забележан во Унгарија, каде што цените се повисоки за 57 проценти, по што ...те на Евростат анализирани од Еуроњуз Бизнис.
0.74 | 10.46 | 1.00 | 26-09-28 08:40 | 3 | 4 | Највисок раст е забележан во Унгарија, каде што цените се повисоки за 57 проценти, по што следуваат Бугарија со 54 проценти и.
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
91.00
3.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
108.00
4.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
212.00
1.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
8.40
17.50
0.81/0.31
4.00
1.89
0.47
784.00
2.50
1.25
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
91.00
3.00
0.00
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.81/0.31
8.40
17.50
4.00
1.89
0.47
784.00
2.50
1.25
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
91.00
3.00
0.00