Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
0.60
0.47/0.00
2.00
0.08
0.04
157.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 1.8
Best clustering variants
2 0.99/0.01 2 1.8
Parameters [6, 0.99, 0.01, 2] average weight 1.8 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
9.01
13.40
0.00/0.26
2.00
1.62
0.81
224.00
2.00
1.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
219.00
4.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
56.00
5.00
0.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
75.00
4.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
119.00
5.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
90.00
6.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.26
9.01
13.40
2.00
1.62
0.81
224.00
2.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.81 | 10.67 | 1.00 | 26-09-15 16:01 | 1 | 0 | Подготовката на зимница на овие простори е традиција која се пренесува со генерации.
0.81 | 10.67 | 1.00 | 26-09-15 16:01 | 3 | 2 | Денешната слика од вторничниот пазарен ден се вклопи во Подготовката на зимница на овие пр...и е традиција која се пренесува со генерации.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
219.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
56.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
75.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
119.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
90.00
6.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
5.54
14.33
0.46/0.26
3.00
1.71
0.57
280.00
3.00
2.67
0
0.00
4.02
1.18/0.00
4.00
0.47
0.12
503.00
4.75
0.69
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.46/0.26
5.54
14.33
3.00
1.71
0.57
280.00
3.00
2.67
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.18/0.00
0.00
4.02
4.00
0.47
0.12
503.00
4.75
0.69