Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
1.04
0.39/0.00
2.00
0.16
0.08
84.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 1.4
Best clustering variants
2 0.99/0.01 2 1.4
Parameters [7, 0.99, 0.01, 2] average weight 1.4 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 7 clusters
Cluster# (7)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
4.78
11.91
0.00/0.16
2.00
1.42
0.71
241.00
5.50
2.25
2
1.98
8.58
0.00/0.08
2.00
1.26
0.63
96.00
2.00
1.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
93.00
1.00
0.00
6
0.00
0.00
0.00/0.00
1.00
0.00
0.00
63.00
5.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
158.00
6.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
93.00
4.00
0.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
44.00
5.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.16
4.78
11.91
2.00
1.42
0.71
241.00
5.50
2.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.71 | 8.73 | 1.00 | 26-08-07 10:13 | 4 | 3 | Постигна прекрасен еврогол од слободен удар и им се извини на Торцидашите Хајдук Сплит как... квалификации помина лесно на едно гостување.
0.71 | 8.73 | 1.00 | 26-08-07 10:13 | 7 | 4 | Постигна прекрасен еврогол од слободен удар и им се извини на Торцидашите.
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.08
1.98
8.58
2.00
1.26
0.63
96.00
2.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.63 | 6.12 | 1.00 | 26-08-07 16:33 | 1 | 5 | Хајдук конечно се разигра.
0.63 | 6.12 | 1.00 | 26-08-07 16:33 | 3 | 7 | Погледнете ја победата на сплитските „били“ Хајдук конечно се разигра.
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
93.00
1.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
6
0.00/0.00
0.00
0.00
1.00
0.00
0.00
63.00
5.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
158.00
6.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
93.00
4.00
0.00
Sentence-cluster ID (count= 7)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
44.00
5.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00
4.82
1.47/0.00
5.00
0.73
0.15
296.00
3.60
2.24
0
0.00
9.28
1.26/0.00
4.00
1.10
0.28
492.00
4.50
5.25
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.47/0.00
0.00
4.82
5.00
0.73
0.15
296.00
3.60
2.24
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.26/0.00
0.00
9.28
4.00
1.10
0.28
492.00
4.50
5.25