Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
7.96
5.98
0.00/0.27
2.00
0.82
0.41
125.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 1.8
Best clustering variants
2 0.99/0.01 2 1.8
Parameters [6, 0.99, 0.01, 2] average weight 1.8 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
6.71
13.69
0.00/0.32
2.00
1.73
0.87
182.00
13.50
0.25
1
2.43
10.49
0.00/0.09
2.00
1.28
0.64
212.00
5.50
6.25
0
0.00
3.51
0.31/0.00
2.00
0.49
0.24
125.00
8.00
16.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
70.00
9.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
109.00
10.00
0.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
21.00
11.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.32
6.71
13.69
2.00
1.73
0.87
182.00
13.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.87 | 10.70 | 1.00 | 26-08-28 19:52 | 13 | 7 | Текстот Сообраќајка во Србија – има загинати, возило комплетно изгоре е превземен од А1он . Текстот.
0.87 | 10.70 | 1.00 | 26-08-28 19:52 | 14 | 8 | Сообраќајка во Србија – има загинати, возило комплетно изгоре е превземен од А1он.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.09
2.43
10.49
2.00
1.28
0.64
212.00
5.50
6.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.64 | 7.79 | 1.00 | 26-08-29 15:35 | 3 | 0 | Тешка сообраќајна несреќа се случи во близина на Рума, Србија, при што според неофицијални...мации од српските медиуми, има загинати лица.
0.64 | 7.79 | 1.00 | 26-08-28 19:52 | 8 | 2 | Денес се случи тешка сообраќајна несреќа во близина на Рума, во Србија.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.31/0.00
0.00
3.51
2.00
0.49
0.24
125.00
8.00
16.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
70.00
9.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
109.00
10.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
21.00
11.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
2.93
13.38
1.70/0.32
6.00
2.01
0.34
491.00
9.83
13.81
1
0.00
3.91
0.61/0.00
3.00
0.49
0.16
228.00
8.33
9.56
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.70/0.32
2.93
13.38
6.00
2.01
0.34
491.00
9.83
13.81
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.61/0.00
0.00
3.91
3.00
0.49
0.16
228.00
8.33
9.56