Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00
2.62
0.75/0.00
3.00
0.35
0.12
163.00
1.00
0.00
All clustering variants
Old ranking - len = 3
2 ■ 0.99/0.01 ■ 2 ■ 2.0
Best clustering variants
2 0.99/0.01 2 2.0
Parameters [5, 0.99, 0.01, 2] average weight 2.0 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 5 clusters
Cluster# (5)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
10.24
14.68
0.00/0.33
2.00
1.76
0.88
231.00
5.00
1.00
0
0.00
8.56
0.06/0.00
2.00
0.99
0.49
279.00
3.50
0.25
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
115.00
3.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
80.00
4.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
123.00
5.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.33
10.24
14.68
2.00
1.76
0.88
231.00
5.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.88 | 11.56 | 1.00 | 26-09-24 16:35 | 6 | 4 | Известен бил јавен обвинител и по документирање на настанот ќе биде поднесена соодветна пријава.
0.88 | 11.56 | 1.00 | 26-09-24 14:54 | 4 | 6 | Известен бил јавен обвинител и по документирање на настанот ќе биде поднесена соодветна пријава, соопшти Полициската станица Кавадарци.
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.06/0.00
0.00
8.56
2.00
0.99
0.49
279.00
3.50
0.25
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
115.00
3.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
80.00
4.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
123.00
5.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00
15.46
1.06/0.00
5.00
2.02
0.40
550.00
4.20
0.96
0
0.00
8.88
0.02/0.00
2.00
1.03
0.51
278.00
4.00
1.00
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
1.06/0.00
0.00
15.46
5.00
2.02
0.40
550.00
4.20
0.96
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.02/0.00
0.00
8.88
2.00
1.03
0.51
278.00
4.00
1.00