Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00
0.76
1.00/0.00
3.00
0.10
0.03
176.00
1.00
0.00
All clustering variants
Old ranking - len = 3
2 ■ 0.99/0.01 ■ 2 ■ 1.7
Best clustering variants
2 0.99/0.01 2 1.7
Parameters [5, 0.99, 0.01, 2] average weight 1.7 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 5 clusters
Cluster# (5)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
8.59
13.35
0.00/0.30
2.00
1.70
0.85
178.00
4.50
0.25
0
0.00
9.76
0.05/0.00
2.00
1.00
0.50
523.00
3.50
0.25
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
29.00
6.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
362.00
3.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
159.00
4.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.30
8.59
13.35
2.00
1.70
0.85
178.00
4.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.85 | 10.77 | 1.00 | 26-09-27 12:08 | 5 | 1 | Возилото е одземено и по документирање на случајот ќе следува соодветен поднесок.
0.85 | 10.77 | 1.00 | 26-09-28 16:11 | 4 | 4 | Возилото било одземено и по документирање на случајот, против него ќе следува соодветен поднесок.
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.05/0.00
0.00
9.76
2.00
1.00
0.50
523.00
3.50
0.25
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
29.00
6.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
362.00
3.00
0.00
Sentence-cluster ID (count= 5)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
159.00
4.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
8.59
13.35
0.00/0.30
2.00
1.70
0.85
178.00
4.50
0.25
0
0.00
9.71
1.09/0.00
5.00
1.11
0.22
1073.00
4.00
1.20
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.30
8.59
13.35
2.00
1.70
0.85
178.00
4.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.85 | 10.77 | 1.00 | 26-09-27 12:08 | 5 | 1 | Возилото е одземено и по документирање на случајот ќе следува соодветен поднесок.
0.85 | 10.77 | 1.00 | 26-09-28 16:11 | 4 | 4 | Возилото било одземено и по документирање на случајот, против него ќе следува соодветен поднесок.
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.09/0.00
0.00
9.71
5.00
1.11
0.22
1073.00
4.00
1.20