Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
2.98
0.15/0.00
2.00
0.40
0.20
141.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 2.6
Best clustering variants
2 0.99/0.01 2 2.6
Parameters [6, 0.99, 0.01, 2] average weight 2.6 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
13.00
15.75
0.00/0.31
2.00
1.72
0.86
374.00
2.00
1.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
144.00
4.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
66.00
5.00
0.00
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
172.00
4.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
107.00
5.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
47.00
6.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.31
13.00
15.75
2.00
1.72
0.86
374.00
2.00
1.00
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.86 | 12.86 | 1.00 | 26-10-06 16:15 | 1 | 0 | На подрачје на Гостивар повторно е повредено дете по пад од електричен тротинет, што е ушт...кви случаи регистрирани во последниот период.
0.86 | 12.86 | 1.00 | 26-10-06 16:15 | 3 | 2 | Иако законските правила се заострија и малолетни лица не смеат да На подрачје на Гостивар ...кви случаи регистрирани во последниот период.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
144.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
66.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
172.00
4.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
107.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
47.00
6.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
6.62
16.94
0.85/0.31
4.00
1.97
0.49
547.00
3.50
2.75
0
0.00
0.32
1.06/0.00
3.00
0.04
0.01
363.00
4.67
0.89
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.85/0.31
6.62
16.94
4.00
1.97
0.49
547.00
3.50
2.75
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
1.06/0.00
0.00
0.32
3.00
0.04
0.01
363.00
4.67
0.89