Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
4.77
8.97
0.14/0.12
3.00
1.08
0.36
266.00
1.00
0.00
All clustering variants
Old ranking - len = 3
2 ■ 0.99/0.01 ■ 2 ■ 3.7
Best clustering variants
2 0.99/0.01 2 3.7
Parameters [4, 0.99, 0.01, 2] average weight 3.7 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 4 clusters
Cluster# (4)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
14.84
16.20
0.00/0.36
2.00
1.81
0.91
327.00
1.50
0.25
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
151.00
2.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
103.00
2.00
0.00
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
71.00
3.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00/0.36
14.84
16.20
2.00
1.81
0.91
327.00
1.50
0.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.91 | 13.08 | 1.00 | 26-09-12 17:46 | 1 | 2 | Во објава на социјалните мрежи, Дарио Амодеи предложи план што вклучува независни оценки на системите за вештачка интелигенција.
0.91 | 13.08 | 1.00 | 26-09-12 17:46 | 2 | 3 | Извршниот директор на компанијата за вештачка интелигенција „Антропик“ Во објава на соција...ценки на системите за вештачка интелигенција.
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
151.00
2.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
103.00
2.00
0.00
Sentence-cluster ID (count= 4)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
71.00
3.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.00
17.81
0.49/0.00
4.00
2.06
0.52
549.00
2.00
0.50
1
0.00
0.00
0.00/0.00
1.00
0.00
0.00
103.00
2.00
0.00
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.49/0.00
0.00
17.81
4.00
2.06
0.52
549.00
2.00
0.50
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.00
0.00
0.00
1.00
0.00
0.00
103.00
2.00
0.00