Article-cluster size (# of articles)
new weight
old weight
lower/upper
Articles count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00
1.23
0.39/0.00
2.00
0.16
0.08
154.00
1.00
0.00
All clustering variants
Old ranking - len = 2
2 ■ 0.99/0.01 ■ 2 ■ 0.7
Best clustering variants
2 0.99/0.01 2 0.7
Parameters [6, 0.99, 0.01, 2] average weight 0.7 execution time 0.0s
NEW: semantic (merge_dist=0.35, adaptive=off) — 6 clusters
Cluster# (6)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
3.30
10.56
0.00/0.12
2.00
1.34
0.67
182.00
4.50
2.25
0
0.00
7.22
0.67/0.00
3.00
0.80
0.27
385.00
5.33
11.56
5
0.00
0.00
0.00/0.00
1.00
0.00
0.00
172.00
5.00
0.00
3
0.00
0.00
0.00/0.00
1.00
0.00
0.00
59.00
7.00
0.00
4
0.00
0.00
0.00/0.00
1.00
0.00
0.00
105.00
8.00
0.00
2
0.00
0.00
0.00/0.00
1.00
0.00
0.00
110.00
9.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
0.00/0.12
3.30
10.56
2.00
1.34
0.67
182.00
4.50
2.25
Simi | Wt/Ln | TimeC | Datetime | S# | S-Id | Text
0.67 | 7.64 | 1.00 | 26-09-12 09:47 | 3 | 2 | За разлика од обуката на големи модели, inference.
0.67 | 7.64 | 1.00 | 26-09-12 09:47 | 6 | 3 | За разлика од обуката на големи модели, inference се користи за услуги како чет-ботови, гласовни асистенти и алатки за пишување код.
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.67/0.00
0.00
7.22
3.00
0.80
0.27
385.00
5.33
11.56
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
5
0.00/0.00
0.00
0.00
1.00
0.00
0.00
172.00
5.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
3
0.00/0.00
0.00
0.00
1.00
0.00
0.00
59.00
7.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
4
0.00/0.00
0.00
0.00
1.00
0.00
0.00
105.00
8.00
0.00
Sentence-cluster ID (count= 6)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
2
0.00/0.00
0.00
0.00
1.00
0.00
0.00
110.00
9.00
0.00
OLD: TF-IDF K-Means (k=2) — 2 clusters
Cluster# (2)
new weight
old weight
lower/upper
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
3.56
8.52
0.25/0.12
3.00
0.96
0.32
354.00
4.67
1.56
1
0.00
7.45
2.07/0.00
6.00
1.05
0.17
659.00
6.67
7.89
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
0
0.25/0.12
3.56
8.52
3.00
0.96
0.32
354.00
4.67
1.56
Sentence-cluster ID (count= 2)
lower/upper
new weight
old weight
Sentence count
Total similarity
Avg. similarity
Total chr length
Sentence # average
Sentence # variation
1
2.07/0.00
0.00
7.45
6.00
1.05
0.17
659.00
6.67
7.89