GassiGiuseppe
|
1de2cc59db
|
doctor and model test
|
2025-10-08 22:51:36 +02:00 |
|
Christian Risi
|
c2e13bc9c6
|
Quick fix to architecture
|
2025-10-08 12:34:09 +02:00 |
|
GassiGiuseppe
|
fc44929a7b
|
moved spanned mask variables in init for better reliability, also tested
|
2025-10-07 23:15:50 +02:00 |
|
Christian Risi
|
99b5198c9a
|
WIP
|
2025-10-07 16:38:08 +02:00 |
|
Christian Risi
|
fdece42462
|
Made model Batch ready
|
2025-10-07 16:37:20 +02:00 |
|
Christian Risi
|
109ad9f36b
|
Changed Imports
|
2025-10-07 16:36:59 +02:00 |
|
Christian Risi
|
f9545aca1d
|
Deleted MultiHeadAttention
|
2025-10-07 16:36:11 +02:00 |
|
GassiGiuseppe
|
56d438f01a
|
WIP NanoSocratesCore
|
2025-10-06 18:21:27 +02:00 |
|
GassiGiuseppe
|
e1549d4458
|
Modified decoder and decoder for sequential architecture
|
2025-10-06 18:20:46 +02:00 |
|
Christian Risi
|
e93710af08
|
Fixed illegal tokens being added in target output
|
2025-10-06 16:16:47 +02:00 |
|
Christian Risi
|
b1e7af0607
|
Merge branch 'dev.embedder' of https://repositories.communitynotfound.work/PoliBa-DeepLearning/NanoSocrates into dev.embedder
|
2025-10-06 15:55:44 +02:00 |
|
Christian Risi
|
c217f5dec9
|
Added 2 types of masking
|
2025-10-06 15:45:45 +02:00 |
|
Christian Risi
|
49f0beb6ea
|
Updated imports
|
2025-10-06 15:45:28 +02:00 |
|
GassiGiuseppe
|
948c3fd7ac
|
update to batch attention mask
|
2025-10-06 13:03:03 +02:00 |
|
GassiGiuseppe
|
7e40a36701
|
wip: NanoSocratesCore
|
2025-10-05 22:58:06 +02:00 |
|
GassiGiuseppe
|
0f243eaac2
|
added padding_mask entry to decoder and encoder
|
2025-10-05 18:46:06 +02:00 |
|
GassiGiuseppe
|
6f219f634f
|
Added attention_mask
|
2025-10-05 17:49:01 +02:00 |
|
Christian Risi
|
c60da8ba82
|
Refactoring
|
2025-10-05 15:40:29 +02:00 |
|