Henk
9e7eb80db4
Nerys V2 part 2
2022-06-25 14:03:19 +02:00
Henk
ecc6ee9474
Nerys V2
2022-06-25 13:47:49 +02:00
henk717
10e85db89d
Merge pull request #162 from VE-FORBRYDERNE/whitespace-cleanup
...
Story whitespace cleanup
2022-06-25 13:36:03 +02:00
henk717
a2d2ea0735
Typo Fix
2022-06-25 13:14:27 +02:00
henk717
2495b9380d
Nerys V2
2022-06-25 13:08:30 +02:00
vfbd
6acfa8c33c
Merge branch 'whitespace' into whitespace-cleanup
2022-06-24 12:44:30 -04:00
vfbd
6e138db1c0
Clean up whitespace in the editor as well
2022-06-24 12:44:00 -04:00
Henk
d3fce44095
Merge branch 'main' into united
2022-06-24 18:31:45 +02:00
Henk
8be0964427
AIDG Import Fix
2022-06-24 18:29:06 +02:00
vfbd
4b16600e49
Clean up whitespace at the end of actions when loading story
...
Specifically, we merge blank actions into the next action and we move
whitespace at the end of non-blank actions to the beginning of the next
action.
2022-06-24 12:03:35 -04:00
henk717
fca7c15fd3
Merge branch 'KoboldAI:main' into united
2022-06-24 18:00:37 +02:00
henk717
27e16aecf2
Model Cleaner (suggested by Shrinkarom)
2022-06-24 14:03:58 +02:00
henk717
c3af92e9af
Koboldai.org | New NeoX | Cloudflare
2022-06-24 10:43:06 +02:00
Henk
6a89ad5b94
Merge branch 'main' into united
2022-06-23 21:07:56 +02:00
henk717
ec0bc1cc17
Merge pull request #130 from VE-FORBRYDERNE/neox-badwords
...
GPT-NeoX HF model badwords fix
2022-06-23 21:05:42 +02:00
vfbd
3da885d408
GPT-NeoX HF model badwords fix
2022-06-23 15:02:43 -04:00
henk717
d94f29a68a
Merge pull request #127 from VE-FORBRYDERNE/tracer
...
Fix JAX UnexpectedTracerError
2022-06-23 19:29:51 +02:00
henk717
8098f4ec8f
Merge branch 'KoboldAI:main' into united
2022-06-23 17:20:48 +02:00
henk717
1d41966d88
Merge pull request #129 from VE-FORBRYDERNE/budget
...
Account for lnheader in budget calculation
2022-06-23 12:39:20 +02:00
vfbd
0eb9f8a879
Account for lnheader in budget calculation
2022-06-22 19:16:24 -04:00
henk717
3de22f2b27
Merge pull request #160 from VE-FORBRYDERNE/gc
...
Delete all torch tensors before loading model
2022-06-22 18:38:21 +02:00
vfbd
53034ee533
Delete all torch tensors before loading model
2022-06-22 12:07:36 -04:00
henk717
f127918114
Merge pull request #159 from VE-FORBRYDERNE/fairseq
...
Don't blacklist </s> token in "s" newline mode
2022-06-22 17:41:20 +02:00
vfbd
922394c68f
Don't blacklist </s> token in "s" newline mode
2022-06-22 11:23:03 -04:00
Henk
d4e18360f0
HF NeoX Support
2022-06-22 01:46:40 +02:00
henk717
5f9a116052
Merge pull request #128 from VE-FORBRYDERNE/neox
...
TPU support for HF GPT-NeoX model
2022-06-22 01:39:53 +02:00
Gnome Ann
8c594c6869
Correct the padding token for GPT-NeoX
2022-06-21 19:37:43 -04:00
Gnome Ann
a7f667c34c
Use NeoX badwords when loading from HF GPT-NeoX model
2022-06-21 19:33:25 -04:00
Gnome Ann
5e3c7c07ae
Merge branch 'main' into neox
2022-06-21 19:30:51 -04:00
henk717
f1d0a327f8
Merge branch 'KoboldAI:main' into united
2022-06-21 23:34:32 +02:00
Henk
75bc472a9f
Transformers bump to 4.20.1
...
Transformers issued an important change for the OPT models breaking their compatibility with all older versions. In order for people to be able to use all models on the menu they need 4.20.1 so this is now forced in the dependencies making the update easier.
2022-06-21 23:33:38 +02:00
henk717
b5b8e5a30b
Merge branch 'KoboldAI:main' into united
2022-06-21 23:19:57 +02:00
henk717
2be1f5088f
Merge pull request #126 from VE-FORBRYDERNE/opt
...
Update OPT models and fix 20B model on TPU
2022-06-21 23:19:03 +02:00
Gnome Ann
33a2a318db
Fix 20B TPU model
2022-06-21 17:16:01 -04:00
Gnome Ann
a7e3ef71aa
Add final layer norm to OPT
2022-06-21 16:36:26 -04:00
henk717
37eb47d0d3
Merge pull request #157 from VE-FORBRYDERNE/sp-fix
...
Bug fixes and new soft prompt implementation
2022-06-21 22:20:36 +02:00
Gnome Ann
8593bf339b
Another typo fix
2022-06-21 15:36:25 -04:00
Gnome Ann
7e0ded6b47
Typo fix
2022-06-21 15:12:55 -04:00
Gnome Ann
91643be10a
Change soft prompt implementation to a more universal one
2022-06-21 15:03:43 -04:00
Gnome Ann
0ea4fa9c87
Automatically calculate badwords and pad_token_id
2022-06-21 14:35:52 -04:00
Gnome Ann
ea7d278ff4
Fix 20B TPU model
2022-06-21 13:16:45 -04:00
Gnome Ann
6b172306f6
move_model_to_devices no longer crashes if you don't have accelerate
2022-06-21 13:15:46 -04:00
henk717
f2c5bb5cb7
Merge pull request #156 from VE-FORBRYDERNE/accelerate
...
Accelerate disk cache support
2022-06-21 00:31:50 +02:00
Gnome Ann
ff69e9fbfe
Put layers_module_names, module_names and named_buffers in utils.py
2022-06-20 17:17:42 -04:00
Gnome Ann
1620ac4148
Lazy loader needs to cache named buffers of layers in the disk cache
2022-06-20 17:08:52 -04:00
Gnome Ann
ab5ab79003
Set primary device to CPU if in CPU-only mode
2022-06-20 16:25:01 -04:00
Gnome Ann
bd7d7b41a1
Don't enable accelerate if no layers are in disk cache or GPUs
2022-06-20 16:21:44 -04:00
Gnome Ann
90fd8b1845
Disk cache support in CPU-only mode
2022-06-20 16:06:09 -04:00
Gnome Ann
af07d7a15f
Disk cache support for computers with at least one GPU
2022-06-20 14:49:54 -04:00
Gnome Ann
47a58a36b8
Add disk cache slider
2022-06-19 22:53:30 -04:00