[`GPTNeoX`] Faster rotary embedding for GPTNeoX (based on llama changes) #25830
Faster rotary embedding for GPTNeoX
eadf0610
there might be un-necessary moves from device
f3a4b42e
fixup
4fdbee69
ArthurZucker
changed the title Faster rotary embedding for GPTNeoX [`GPTNeoX`] Faster rotary embedding for GPTNeoX (based on llama changes) 3 years ago
fix dtype issue
3934d84a
Merge branch 'main' of https://github.com/huggingface/transformers in…
a1f513d9
add copied from statements
6ac79564
ArthurZucker
marked this pull request as ready for review 3 years ago
fox copies
58981571
oupsy
461f7384
gante
approved these changes
on 2023-08-31
add copied from Llama for scaled ones as well
98aea0e1
fixup
0cd14125
Merge branch 'main' of github.com:huggingface/transformers into impro…
0d0f9cd4
Merge branch 'main' of github.com:huggingface/transformers into impro…
5a87a9b2
fix
f74f0232
Merge branch 'main' of github.com:huggingface/transformers into impro…
6735cc01
fix copies
c83bcfd1
Merge branch 'improve-gpt-neox' of github.com:ArthurZucker/transforme…
35ebed7e
Merge branch 'improve-gpt-neox' of github.com:ArthurZucker/transforme…
165c6220
ArthurZucker
deleted the improve-gpt-neox branch 2 years ago
Assignees
No one assigned
Login to write a write a comment.
Login via GitHub