Team Ai
Modelpublic

microsoft/phi-1

sourceHugging Facemitupdated 11mo agoView on Hugging Face
222likes25kdownloads
50 commits on main
d4c0adc11mo ago

Update README.md (#15)

gugarosa, bsnelling
c542f6111mo ago

Upload data_summary_card.md (#14)

gugarosa, bsnelling
b9ac0e62y ago

fix(config): Removes auto_map since it is not used anymore.

gugarosa
cb3004d2y ago

Update README.md

gugarosa
8399ff62y ago

Delete modeling_phi.py

gugarosa
8c835682y ago

Delete configuration_phi.py

gugarosa
92f23552y ago

Update README.md

gugarosa
0c849e42y ago

Delete pytorch_model.bin

gugarosa
b336aa72y ago

Adding `safetensors` variant of this model (#9)

gugarosa, SFconvertbot
a1a59313y ago

Update LICENSE

gugarosa
ac0061a3y ago

Update README.md

gugarosa
ffff2603y ago

Update README.md

gugarosa
2aef7073y ago

Update README.md

gugarosa
64cc1d83y ago

Update README.md

gugarosa
0d846e03y ago

Update config.json

gugarosa
944a0133y ago

Update modeling_phi.py

gugarosa
fb32aac3y ago

Update README.md

gugarosa
07d93633y ago

Update README.md

gugarosa
03b9f693y ago

Update modeling_phi.py

gugarosa
54bed1a3y ago

Update modeling_phi.py

gugarosa
50bb2673y ago

Update modeling_phi.py

gugarosa
2cfa65f3y ago

Upload modeling_phi.py

gugarosa
957a7833y ago

Delete Research License.docx

gugarosa
1cb06683y ago

Upload 5 files

gugarosa
3e53f583y ago

Update config.json

gugarosa
e5752413y ago

Update modeling_phi.py

gugarosa
4e6ed663y ago

Update modeling_phi.py

gugarosa
fbf395a3y ago

Update configuration_phi.py

gugarosa
b9088383y ago

fix(root): Fixes relative paths.

gugarosa
8a2c68b3y ago

chore(root): Updates files to internal transformers implementation.

gugarosa
530294c3y ago

Update README.md

gugarosa
b3ebf083y ago

Upload 4 files

gugarosa
304b0583y ago

Update README.md

gugarosa
654b6903y ago

Update README.md

gugarosa
eac52183y ago

chore(readme): Updates with clear information.

gugarosa
e8a38cd3y ago

Disables inference API to prevent mismatch with HF implementation.

gugarosa
f4e55a83y ago

fix(modeling_phi): Fixes initial generation with length larger than context length.

gugarosa
ecfe56e3y ago

fix(modeling_phi): Fixes cached generation when above maximum context length.

gugarosa
759d1483y ago

Fixes exceeding maximum sequence length when using generate().

gugarosa
5819d043y ago

Uses native torch decorator for disabling autocast.

gugarosa
67ecc753y ago

Adds disable_autocast support for different device types.

gugarosa
b5c51613y ago

Fixes any potential overflow when calculating attention weights.

gugarosa
470e18a3y ago

Delete modeling_mixformer_sequential.py

gugarosa
bd98e4e3y ago

Delete configuration_mixformer_sequential.py

gugarosa
34b22f43y ago

Upload pytorch_model.bin

gugarosa
bbace883y ago

Update to new model interface.

gugarosa
8d2c4ce3y ago

Improves type hinting on configuration arguments.

gugarosa
9ed59873y ago

Fixes flash-attn import with a try/except statement

gugarosa
90c38d93y ago

Adds support for flash-attn rotary embedding and fused dense layers.

gugarosa
371fd513y ago

Adds support for MQA/GQA and attention mask during training / fine-tuning.

gugarosa