Lysandre Debut and Nicolas Patry
5a63232a8a
Fix QA argument handler ( #8765 )
...
* Fix QA argument handler
* Attempt to get a better fix for QA (#8768 )
Co-authored-by: Nicolas Patry <patry.nicolas@protonmail.com >
2020-11-29 20:06:10 -05:00
Lysandre Debut
e46890f699
MT5 should have an autotokenizer ( #8743 )
...
* MT5 should have an autotokenizer
* Different configurations should be able to point to same tokenizers
2020-11-24 09:51:34 -05:00
Lysandre Debut
df2cdd84f3
Fix slow tests v2 ( #8746 )
...
* Fix BART test
* Fix MBART tests
* Remove erroneous line from yaml
* Update tests/test_modeling_bart.py
* Quality
2020-11-24 09:51:28 -05:00
LysandreJik
c6e2876cd4
TF BERT test update
2020-11-23 18:19:54 -05:00
LysandreJik
5580cccd81
Update TF BERT test
2020-11-23 18:19:34 -05:00
Stas Bekman
ccc4f64044
consistent ignore keys + make private ( #8737 )
...
* consistent ignore keys + make private
* style
* - authorized_missing_keys => _keys_to_ignore_on_load_missing
- authorized_unexpected_keys => _keys_to_ignore_on_load_unexpected
* move public doc of private attributes to private comment
2020-11-23 17:55:15 -05:00
Sylvain Gugger and Lysandre Debut
3408e6ffcd
Change default cache path ( #8734 )
...
* Change default cache path
* Document changes
* Apply suggestions from code review
Co-authored-by: Lysandre Debut <lysandre@huggingface.co >
Co-authored-by: Lysandre Debut <lysandre@huggingface.co >
2020-11-23 17:54:45 -05:00
Santiago Castro
a986b02e49
Fix many typos ( #8708 )
2020-11-23 17:54:20 -05:00
Sylvain Gugger
b6ec39e41f
Document adam betas TrainingArguments ( #8688 )
2020-11-23 17:53:49 -05:00
Sylvain Gugger
f80ea27f80
Add sentencepiece to the CI and fix tests ( #8672 )
...
* Fix the CI and tests
* Fix quality
* Remove that m form nowhere
2020-11-23 17:53:27 -05:00
Sylvain Gugger
0603564e93
Merge remote-tracking branch 'origin/master'
2020-11-19 12:18:57 -05:00
Sylvain Gugger
1e08af383a
Forgot to save...
2020-11-19 12:18:50 -05:00
LysandreJik
d86b5ffc6f
Release: v4.0.0-rc-1
v4.0.0-rc-1
2020-11-19 12:00:07 -05:00
Sylvain Gugger
cb3e5c33f7
Fix a few last paths for the new repo org ( #8666 )
2020-11-19 11:56:42 -05:00
Matthias
a79a96ddaa
fix small typo ( #8644 )
...
Fixed a small typo on the XLNet and permutation language modelling section
2020-11-19 11:24:11 -05:00
Sylvain Gugger
4208f496ee
Better filtering of the model outputs in Trainer ( #8633 )
...
* Better filtering of the model outputs in Trainer
* Fix examples tests
* Add test for Lysandre
2020-11-19 10:43:15 -05:00
Lysandre Debut and patrickvonplaten
f2e07e7272
Fix a bunch of slow tests ( #8634 )
...
* CI should install `sentencepiece`
* Requiring TF
* Fixing some TFDPR bugs
* remove return_dict=False/True hack
Co-authored-by: patrickvonplaten <patrick.v.platen@gmail.com >
2020-11-19 10:41:41 -05:00
elk-cloner and Patrick von Platen
5362bb8a6b
Tf longformer for sequence classification ( #8231 )
...
* working on LongformerForSequenceClassification
* add TFLongformerForMultipleChoice
* add TFLongformerForTokenClassification
* use add_start_docstrings_to_model_forward
* test TFLongformerForSequenceClassification
* test TFLongformerForMultipleChoice
* test TFLongformerForTokenClassification
* remove test from repo
* add test and doc for TFLongformerForSequenceClassification, TFLongformerForTokenClassification, TFLongformerForMultipleChoice
* add requested classes to modeling_tf_auto.py
update dummy_tf_objects
fix tests
fix bugs in requested classes
* pass all tests except test_inputs_embeds
* sync with master
* pass all tests except test_inputs_embeds
* pass all tests
* pass all tests
* work on test_inputs_embeds
* fix style and quality
* make multi choice work
* fix TFLongformerForTokenClassification signature
* fix TFLongformerForMultipleChoice, TFLongformerForSequenceClassification signature
* fix mult choice
* fix mc hint
* fix input embeds
* fix input embeds
* refactor input embeds
* fix copy issue
* apply sylvains changes and clean more
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com >
2020-11-19 10:37:27 -05:00
Quentin Lhoest
62cd9ce9f8
fix missing return dict ( #8653 )
2020-11-19 15:17:18 +01:00
Amine Abdaoui
0c2677f529
[model card] : fix bert-base-15lang-cased ( #8655 )
...
the table was badly formatted because of a single line break
2020-11-19 05:41:02 -05:00
Amine Abdaoui
0a80959bdd
Add cards for all Geotrend models ( #8617 )
...
* docs(bert-base-15lang-cased): add model card
* add cards for all Geotrend models
* [model cards] fix language tag for all Geotrend models
2020-11-19 04:47:24 -05:00
cronoik
dcc9c64299
Updated the Extractive Question Answering code snippets ( #8636 )
...
* Updated the Extractive Question Answering code snippets
The Extractive Question Answering code snippets do not work anymore since the models return task-specific output objects. This commit fixes the pytorch and tensorflow examples but adding `.values()` to the model call.
* Update task_summary.rst
2020-11-18 18:56:47 -05:00
Tim Isbister
28d16e7ac5
Update README.md ( #8635 )
2020-11-18 18:35:23 -05:00
cronoik
b290195ac7
grammar ( #8639 )
2020-11-18 18:04:25 -05:00
Stas Bekman
d86d57faa3
[s2s] distillation apex breaks return_dict obj ( #8631 )
...
* apex breaks return_dict obj
* style
2020-11-18 12:51:29 -08:00
Perez Ogayo and Julien Chaumond
bf3611b2ab
Created ModelCard for Hel-ach-en MT model ( #8496 )
...
* Updated ModelCard
* Apply suggestions from code review
Co-authored-by: Julien Chaumond <chaumond@gmail.com >
2020-11-18 14:42:13 -05:00
Yifan Peng
c95b26a719
Create README.md ( #8362 )
2020-11-18 13:37:14 -05:00
Manuel Romero and Julien Chaumond
fdbbb6c17a
Model card: T5-base fine-tuned on QuaRTz ( #8369 )
...
* Model card: T5-base fine-tuned on QuaRTz
* Update model_cards/mrm8488/t5-base-finetuned-quartz/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com >
2020-11-18 13:34:27 -05:00
Yifan Peng
6e6d24c5d8
Create README.md ( #8363 )
2020-11-18 13:33:04 -05:00
Divyanshu Kakwani
35fd3d64e3
Add model card for ai4bharat/indic-bert ( #8464 )
2020-11-18 13:28:49 -05:00
dartrevan
38f01dfe03
Update README.md ( #8405 )
...
* Update README.md
* Update README.md
2020-11-18 13:23:08 -05:00
Abhilash Majumder and Julien Chaumond
2d8fbf012a
Model Card for abhilash1910/financial_roberta ( #8625 )
...
* Model Card for abhilash1910/financial_roberta
* Update model_cards/abhilash1910/financial_roberta/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com >
2020-11-18 13:22:28 -05:00
Vishal Singh
26dc6593f3
Update README.md ( #8544 )
...
Modified Model in Action section. The class `AutoModelWithLMHead` is deprecated so changed it to `AutoModelForSeq2SeqLM` for encoder-decoder models. Removed duplicate eos token.
2020-11-18 13:19:32 -05:00
smanjil and Julien Chaumond
6c8fad4f0d
replace performance table with markdown ( #8565 )
...
* replace performance table with markdown
* Update model_cards/smanjil/German-MedBERT/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com >
2020-11-18 13:17:46 -05:00
hhou435
e7f77fc52a
model_cards for Chinese Couplet and Poem GPT2 models ( #8620 )
2020-11-18 13:06:30 -05:00
Sylvain Gugger
a0c62d2493
Fix training from scratch in new scripts ( #8623 )
2020-11-18 12:15:26 -05:00
Sylvain Gugger
1e62e999e8
Fixes the training resuming with gradient accumulation ( #8624 )
2020-11-18 12:00:11 -05:00
Patrick von Platen
cdfa56afe0
[Tokenizer Doc] Improve tokenizer summary ( #8622 )
...
* improve summary
* small fixes
* cleaned line length
* correct "" formatting
* apply sylvains suggestions
2020-11-18 17:14:15 +01:00
Nicola De Cao and Patrick von Platen
2f9d49b389
Adding PrefixConstrainedLogitsProcessor ( #8529 )
...
* Adding PrefixConstrainedLogitsProcessor
* fixing RAG and style_doc
* fixing black (v20 instead of v19)
* Improving doc in generation_logits_process.py
* Improving docs and typing in generation_utils.py
* docs improvement
* adding test and fixing doc typo
* fixing doc_len
* isort on test
* fixed test
* improve docstring a bit
Co-authored-by: Patrick von Platen <patrick.v.platen@gmail.com >
2020-11-18 17:06:25 +01:00
Julien Plu
3bc1540070
New TF loading weights ( #8490 )
...
* New TF loading weights
* apply style
* Better naming
* Largely comment the loading method
* Apply style
* Address Patrick's comments
* Remove useless line of code
* Update Docstring
* Address Sylvain's and Lysandre's comments
* Simplify the names computation
* Typos
2020-11-18 10:48:31 -05:00
Ratthachat (Jung)
0df91ee4f7
self.self.activation_dropout -> self.activation_dropout ( #8611 )
...
(one line typo)
2020-11-18 10:30:29 -05:00
Stas Bekman
cdf1b7ae82
fix to adjust for #8530 changes ( #8612 )
2020-11-18 10:25:00 -05:00
Stas Bekman
2819da02f7
[s2s] broken test ( #8613 )
2020-11-18 10:15:53 -05:00
Michał Pogoda
9fa3ed1a7f
Fix missing space in multiline warning ( #8593 )
...
Multiline string informing about missing PyTorch/TensorFlow had missing space.
2020-11-18 10:09:26 -05:00
Sylvain Gugger
8fcb6935a1
Fix DataCollatorForLanguageModeling ( #8621 )
2020-11-18 10:02:50 -05:00
Benjamin Minixhofer
f6fe41c96b
Reset loss to zero on logging in Trainer to avoid bfloat16 issues ( #8561 )
...
* make tr_loss regular float
* Revert "make tr_loss regular float"
This reverts commit c9d7ccfaf0c4387187b0841694f01ec0ffd5f4ba.
* reset loss at each logging step
* keep track of total loss with _total_loss_scalar
* add remaining tr_loss at the end
2020-11-18 09:58:08 -05:00
cronoik
b592728eff
Fixed link to the wrong paper. ( #8607 )
2020-11-17 19:00:44 -05:00
Sylvain Gugger
0512444ee5
Remove old doc
2020-11-17 17:34:25 -05:00
Caitlin Ostroff and Julien Chaumond
5cf9c79665
Add Harry Potter Model Card ( #8605 )
...
* Add Harry Potter Model
* Update model_cards/ceostroff/harry-potter-gpt2-fanfiction/README.md
* Update model_cards/ceostroff/harry-potter-gpt2-fanfiction/README.md
* Update model_cards/ceostroff/harry-potter-gpt2-fanfiction/README.md
Co-authored-by: Julien Chaumond <chaumond@gmail.com >
2020-11-17 16:50:58 -05:00
Sylvain Gugger and LysandreJik
dd52804f5f
Remove deprecated ( #8604 )
...
* Remove old deprecated arguments
Co-authored-by: LysandreJik <lysandre.debut@reseau.eseo.fr >
* Remove needless imports
* Fix tests
Co-authored-by: LysandreJik <lysandre.debut@reseau.eseo.fr >
2020-11-17 15:11:29 -05:00