Ross Wightman
0ba73e6bcb
Update README.md
2021-10-19 14:38:56 -07:00
Ross Wightman
b6caa356d2
Fixed eca_botnext26ts_256 weights added, 79.27
2021-10-19 12:44:28 -07:00
Ross Wightman
c02334d9fa
Add weights for regnetz_d and haloregnetz_c, update regnetz_c weights. Add commented PyTorch XLA code for halo attention
2021-10-19 12:32:09 -07:00
Ross Wightman
02daf2ab94
Add option to include relative pos embedding in the attention scaling as per references. See discussion #912
2021-10-12 15:37:01 -07:00
Ross Wightman
2c33ca6d8c
Merge pull request #913 from ground0state/master
...
Fix bugs that Mixup does not work when device is cpu
2021-10-12 14:09:56 -07:00
masafumi
047a5ec05f
Fix bugs that Mixup does not work device=cpu
2021-10-12 23:51:46 +09:00
Ross Wightman
cd34913278
Remove some outdated comments, botnet networks working great now.
2021-10-11 22:43:41 -07:00
Ross Wightman
6ed4cdccca
Update lambda_resnet26t weights with better set
2021-10-10 16:32:54 -07:00
Ross Wightman
288ece0e9f
Merge pull request #910 from tmp-iclr/master
...
Add ConvMixer
2021-10-10 16:00:58 -07:00
ICLR Author
44d6d51668
Add ConvMixer
2021-10-09 21:09:51 -04:00
Ross Wightman
a85df34993
Update lambda_resnet26rpt weights to 78.9, add better halonet26t weights at 79.1 with tweak to attention dim
2021-10-08 17:44:13 -07:00
Ross Wightman
38804c721b
Checkpoint clean fn useable stand alone
2021-10-08 17:43:53 -07:00
Ross Wightman
b544ad4d3f
regnetz model default cfg tweaks
2021-10-06 21:14:59 -07:00
Ross Wightman
d80653cb99
Merge branch 'alexander-soare-freeze-functionality'
2021-10-06 17:01:41 -07:00
Ross Wightman
e5da481073
Small post-merge tweak for freeze/unfreeze, add to __init__ for utils
2021-10-06 17:00:27 -07:00
Ross Wightman
5ca72dcc75
Merge branch 'freeze-functionality' of https://github.com/alexander-soare/pytorch-image-models into alexander-soare-freeze-functionality
2021-10-06 16:51:03 -07:00
Ross Wightman
e2b8d44ff0
Halo, bottleneck attn, lambda layer additions and cleanup along w/ experimental model defs
...
* align interfaces of halo, bottleneck attn and lambda layer
* add qk_ratio to all of above, control q/k dim relative to output dim
* add experimental haloregnetz, and trionet (lambda + halo + bottle) models
2021-10-06 16:32:48 -07:00
Ross Wightman
e0b3a3fab3
Make test-pooling flag for validate.py opt in
2021-10-06 16:12:20 -07:00
Alexander Soare
431e60c83f
Add acknowledgements for freeze_batch_norm inspiration
2021-10-06 14:28:49 +01:00
Ross Wightman
fbf59c04ee
Change crop ratio on correct resnet50 variant.
2021-10-04 22:31:08 -07:00
Ross Wightman
ae1ff5792f
Clean a1/a2/3 rsb _0 checkpoints properly, fix v2 loading.
2021-10-04 16:46:00 -07:00
Ross Wightman
d123042605
Update README.md
2021-10-03 21:38:47 -07:00
Ross Wightman
cd638d50a5
Merge pull request #880 from rwightman/fixes_bce_regnet
...
A collection of fixes, model experiments, etc
2021-10-03 19:37:01 -07:00
Ross Wightman
93901e992f
Version bump to 0.5.0 for pending release post RSB and ATTN updates
2021-10-03 17:34:57 -07:00
Ross Wightman
da0d39bedd
Update default crop_pct for byoanet
2021-10-03 17:33:16 -07:00
Ross Wightman
cc9bedf373
Add initial ResNet Strikes Back weights for ResNet50 and ResNetV2-50 models
2021-10-03 17:32:02 -07:00
Ross Wightman
64495505b7
Add updated lambda resnet26 and botnet26 checkpoints with fixes applied
2021-10-03 17:31:39 -07:00
Ross Wightman
b2094f4ee8
support bits checkpoints in avg/load
2021-10-03 17:31:22 -07:00
Ross Wightman
007bc39323
Some halo and bottleneck attn code cleanup, add halonet50ts weights, use optimal crop ratios
2021-10-02 15:51:42 -07:00
Alexander Soare
6d2acec1bb
Fix ordering of tests
2021-10-02 16:10:11 +01:00
Alexander Soare
65c3d78b96
Freeze unfreeze functionality finalized. Tests added
2021-10-02 15:55:08 +01:00
Alexander Soare
0cb8ea432c
wip
2021-10-02 15:55:08 +01:00
Ross Wightman
d9abfa48df
Make broadcast_buffers disable its own flag for now (needs more testing on interaction with dist_bn)
2021-10-01 13:43:55 -07:00
Ross Wightman
b1c2e3eb92
Match rel_pos_indices attr rename in conv branch
2021-09-30 23:19:05 -07:00
Ross Wightman
b49630a138
Add relative pos embed option to LambdaLayer, fix last transpose/reshape.
2021-09-30 22:45:09 -07:00
Ross Wightman
d657e2cc0b
Remove dead code line from efficientnet
2021-09-30 21:54:42 -07:00
Ross Wightman
0ca687f224
Make 'regnetz' model experiments closer to actual RegNetZ, bottleneck expansion, expand from in_chs, no shortcut on stride 2, tweak model sizes
2021-09-30 21:49:38 -07:00
Ross Wightman
b5bf4dce98
Merge pull request #898 from leondgarse/master
...
Remove a duplicate layer creation in byobnet.py
2021-09-30 13:32:15 -07:00
leondgarse
51eaf9360d
Remove a duplicate layer creation in byobnet.py
...
`self.conv2_kxk` is repeated in `byobnet.py`. Remove the duplicate code.
2021-09-30 18:30:48 +08:00
Ross Wightman
b81e79aae9
Fix bottleneck attn transpose typo, hopefully these train better now..
2021-09-28 16:38:41 -07:00
Ross Wightman
80075b0b8a
Add worker_seeding arg to allow selecting old vs updated data loader worker seed for (old) experiment repeatability
2021-09-28 16:37:45 -07:00
Ross Wightman
6478bcd02c
Fix regnetz_d conv layer name, use inception mean/std
2021-09-26 14:54:17 -07:00
Ross Wightman
3f9959cdd2
Merge pull request #882 from ShoufaChen/master
...
fix `use_amp`
2021-09-25 21:37:44 -07:00
Shoufa Chen
908563d060
fix `use_amp`
...
Fix https://github.com/rwightman/pytorch-image-models/issues/881
2021-09-26 12:32:22 +08:00
Ross Wightman
0387e6057e
Update binary cross ent impl to use thresholding as an option (convert soft targets from mixup/cutmix to 0, 1)
2021-09-23 15:45:39 -07:00
Ross Wightman
5d6983c462
Batch validate a list of files if model is a text file with model per line
2021-09-23 15:45:17 -07:00
Ross Wightman
f8a63a3b71
Add worker_init_fn to loader for numpy seed per worker
2021-09-23 15:44:38 -07:00
Ross Wightman
515121cca1
Use reshape instead of view in std_conv, causing issues in recent PyTorch in channels_last
2021-09-23 15:43:48 -07:00
Ross Wightman
da06cc61d4
ResNetV2 seems to work best without zero_init residual
2021-09-23 15:43:22 -07:00
Ross Wightman
8e11da0ce3
Add experimental RegNetZ(ish) models for training / perf trials.
2021-09-23 15:42:57 -07:00