fix the bug of named_parameters #1710

mileyan · 2017-06-03T07:00:34Z

When I use the named_parametes to modify the lr and weight decay, I will face a bug. Because the value of the named_parameters return is torch.nn.paramter.Parameter, not a generator of the Parameter.
the error information:
https://discuss.pytorch.org/t/problem-on-different-learning-rate-and-weight-decay-in-different-layers/3619

When I use the named_parametes to modify the lr and weight decay, I will face a bug. Because the value of the named_parameters return is torch.nn.paramter.Parameter, not a generator of the Parameter.

apaszke

Can you also add a test case for this?

torch/optim/optimizer.py

@@ -33,7 +33,10 @@ def __init__(self, params, defaults):

        param_set = set()
        for group in self.param_groups:
-            group['params'] = list(group['params'])
+            if isinstance(group['params'], torch.nn.paramter.Parameter):


apaszke · 2017-06-03T12:00:49Z

@pytorchbot test this please

apaszke · 2017-06-03T12:01:44Z

It's still missing the test case. I'll merge it once it's added.

soumith · 2017-06-06T03:48:18Z

i've added a test for this and merged your PR into master. Thank you.

soumith · 2017-06-06T03:48:32Z

merged via a76098a

…ytorch#1710) * unify segmented and single fusion path

Syncing nvfuser devel branch to upstream master. https://github.com/csarofeen/pytorch/ A few bigger updates: 1. Initial support of cp.async and cp.async.wait: csarofeen#1619 2. Emulate ampere's mma 16816 with Turing's mma 1688, for a unified interface: csarofeen#1643 3. Extending the infrastructure to support mma operators on turing and ampere arch: csarofeen#1440 Commits that's actually in this PR from the csarofeen branch ``` * dd23252 (csarofeen/devel) Fusion Segmenter: Unify single kernel and multi-kernel runtime path (#1710) * b3d1c3f Fix missing cooperative launch (#1726) * dc670a2 Async gmem copy support on sm80+ (#1619) * 5e6a8da Add turing mma support and test (#1643) * d6d6b7d Fix rFactor when there are indirect root domain(s), and refactor (#1723) * 7093e39 Mma op integration on ampere (#1440) * fade8da patch python test for bfloat16 (#1724) * 8fbd0b1 Fine-grained kernel profiling (#1720) * 77c1b4f Adding dry run mode to skip arch dependent checks (#1702) * 151d95b More precise concretization analysis (#1719) * f4d3630 Enable complex python tests (#1667) * 4ceeee5 Minor bugfix in transform_rfactor.cpp (#1715) * 3675c70 Separate root domain and rfactor domain in TransformPrinter (#1716) * f68b830 Fix scheduling with polymorphic broadcast (#1714) * 4ab5ef7 updating_ci_machine (#1718) * 56585c5 Merge pull request #1711 from csarofeen/upstream_master_bump_0517 * 174d453 Allow using nvFuser on CUDA extension (#1701) * 18bee67 Validate LOOP concrete IDs have complete IterDomains (#1676) ``` Pull Request resolved: #78244 Approved by: https://github.com/csarofeen, https://github.com/malfet

Summary: Syncing nvfuser devel branch to upstream master. https://github.com/csarofeen/pytorch/ A few bigger updates: 1. Initial support of cp.async and cp.async.wait: csarofeen#1619 2. Emulate ampere's mma 16816 with Turing's mma 1688, for a unified interface: csarofeen#1643 3. Extending the infrastructure to support mma operators on turing and ampere arch: csarofeen#1440 Commits that's actually in this PR from the csarofeen branch ``` * dd23252 (csarofeen/devel) Fusion Segmenter: Unify single kernel and multi-kernel runtime path (#1710) * b3d1c3f Fix missing cooperative launch (#1726) * dc670a2 Async gmem copy support on sm80+ (#1619) * 5e6a8da Add turing mma support and test (#1643) * d6d6b7d Fix rFactor when there are indirect root domain(s), and refactor (#1723) * 7093e39 Mma op integration on ampere (#1440) * fade8da patch python test for bfloat16 (#1724) * 8fbd0b1 Fine-grained kernel profiling (#1720) * 77c1b4f Adding dry run mode to skip arch dependent checks (#1702) * 151d95b More precise concretization analysis (#1719) * f4d3630 Enable complex python tests (#1667) * 4ceeee5 Minor bugfix in transform_rfactor.cpp (#1715) * 3675c70 Separate root domain and rfactor domain in TransformPrinter (#1716) * f68b830 Fix scheduling with polymorphic broadcast (#1714) * 4ab5ef7 updating_ci_machine (#1718) * 56585c5 Merge pull request #1711 from csarofeen/upstream_master_bump_0517 * 174d453 Allow using nvFuser on CUDA extension (#1701) * 18bee67 Validate LOOP concrete IDs have complete IterDomains (#1676) ``` Pull Request resolved: #78244 Reviewed By: ejguan Differential Revision: D36678948 Pulled By: davidberard98 fbshipit-source-id: 0ccde965acbd31da67d99c6adb2eaaa888948105

mileyan added 2 commits June 3, 2017 15:00

fix the bug of named_parameters

13de032

When I use the named_parametes to modify the lr and weight decay, I will face a bug. Because the value of the named_parameters return is torch.nn.paramter.Parameter, not a generator of the Parameter.

Update optimizer.py

930fd65

apaszke suggested changes Jun 3, 2017

View reviewed changes

Update optimizer.py

02be20f

soumith closed this Jun 6, 2017

soumith reopened this Jun 6, 2017

soumith closed this Jun 6, 2017

jjsjann123 pushed a commit to jjsjann123/pytorch that referenced this pull request May 24, 2022

Fusion Segmenter: Unify single kernel and multi-kernel runtime path (p…

dd23252

…ytorch#1710) * unify segmented and single fusion path

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

fix the bug of named_parameters #1710

fix the bug of named_parameters #1710

mileyan commented Jun 3, 2017 •

edited

Loading

apaszke left a comment

This comment was marked as off-topic.

apaszke commented Jun 3, 2017

apaszke commented Jun 3, 2017

soumith commented Jun 6, 2017

soumith commented Jun 6, 2017

fix the bug of named_parameters #1710

fix the bug of named_parameters #1710

Conversation

mileyan commented Jun 3, 2017 • edited Loading

apaszke left a comment

Choose a reason for hiding this comment

This comment was marked as off-topic.

apaszke commented Jun 3, 2017

apaszke commented Jun 3, 2017

soumith commented Jun 6, 2017

soumith commented Jun 6, 2017

mileyan commented Jun 3, 2017 •

edited

Loading