Spruce up docs for emulate_precision_casts by ezyang · Pull Request #145579 · pytorch/pytorch

ezyang · 2025-01-24T02:20:58Z

Stack from ghstack (oldest at bottom):

-> Spruce up docs for emulate_precision_casts #145579

Signed-off-by: Edward Z. Yang ezyang@meta.com

cc @voznesenskym @penguinwu @EikanWang @jgong5 @Guobing-Chen @XiaobingSuper @zhuhaozhe @blzheng @wenzhe-nrv @jiayisunx @ipiszy @yf225 @chenyang78 @kadeng @muchulee8 @ColinPeppler @amjames @desertfire @chauhang @aakhundov

[ghstack-poisoned]

pytorch-bot · 2025-01-24T02:21:02Z

🔗 Helpful Links

🧪 See artifacts and rendered test results at hud.pytorch.org/pr/145579

📄 Preview Python docs built from this PR
📄 Preview C++ docs built from this PR
❓ Need help or want to give feedback on the CI? Visit the bot commands wiki or our office hours

Note: Links to docs will display an error until the docs builds have been completed.

✅ No Failures

As of commit 77abf9b with merge base d6bea39 ():
💚 Looks good so far! There are no failures yet. 💚

This comment was automatically generated by Dr. CI and updates every 15 minutes.

Signed-off-by: Edward Z. Yang <ezyang@meta.com> ghstack-source-id: 28df7aa Pull Request resolved: #145579

ezyang · 2025-01-24T15:21:05Z

@pytorchbot merge

pytorchmergebot · 2025-01-24T15:22:52Z

Merge started

Your change will be merged once all checks pass (ETA 0-4 Hours).

Learn more about merging in the wiki.

Questions? Feedback? Please reach out to the PyTorch DevX Team

Advanced Debugging

Check the merge workflow status
here

bdhirsh · 2025-01-24T20:29:25Z

torch/_inductor/config.py

-# For multiple, fused pointwise nodes, inductor will elide the intermediary upcasts and downcasts
-# Typically this should be closer to fp64 ref numerics. However, it can be useful for debugging
-# to emulate the eager numerics.
+# Mode to emulate PyTorch eager numerics when doing lower precision compute


to what extent do you think we should have a flag (this one or another) for inductor to emulate all eager numerics? In particular, even if there is a potential perf cost?

One that came up recently was var_mean, which it looks like inductor lowers in a way that gives slightly different (but more accurate?) numerics:

import torch torch._inductor.config.emulate_precision_casts = True class GraphModule(torch.nn.Module): def forward(self, inp): return torch.ops.aten.var_mean.correction(inp, [3], correction = 0, keepdim = True) from torch._dynamo.testing import rand_strided inp = rand_strided((1, 256, 256, 144,), (9437184, 36864, 144, 1,), device='cuda:0', dtype=torch.float32) m = GraphModule() out_eager = m(inp) out_ref = m(inp.to(dtype=torch.float64)) out_compile = torch.compile(m)(inp) print(torch.allclose(out_eager[0], out_compile[0])) print(torch.allclose(out_ref[0], out_compile[0].to(dtype=torch.float64))) print(torch.allclose(out_ref[0], out_eager[0].to(dtype=torch.float64))) print(torch.max(torch.abs(out_eager[0] - out_compile[0]))) print(torch.max(torch.abs(out_ref[0] - out_compile[0]))) print(torch.max(torch.abs(out_ref[0] - out_eager[0]))) # prints: True True True tensor(3.5763e-07, device='cuda:0') tensor(2.3738e-07, device='cuda:0', dtype=torch.float64) tensor(2.7881e-07, device='cuda:0', dtype=torch.float64)

I think this would be very useful lol

Signed-off-by: Edward Z. Yang <ezyang@meta.com> Pull Request resolved: pytorch#145579 Approved by: https://github.com/gchanan

eellison · 2025-01-29T18:21:50Z

torch/_inductor/config.py

+# and downcasting after.  When two low precision operators are fused together,
+# Inductor will elide the downcast-upcast pairs (effectively a precision
+# truncation) that would occur between these two operators.  Typically,
+# Inductor's behavior should be closer to fp64 ref numerics.  However, with


if we're going to expand maybe we should actually have a doc somewhere instead of just a longer config str

Update

77abf9b

[ghstack-poisoned]

ezyang added a commit that referenced this pull request Jan 24, 2025

Spruce up docs for emulate_precision_casts

305bf08

Signed-off-by: Edward Z. Yang <ezyang@meta.com> ghstack-source-id: 28df7aa Pull Request resolved: #145579

pytorch-bot bot added ciflow/inductor module: inductor labels Jan 24, 2025

github-actions bot requested review from SherlockNoMad, albanD, antoniojkim, bdhirsh and miladm January 24, 2025 02:21

pytorch-bot bot temporarily deployed to upload-benchmark-results January 24, 2025 02:50 Inactive

ezyang added the topic: not user facing topic category label Jan 24, 2025

gchanan approved these changes Jan 24, 2025

View reviewed changes

pytorch-bot bot added the ciflow/trunk Trigger trunk jobs on your pull request label Jan 24, 2025

pytorchmergebot added the merging label Jan 24, 2025

pytorch-bot bot temporarily deployed to upload-benchmark-results January 24, 2025 15:51 Inactive

albanD removed their request for review January 24, 2025 16:23

pytorchmergebot added the Merged label Jan 24, 2025

pytorchmergebot closed this in cf063d4 Jan 24, 2025

pytorchmergebot removed the merging label Jan 24, 2025

bdhirsh reviewed Jan 24, 2025

View reviewed changes

nWEIdia pushed a commit to nWEIdia/pytorch that referenced this pull request Jan 27, 2025

Spruce up docs for emulate_precision_casts (pytorch#145579)

bdd5de0

Signed-off-by: Edward Z. Yang <ezyang@meta.com> Pull Request resolved: pytorch#145579 Approved by: https://github.com/gchanan

eellison reviewed Jan 29, 2025

View reviewed changes

github-actions bot deleted the gh/ezyang/3071/head branch March 1, 2025 02:10

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Spruce up docs for emulate_precision_casts#145579

Spruce up docs for emulate_precision_casts#145579
ezyang wants to merge 1 commit intogh/ezyang/3071/basefrom
gh/ezyang/3071/head

ezyang commented Jan 24, 2025 •

edited by pytorch-bot bot

Loading

Uh oh!

pytorch-bot bot commented Jan 24, 2025 •

edited

Loading

Uh oh!

ezyang commented Jan 24, 2025

Uh oh!

pytorchmergebot commented Jan 24, 2025

Uh oh!

bdhirsh Jan 24, 2025

Uh oh!

ezyang Jan 24, 2025

Uh oh!

eellison Jan 29, 2025

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

5 participants

Conversation

ezyang commented Jan 24, 2025 • edited by pytorch-bot bot Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

Uh oh!

pytorch-bot bot commented Jan 24, 2025 • edited Loading Uh oh! There was an error while loading. Please reload this page.

Uh oh!

🔗 Helpful Links

🧪 See artifacts and rendered test results at hud.pytorch.org/pr/145579

✅ No Failures

Uh oh!

ezyang commented Jan 24, 2025

Uh oh!

pytorchmergebot commented Jan 24, 2025

Merge started

Uh oh!

bdhirsh Jan 24, 2025

Choose a reason for hiding this comment

Uh oh!

ezyang Jan 24, 2025

Choose a reason for hiding this comment

Uh oh!

eellison Jan 29, 2025

Choose a reason for hiding this comment

Uh oh!

Reviewers

Assignees

Labels

Projects

Milestone

Development

Uh oh!

5 participants

ezyang commented Jan 24, 2025 •

edited by pytorch-bot bot

Loading

pytorch-bot bot commented Jan 24, 2025 •

edited

Loading