Skip to content
Research · Jul 25, 2026

Apple proposes RayRoPE for geometry-aware positional encoding in multi-view transformers

RayRoPE encodes patch positions via rays and predicted 3D points to enable SE(3)-invariant attention and geometry-adaptive multi-frequency similarity, improving novel-view synthesis and stereo depth estimation.

Trust79
HypeLow hype

1 source · cross-referenced

ShareXLinkedInEmail
TL;DR
  • Apple’s ML Research introduces RayRoPE, a positional encoding method for multi-view transformers that uses rays and predicted 3D points to encode patch positions.

Apple’s Machine Learning Research team proposed RayRoPE, a positional encoding method for multi-view transformers that process tokens from posed input images. The approach encodes patch positions based on associated rays but uses a predicted point along the ray rather than the ray direction to achieve geometry-aware encoding.

To ensure SE(3)-invariant attention, RayRoPE computes query-frame projective coordinates for multi-frequency similarity. Because the predicted 3D point along a ray may not be precise, the method includes a mechanism to analytically compute the expected position encoding under uncertainty.

The team validated RayRoPE on novel-view synthesis and stereo depth estimation, reporting consistent improvements over alternative positional encoding schemes. On the CO3D dataset, RayRoPE achieved a 15% relative improvement on LPIPS compared to alternatives.

RayRoPE also supports seamless integration of RGB-D input, which further increases performance gains over methods that cannot positionally encode this information.

Sources
  1. 01Apple — Machine Learning ResearchRayRoPE: Projective Ray Positional Encoding for Multi-View Attention
Also on Research

Stories may contain errors. Dispatch is assembled with AI assistance and curated by human editors; despite the trust-score filter, mistakes happen. We correct publicly — every article links to its revision history. Nothing here is financial, legal, or medical advice. Verify before relying on any claim.

© 2026 Dispatch. No ads. No sponsorships. No paid placement. Reader-supported via Ko-fi.

Built by a person who cares about honest AI news.