Rendering 3D Gaussians on a Graph Processor
arXiv:2607.15951v1 Announce Type: cross Abstract: We present the first implementation of a 3D Gaussian renderer on an Intelligence Processing Unit (IPU), comprising 1,472 independent tiles with…
arXiv:2607.15951v1 Announce Type: cross Abstract: We present the first implementation of a 3D Gaussian renderer on an Intelligence Processing Unit (IPU), comprising 1,472 independent tiles with…
arXiv:2607.15421v1 Announce Type: cross Abstract: Compact medical-image classifiers need efficiency and interpretable evidence, yet these goals are often addressed separately. We introduce qZACH-ViT, a quantization-aware extension…
arXiv:2607.15512v1 Announce Type: new Abstract: Attacks on general computer vision algorithms are often relegated to the digital domain, with the optimization performed purely in the digital…
arXiv:2607.16146v1 Announce Type: cross Abstract: Vision and touch are complementary modalities essential for robotic perception and manipulation. While vision provides global object context, touch offers precise…
arXiv:2510.05509v3 Announce Type: replace Abstract: Diffusion models are powerful deep generative models, but unlike classical models, they lack an explicit low-dimensional latent space that parameterizes the…
arXiv:2607.15410v1 Announce Type: new Abstract: Our study provides evidence that CNNs struggle to extract orientation features effectively. We show that using the Complex Structure Tensor, which…
arXiv:2607.15517v1 Announce Type: new Abstract: Four-finger SLAP fingerprints are flat live-scan impressions of the index, middle, ring, and little fingers of one hand, used for identity…
arXiv:2607.15065v1 Announce Type: cross Abstract: Predictive world models enable robots to plan by imagining the outcomes of their actions, but their value for control hinges on…
arXiv:2607.14334v1 Announce Type: new Abstract: Learned image compression (LIC) is bottlenecked by the need to store independent models for each rate-distortion operating point. Existing variable bit-rate…
arXiv:2508.18242v2 Announce Type: replace Abstract: We introduce GSVisLoc, a visual localization method designed for 3D Gaussian Splatting (3DGS) scene representations. Given a 3DGS model of a…