Pith. sign in

Paper Citation Record · LEDGER

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision

As of 9 August 2026, this Paper Citation Record lists 15 of 15 outbound references and 1 inbound Pith citation observation for arXiv:2509.04658.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.04658 v1

Coverage vector

measured 15 of 15 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T06:00:14.837743Z

measured 16 of 16 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-20T12:52:15.138790Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-20T12:53:17.402749Z

Reference resolution

15 of 15 outbound references displayed

  • verified exact5
  • verified fuzzy4
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 90a39ecc-a9ff-4eca-bf9d-8a13bca83ec9 · outbound

This paper cites Surformer v1: Transformer-Based Surface Classification Using Tactile and Vision Features.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Surformer v1: Transformer-Based Surface Classification Using Tactile and Vision Features

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-08-05T06:00:14.885388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.786778Z digest=sha256:2899a08acad723cf1862f7792d0b23f66f9fcdd0280a4e205099f27ac548b4c1

Observation 62731c6a-0b80-4370-9c7d-8ee10f2325e5 · outbound

This paper cites Touch and Go: Learning from Human-Collected Vision and Touch.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Touch and Go: Learning from Human-Collected Vision and Touch

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:14.790891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:14.790891Z digest=sha256:67a03581bf4e64550ab8f9a7f4de726d3eb8e5b8245d0d91ec98c2782ecf9dfd

Observation 23471172-3861-4079-93b2-15449c796d56 · outbound

This paper cites Deep residual learning for image recognition,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Deep residual learning for image recognition,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:14.795366Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:14.795366Z digest=sha256:8aab4542bab609c8f15bce5ca2b40ae95801ed5f815dcb5499cec4dc9cab6ff1

Observation 259e1bb6-c6fd-43c6-85e1-774488e7b737 · outbound

This paper cites Convolutional Networks with Dense Connectivity.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Convolutional Networks with Dense Connectivity

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-05T06:00:15.088284Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.798802Z digest=sha256:9fadeeac4eab2ce6593c11a99671b690e84a72cdfa2d0be51d76840fc5ed6e8a

Observation 9ffcbcc7-4b59-448c-ab6f-e626d9426984 · outbound

This paper cites Gradient-based learning applied to document recognition,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Gradient-based learning applied to document recognition,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:15.204959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.802571Z digest=sha256:dfc1fc402a1a3ade60ac4285fa71c37235750c4aae91e3b6c7b08b0778fe789b

Observation 17824df3-b995-4898-8f4b-f6ca687f56a5 · outbound

This paper cites EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision EfficientNet: Rethinking Model Scaling for Convolutional Neural Networks

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:14.806676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:14.806676Z digest=sha256:1171f907c14d7112d9c452baad92fec21795a7aa7d709bb816e18163615fd7f9

Observation 2545e364-6b30-471d-b503-c2caa70e3cb5 · outbound

This paper cites Majority voting: Material classification by tactile sensing using surface textures,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Majority voting: Material classification by tactile sensing using surface textures,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:15.194631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.811600Z digest=sha256:2b25b73d7f02f4dccd15c04d91924b62a4698593ea78d9b646c0d206ea14d474

Observation 7da64e3e-0ffc-46e6-9de8-c7dd8b174f35 · outbound

This paper cites Tactile-data classification of contact materials using computational intelligence,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Tactile-data classification of contact materials using computational intelligence,

Reference 8

Resolution
verified exact
raw_fallback, observed 2026-08-05T06:00:15.060851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.814929Z digest=sha256:0c8c3ee69ed8f1db3dcbdb5ce2a2f9e8338c46e99214ca552052aa9d1d62ffdd

Observation d8bb8402-25db-4517-83e8-1e93994c2a9a · outbound

This paper cites Tactile-data classification of contact materials using principal component analysis and self-organizing maps,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Tactile-data classification of contact materials using principal component analysis and self-organizing maps,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:15.184948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.818036Z digest=sha256:8b4c65a8cb6db2a1e571705cbcc9201ed29407cc36b1ba7890eefbfb1136f19e

Observation be0a046b-c960-4a9e-84e7-4257e6b6fd36 · outbound

This paper cites Estimating perceptual attributes of haptic textures using visuo-tactile data,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Estimating perceptual attributes of haptic textures using visuo-tactile data,

Reference 10

Resolution
verified exact
raw_fallback, observed 2026-08-05T06:00:14.986302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.821187Z digest=sha256:5e79b20c0ee55cad93bd9a9bb8a8facbef7cd4e65b9b8f9423e5ce38a9cc1243

Observation a0efb47b-42af-4855-a341-452efbb2a5d9 · outbound

This paper cites Visuo-Tactile Transformers for Manipulation.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Visuo-Tactile Transformers for Manipulation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:14.824282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:14.824282Z digest=sha256:b9ff88f0cac7e68e37b722d38b60ce78f436d3af2d3390348de5008b8168afcd

Observation f9023a75-e8ef-4511-8c93-311b30a82b1f · outbound

This paper cites ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision ViTacFormer: Learning Cross-Modal Representation for Visuo-Tactile Dexterous Manipulation

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:14.827748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:14.827748Z digest=sha256:2ea5fe43e4ebeac0b38ffec7cfa927fc332716312fd33a29c384d709a3185018

Observation 41855a90-f11d-4be1-b5e8-e3d9058d4178 · outbound

This paper cites GelFusion: Enhancing Robotic Manipulation under Visual Constraints via Visuotactile Fusion.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision GelFusion: Enhancing Robotic Manipulation under Visual Constraints via Visuotactile Fusion

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T06:00:14.831282Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T06:00:14.831282Z digest=sha256:2478ef380894aff6ab676c87b582636ff5b7610debb116c0a48d0d89ebfaa138

Observation acb187f6-3e3f-4111-afd3-bc22ace99e04 · outbound

This paper cites Efficient visual-tactile transformer with token reorganization for robotic slip detection,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision Efficient visual-tactile transformer with token reorganization for robotic slip detection,

Reference 14

Resolution
verified exact
doi, observed 2026-08-05T06:00:14.869635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.834507Z digest=sha256:c783106b9fb6b3dd9ac3e909f80d879a6f7998666d1c078241914fbf3c817a1a

Observation e021d8c5-3097-402e-852c-5d1986145b58 · outbound

This paper cites EfficientNetV2: Smaller models and faster training,.

Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision EfficientNetV2: Smaller models and faster training,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T06:00:15.175128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T06:00:14.837743Z digest=sha256:a8f08a8f248d52a676a57b72dad795834d41cb79f98f428d7285161946cbc72f

Pith citing papers

Observation 0ace1f1b-26c1-499b-baa1-0b0bfe490ffc · inbound

Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms cites this paper.

Tactile-based Multimodal Fusion in Embodied Intelligence: A Survey of Vision, Language, and Contact-Driven Paradigms Surformer v2: A Multimodal Classifier for Surface Understanding from Touch and Vision

Reference 88

Resolution
verified exact
arxiv_id, observed 2026-05-20T12:53:17.404600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T12:52:15.138790Z digest=sha256:13d8d803926fb8c212c1188a679b960445e82abd0dee002fbe6a18032a82beee