Pith. sign in

Paper Citation Record · LEDGER

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System

As of 17 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 0 inbound Pith citation observations for arXiv:2508.11885.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.11885 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T17:29:34.062251Z

measured 39 of 39 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy27
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a1342987-a4d9-4c15-afab-6330f34a3b28 · outbound

This paper cites Flamingo: a visual language model for few-shot learning.Advances in Neural Informa- tion Processing Systems, 35:23716–23736, 2022.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Flamingo: a visual language model for few-shot learning.Advances in Neural Informa- tion Processing Systems, 35:23716–23736, 2022

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.516943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.798435Z digest=sha256:474e2bfa1711b48715ca7076517c4e0f0246ea5b8a1af9814f014241830f0f90

Observation c669e027-0e8e-4318-a411-87b7b1f8ecef · outbound

This paper cites Deep Variational Information Bottleneck.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Deep Variational Information Bottleneck

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.804065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.804065Z digest=sha256:d4a574563768b2eb72a0e2b0367a0488bf7e6874363f622a605763c375adcf10

Observation 417d974b-8cff-4c08-86f8-c18136e04d28 · outbound

This paper cites Divprune: Diversity-based visual token pruning for large multimodal models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Divprune: Diversity-based visual token pruning for large multimodal models

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.501792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.809188Z digest=sha256:f9842d4dcdf6f7debf81a3554e30cd1af371d15ee7ecdf5b189309140cd20487

Observation fa8cd2cc-21ab-4ef7-8593-b4369c6a903b · outbound

This paper cites A closer look at referring expressions for video object segmentation.Multimedia Tools and Applications, 82 (3):4419–4438, 2023.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System A closer look at referring expressions for video object segmentation.Multimedia Tools and Applications, 82 (3):4419–4438, 2023

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.490357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.813249Z digest=sha256:ea799879285370914dfbb8c9fe100ed4dc1fe8f6001af97474e458f1a49c8618

Observation c8a16dc0-b6ce-4fb1-8c34-ec37cdc63fd4 · outbound

This paper cites Token Merging: Your ViT But Faster.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Token Merging: Your ViT But Faster

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.817206Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.817206Z digest=sha256:d91ed23874f6af9716ae5c7773d6d2760223568d439f516d3f64b58eca6eda9f

Observation 89887543-2312-40b9-9386-9a6c66ff89e3 · outbound

This paper cites Matryoshka multimodal models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Matryoshka multimodal models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.479102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.932998Z digest=sha256:4bc5ea461aa3561992a82129a99e1ea8d11457c5e3c3f285acad19fa92e852b5

Observation 3194bba4-5780-41b9-a2da-efe947697eff · outbound

This paper cites An im- age is worth 1/2 tokens after layer 2: Plug-and-play in- ference acceleration for large vision-language models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System An im- age is worth 1/2 tokens after layer 2: Plug-and-play in- ference acceleration for large vision-language models

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.469859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.936978Z digest=sha256:802a6df2e9ced53b47f295170cdf3ebb80d7f247c22ecb79e06bd50648a74af9

Observation 7f07d4f5-a8f3-4c80-82f1-b7de5527d8c0 · outbound

This paper cites Sequence complementor: Complementing transformers for time series forecasting with learnable sequences.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Sequence complementor: Complementing transformers for time series forecasting with learnable sequences

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.459928Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.940928Z digest=sha256:6b82f98cf9713923e88da577561806cd5bd00514c9d83ec37d042a74d7b6a80d

Observation 128e75b4-12f7-4efb-8fc4-5e939be8866f · outbound

This paper cites Masked- attention mask transformer for universal image seg- mentation.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Masked- attention mask transformer for universal image seg- mentation

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.449726Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.944480Z digest=sha256:4cbf922f702746520d055eb7fde69cc23b38aa53fb70ff4d6dae59c9c517630d

Observation 86784cca-0919-4f77-b1d6-d1a3a87d05d1 · outbound

This paper cites Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Xmem: Long-term video object segmentation with an atkinson-shiffrin memory model

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.438582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.948263Z digest=sha256:df75e41277b5b99f713304e410da59b920fee6745678c543ac3d1d041131e50a

Observation 819eefe8-fdf4-423e-bacb-9d06bef93de4 · outbound

This paper cites John Wiley & Sons, 2nd edition,.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System John Wiley & Sons, 2nd edition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.427386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.951951Z digest=sha256:643a9927fb201b60681f33a12c490119bc652176f728f2e05715cafa109b36d0

Observation 45307e83-cfdc-4395-80f4-ed9d8a2d5e5d · outbound

This paper cites Phi-2: The surprising power of small language models.Microsoft Research Blog, 1:3, 2023.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Phi-2: The surprising power of small language models.Microsoft Research Blog, 1:3, 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.416035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.955972Z digest=sha256:5c8e5dbc7be579c40df1a011f5b7ebd57b37ab4747fe43d6399ebed2dec33211

Observation 70165da1-020b-4f72-8c26-eee384ecae7e · outbound

This paper cites Segment anything.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Segment anything

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.404631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.959773Z digest=sha256:ce8302d1d67d0b3a7744346e285a1359120541a6949b688648813de5c30e64aa

Observation d6dd719d-9f45-4fdc-b39d-8b796b6657d7 · outbound

This paper cites LISA: Reasoning Segmentation via Large Language Model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System LISA: Reasoning Segmentation via Large Language Model

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.963290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.963290Z digest=sha256:9729d8911d6e720eab258d311855be0aabf29cc551d8c8fd9781841d5cffc589

Observation b5584475-201e-4788-ae9e-e6b8bd5bb732 · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.967349Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.967349Z digest=sha256:21f0c26ec62cca013c7ce2afaab3e046693a0050fa6bc6159e87b7d5b6cf6986

Observation 69159d6c-a40a-4b04-8776-ad1578e16ab0 · outbound

This paper cites Referring transformer: A one-step approach to multi-task visual grounding.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Referring transformer: A one-step approach to multi-task visual grounding

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.392940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.971599Z digest=sha256:8b258b57493d7841308143081a57e69a3610cfaadfbadebba9e452b395601f18

Observation e61207bf-1808-4bdb-a967-d6947c340696 · outbound

This paper cites To- kenpacker: Efficient visual projector for multimodal llm.International Journal of Computer Vision, pages 1–19, 2025.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System To- kenpacker: Efficient visual projector for multimodal llm.International Journal of Computer Vision, pages 1–19, 2025

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.381446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.975868Z digest=sha256:590f54c208225bd627eda4342e277012b992f7d5cffa9fca88a358a570e22dde

Observation e4bceed2-a8c3-467f-b214-0d01ecdc27b9 · outbound

This paper cites Boosting multimodal large language models with visual tokens withdrawal for rapid inference.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Boosting multimodal large language models with visual tokens withdrawal for rapid inference

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.370933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.979318Z digest=sha256:1d542c7e60b49afee406ff780eaf50b8de9a84e83ae54a85074763c9ebbbf00a

Observation a72d34e1-9562-447f-9e85-6886b6f6c7c3 · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Swin transformer: Hierarchical vision transformer using shifted windows

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.360956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.982412Z digest=sha256:046d11c1bedf0dc65656130c8a6132d2ec372c2280e9c488b864cf184a0e2e97

Observation cb9f644e-fb0f-448d-afd3-3fa6d21b2bde · outbound

This paper cites Modeling context between objects for referring ex- pression understanding.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Modeling context between objects for referring ex- pression understanding

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.350384Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.985618Z digest=sha256:d34c2cd626d06dfe30c3653fe9e3be8d08dce2eae4502044b8054dc5db064597

Observation 18f123f2-0628-49c6-89d8-8e1ca92ee7d6 · outbound

This paper cites PerceptionGPT: Effectively Fusing Visual Perception into LLM.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System PerceptionGPT: Effectively Fusing Visual Perception into LLM

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.989125Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.989125Z digest=sha256:3962ee70c59c5f62e32db791d98ba7347940c8ea0c0842efc450845ce0a86c2f

Observation c34e7079-30e6-4004-92c1-01f0c58eccf7 · outbound

This paper cites PixelLM: Pixel Reasoning with Large Multimodal Model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System PixelLM: Pixel Reasoning with Large Multimodal Model

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:33.993033Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:33.993033Z digest=sha256:f24a16d2f793d0d0d744ddb3103ef7677d12c0c74a28388fc0033d8246b58605

Observation d9d77504-c35e-4ef2-ab93-bd11a2850368 · outbound

This paper cites Urvos: Unified referring video object segmentation network with a large-scale benchmark.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Urvos: Unified referring video object segmentation network with a large-scale benchmark

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.340883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:33.996873Z digest=sha256:5b666b69231b84bb116bcf15ef8754965fd1504d3341d7fc2a64c0903d8ae15d

Observation f4b03301-9631-425c-b52a-3bdea729dcae · outbound

This paper cites Llava-prumerge: Adaptive token re- duction for efficient large multimodal models.ICCV,.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Llava-prumerge: Adaptive token re- duction for efficient large multimodal models.ICCV,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.330531Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.000423Z digest=sha256:88ebc8f4f44951efa637f0de8375469b2750dbb290c20a54e45a2485d6d7dc12

Observation a5bc1ebb-aaef-4bb8-a004-3b6ea746b88e · outbound

This paper cites Contrastive grouping with transformer for referring image segmentation.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Contrastive grouping with transformer for referring image segmentation

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.319946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.004669Z digest=sha256:8848662d5b5a171f2e7f48b90696669a6817e68bb38f37d7f97e5b3315490b62

Observation 07155c16-30d9-4768-9404-dcc028f130ad · outbound

This paper cites Deep learning and the information bottleneck principle.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Deep learning and the information bottleneck principle

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.307134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.008176Z digest=sha256:9060f7c00e13681008b4c0f8377407584d07510362a8e459b2ac6b158124a5dd

Observation bde20625-28c0-429e-9e9e-f5bd4619eb66 · outbound

This paper cites Cris: Clip-driven referring image segmentation.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Cris: Clip-driven referring image segmentation

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.296740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.011745Z digest=sha256:3c65e7e853368ffca165754757288cd545eac401bceaf64e9a18b324b86460db

Observation 95dff278-d6b0-4fd9-a107-a561770720b9 · outbound

This paper cites LaSagnA: Language-based Segmentation Assistant for Complex Queries.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System LaSagnA: Language-based Segmentation Assistant for Complex Queries

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.015425Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.015425Z digest=sha256:b2c28aa5c3afc874c2e1e461e46c317aa3a6882b20dc5b826edb29f9502fd972

Observation a7f7e31d-8a5e-444a-9e1c-af8ca44499ff · outbound

This paper cites Instructseg: Unifying instructed visual segmentation with multi- modal large language models.ICCV 2025, 2025.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Instructseg: Unifying instructed visual segmentation with multi- modal large language models.ICCV 2025, 2025

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.284853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.020519Z digest=sha256:bb6ce164eccd32d320d9b712d333712f3f0a702785d0ef430b59f2b12e76a78f

Observation 337b2236-973c-46ee-ad98-70736df1d5e5 · outbound

This paper cites GSVA: Generalized Segmentation via Multimodal Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System GSVA: Generalized Segmentation via Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.025187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.025187Z digest=sha256:c74e90c95e0b8a0adb2b2e9fd2d8ff04e275fb91cf806dd0d221ad7e028a301e

Observation 555d31c1-f140-4e16-a401-0627bf54e0cf · outbound

This paper cites VISA: Reasoning Video Object Segmentation via Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System VISA: Reasoning Video Object Segmentation via Large Language Models

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.029272Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.029272Z digest=sha256:f5ea0cd4eafc01bc63ad23c9bfe0ad0da9ddc8b4a977b06e682cd879d76adde2

Observation 4dac4429-0caa-4fba-a3ba-531d83d4e636 · outbound

This paper cites Visionzip: Longer is better but not necessary in vision language models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Visionzip: Longer is better but not necessary in vision language models

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.272300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.033613Z digest=sha256:9ba5b5a123e76dbdc5426ef53a816a8bbb5d37ea31b8ee855e2675bbc11af230

Observation 8a94a600-f29b-404d-8fb4-d65a86e297fe · outbound

This paper cites Fit and prune: Fast and training-free visual token pruning for multi-modal large language models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Fit and prune: Fast and training-free visual token pruning for multi-modal large language models

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.260031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.037936Z digest=sha256:b2602038e93b190b0909ef27e76ba9108fa002168fd0df1a1053ec0b77262b73

Observation 18f9d2a3-d968-471b-8379-df5177abe9e9 · outbound

This paper cites Modeling context in re- ferring expressions.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Modeling context in re- ferring expressions

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.247202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.041698Z digest=sha256:8f6eacfe6c47cf5475ccbf4497a430802aee5dc5c2e69e8f26a1de9a4a0bfd40

Observation 5382ee08-f445-477c-829b-c1a83b00a8c4 · outbound

This paper cites Sigmoid loss for language im- age pre-training.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Sigmoid loss for language im- age pre-training

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.234978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.045442Z digest=sha256:0758cad43fd3df367eb7a02c452ac2bb22f47f1b5ad3d66d9bb5a71194fdc767

Observation b844293a-ca70-4fa6-93ae-5f2b30ab76a5 · outbound

This paper cites Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System Beyond Text-Visual Attention: Exploiting Visual Cues for Effective Token Pruning in VLMs

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.049030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.049030Z digest=sha256:9c94bbb33871355af2b0c91d557d4f9ef21fde30dc40a396025fd57a5a0dbdfe

Observation cb30834c-3eac-4f51-8643-68ed8c81b4de · outbound

This paper cites PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System PSALM: Pixelwise SegmentAtion with Large Multi-Modal Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.053765Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.053765Z digest=sha256:b95b07d32f1c208f2acfc30cd698a04b6c613a9d53835aa8ab15ce9f7d11d4f9

Observation 29e03bcb-c8e3-49f7-96ae-7cdd54b1217f · outbound

This paper cites MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-15T17:29:34.058509Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:29:34.058509Z digest=sha256:a1a0df61019c4c8077048383416f1c94f420efe8b1d71f724fb16ad0a8cfca85

Observation 8f09e585-558a-4f4e-b564-6776761766f7 · outbound

This paper cites dog with its mouth open,.

Contact-Rich and Deformable Foot Modeling for Locomotion Control of the Human Musculoskeletal System dog with its mouth open,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T17:29:34.220714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-15T17:29:34.062251Z digest=sha256:d1cdcbb6b98a7a6f10310ca8348b6a878465dcf7dfbb074303163e5381584dfd

Pith citing papers

No inbound Pith citation observations are available.