Pith. sign in

Paper Citation Record · LEDGER

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

As of 14 August 2026, this Paper Citation Record lists 42 of 42 outbound references and 1 inbound Pith citation observation for arXiv:2607.03657.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.03657 v1

Coverage vector

measured 42 of 42 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-12T00:53:05.419742Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

42 of 42 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved42
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 44f31bfa-a977-4f26-af64-c025b218ebd9 · outbound

This paper cites Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign Language Trans- lation (SLT) aims to bridge communication gaps by trans- lating sign-language videos into spoken sentences

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:0584dce13c23648303093f5449230fa13000dac15bf2a63acf793e7782f60caf

Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · outbound

This paper cites ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:b4c303b4c510cbb5c2321deb470bce942ce2d839df0e999c563404619d478192

Observation eeddb7e1-89c9-4e83-82c4-fe481fe0d951 · outbound

This paper cites SignLLM [5] maps sign videos to discrete tokens aligned with LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation SignLLM [5] maps sign videos to discrete tokens aligned with LLMs

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:cec481203aed8a8ef7a2b2dda0dff2af3ab60af1a4b36234615f9e4d8c5e8a5b

Observation 9d36b98f-90a7-47a3-bc64-605afec9d8ea · outbound

This paper cites Re- cent work leverageslarge-scale pretraining and multimodal LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Re- cent work leverageslarge-scale pretraining and multimodal LLMs

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:a3abdf5b12bd61c7c82e98d2747064c1a30462f87eb7ae17c023070dce597e41

Observation 2d32709d-b02d-45a6-b90b-42852d90b0fb · outbound

This paper cites Translate the given sentence into<language>.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Translate the given sentence into<language>

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:d9bc46e4597b61073612c00cc4b577566b39a4946a49ba5d7cd8ac594b20f45b

Observation df985ce5-4bd3-4222-9df0-a871dd45dd92 · outbound

This paper cites Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26].

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Experimental Setup Datasets.We evaluate on two benchmark SLT datasets: PHOENIX14T [25] and CSL-Daily [26]

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:4f89c4c52abeed416b86289ce269b3e8a15295be603de6f8aa94191adbaa1b35

Observation c81409e8-09fd-4e08-a012-c74491eedae3 · outbound

This paper cites an unresolved cited work.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work

Reference 7

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:26ef2fa8526bf40c0491694bc4b6fc668c06679cab57941fcf18822884a34725

Observation 642b8bcb-b2c0-4ebe-aef9-20b67b7de116 · outbound

This paper cites an unresolved cited work.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Unresolved cited work

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:110d1bfd1e4bfa8da200e6301df3da138c3cd057f011187816669a187c216dd0

Observation 45e7175b-e3e7-4d58-8488-ebfc66214358 · outbound

This paper cites Sign language transformers: Joint end-to-end sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign language transformers: Joint end-to-end sign language recognition and translation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:1d4d0c8016c58d9fee35c88344c35c1c6ebb317bdfe957266c7a1db12ed7766c

Observation 1266af36-d757-4300-b91f-8f7aa15ddb6a · outbound

This paper cites Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Factorized Learning Assisted with Large Language Model for Gloss-free Sign Language Translation

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:748924d39e57b412c0534b0083abf40b8853869cb2a2061dc198570f449de754

Observation 21f54591-24a9-4b3b-9817-13aa0b883718 · outbound

This paper cites Gloss-free sign language translation: Improving from visual-language pretraining,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss-free sign language translation: Improving from visual-language pretraining,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:dea9c3edb6bcf07c4d162f14fd410c2b01058402977d17b9e5360918aabd56ab

Observation abf0959e-9119-4624-ad2f-a1b7f504f03c · outbound

This paper cites Leveraging the power of mllms for gloss-free sign language transla- tion,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Leveraging the power of mllms for gloss-free sign language transla- tion,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:b174f6702d88f2dbd487b2e0e5f7e6f463318dbb5468e5af961955800a356636

Observation 53095427-6741-48d3-8723-c0745c10db7b · outbound

This paper cites Llms are good sign language translators,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Llms are good sign language translators,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:1efdd309d1cb86b0be1348459e4ffa69cab655dc1477b99bd60157d5598a6961

Observation dbfc33fa-6e84-47ad-9f2a-ec8cdfbf987c · outbound

This paper cites Lost in translation, found in context: Sign language translation with contextual cues,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in context: Sign language translation with contextual cues,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:08752cc215ed2de1c204ffbfebe87bca62a2ab7095ae7030933decf421ed3910

Observation f5e1c48a-cb80-4ae4-bc19-befe5bae6ba5 · outbound

This paper cites Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Sign2GPT: Leveraging Large Language Models for Gloss-Free Sign Language Translation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:7c04ccc003516b0cb580869d0fdc471951d23ceb89825de32bcd188f87d6cd31

Observation e9d58d38-1a59-4ecb-b248-bf0117adca00 · outbound

This paper cites Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Tspnet: Hierarchical feature learning via temporal semantic pyramid for sign language translation,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5de4a3c7a1a9f9b7f6e1c71887c7d1b585362f7fe1582695fa14ab821965741d

Observation 4b58aec4-71d2-41e9-aac2-f72c8c68e281 · outbound

This paper cites Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Conditional sen- tence generation and cross-modal reranking for sign lan- guage translation,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:68e47795713b825c5b295af5c5e6de161151d0c875466c6ba07ba72dbfe718c6

Observation 0b2a8b37-1673-4aeb-b5e4-3ae2e8ec5e11 · outbound

This paper cites A token-level contrastive framework for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation A token-level contrastive framework for sign language translation,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:d082878a610a93ace704963e38f11a33e9ef7d31155115237c6c221f56e3f0fe

Observation 0dc7d26f-fe5a-474f-a41b-5cd7620aeb47 · outbound

This paper cites Gloss attention for gloss-free sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Gloss attention for gloss-free sign language translation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:202ab82428507895c29643b2345c3e53b9941ee014fd9d1d0e2edd3e0c3d4337

Observation ca9a3f79-3091-4547-9b77-5bad231a13da · outbound

This paper cites Visual alignment pre- training for sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Visual alignment pre- training for sign language translation,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:4c21b3c3ee078753c0b2598b10401c4d5633ba31f73239a686a374e7ad9932b6

Observation 4b541d6c-79b4-4367-a7d8-495388c5ac6d · outbound

This paper cites An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation An Efficient Sign Language Translation Using Spatial Configuration and Motion Dynamics with LLMs

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:1968c83bc3da4390073c731d12ee4928ea9c9984dc03c25c739b72818ec1b88c

Observation 58457a83-56e3-413b-9060-9485926f7cc6 · outbound

This paper cites Better sign language translation with stmc-transformer,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Better sign language translation with stmc-transformer,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:78ea6d30b3eece770932dcf248c8114d83b04f853466d6a2d920faa5ea9f9f8d

Observation e2d25ca2-ac59-4de3-9a14-e6329f930d5d · outbound

This paper cites Lost in translation, found in embeddings: Sign language translation and alignment,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lost in translation, found in embeddings: Sign language translation and alignment,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9b2e571565221e015613a9734bc345a11af9c083900fc60506aa4394888bef8b

Observation 56c40903-7d60-4aff-8480-32c7eba70a65 · outbound

This paper cites Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multimodal sign language recognition via temporal deformable convolu- tional sequence learning.,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:07628fc82bbfe0f13d1607dbdb0031f9a868e8847e09545c1b806c898399c485

Observation 2468cb21-2394-467c-9cb3-81cdd806bb20 · outbound

This paper cites Multi-channel transformers for multi-articulatory sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multi-channel transformers for multi-articulatory sign language translation,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9a2dc06de1135385bc9cf7f1effb1867ac63cc6cf5992d9d3dca62348b8fd492

Observation 5d8d1f0f-2055-45ea-8387-e00c117fe462 · outbound

This paper cites Spatial- temporal multi-cue network for sign language recogni- tion and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Spatial- temporal multi-cue network for sign language recogni- tion and translation,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9f0430b7c19cf0825de2da4a96713d339ad11a44b7235f7aaa8b407b1bc96105

Observation 61c3a763-a5c2-4172-86aa-1f2ef5fe52c9 · outbound

This paper cites Graph- based multimodal sequential embedding for sign lan- guage translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Graph- based multimodal sequential embedding for sign lan- guage translation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:9258ae708b4627a486f886e17f8a2ae854e8ea03c03ec1012e121f07598fe533

Observation 565ab00d-b204-4af9-9c4b-e5fdf8c5744c · outbound

This paper cites Skeleton-aware neural sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Skeleton-aware neural sign language translation,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:90df697530f1a33bedd155cc3c9f0541a10f817635d8247e080c7d2160ee2a9f

Observation 5c2752ac-d1c2-4699-a948-3259f7b9a05f · outbound

This paper cites Two-stream net- work for sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Two-stream net- work for sign language recognition and translation,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:ae7eb0313f3d7970f55074d08bf5a68dd33d587b5fad29e8257ba7dd9f26e011

Observation 0bf15414-6b57-4033-9535-3d91fd91406d · outbound

This paper cites Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ustm: Unified spa- tial and temporal modeling for continuous sign language recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:2ad31846fc830c79a8787b4b9d98a59fdff19d17b2b0a288055e568649036a2f

Observation d51bd3aa-e5ff-4acf-8d1d-e14eebc5a5e9 · outbound

This paper cites Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Openpose: Re- altime multi-person 2d pose estimation using part affin- ity fields,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:26e88a271a012589bfaea9fd5bf59a340fa3e58fd034578eeda0f8206457d14f

Observation 005b2082-0c03-48e2-8127-b7701bc67a87 · outbound

This paper cites Temporal convolutional networks for action segmentation and de- tection,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Temporal convolutional networks for action segmentation and de- tection,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:533f90660aad4dc7aa2d28b5d19c4f9be57d7de77077a9673991745b20298cb4

Observation 4cc18bda-1be3-4653-8566-decea2ce0c1b · outbound

This paper cites Neural sign language translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Neural sign language translation,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:0f918ed63b7f131420c994b1d1129392a7935ee3304bb3a0783bb85fd326418d

Observation d932677a-5331-4ba5-a516-5f5aced9e2ed · outbound

This paper cites Improving sign language translation with monolingual data by sign back-translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Improving sign language translation with monolingual data by sign back-translation,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:d358a526ffe9ecc4bc51dfb0192f199421cda254fe6c7d824f625585a35778e0

Observation ab9a8202-c176-4fa8-9314-30088f71a9c9 · outbound

This paper cites Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Stochastic transformer networks with linear compet- ing units: Application to end-to-end sl translation,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:4e0f01f709b4a74f7b9229a8ac0e3c4608aceb3a86ef2348920fcb626d8f179e

Observation 4e631fb4-d64c-49c8-b802-af43b1aed53f · outbound

This paper cites Mska: Multi- stream keypoint attention network for sign language recognition and translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Mska: Multi- stream keypoint attention network for sign language recognition and translation,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:601fc7601e814dc3ed84181db711634d750b2a315bbfd5e3dc811b92596f73de

Observation 789c2d50-9bc0-45b0-925c-bdb1339e3292 · outbound

This paper cites Cross-modality Data Augmentation for End-to-End Sign Language Translation.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Cross-modality Data Augmentation for End-to-End Sign Language Translation

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:a65067c27adf4a167c244a9d9907c0ddfb87afdd045585b086997a22c9b0bbc3

Observation 490c6297-7a99-4be4-a976-c157a1ff7f45 · outbound

This paper cites Crosslingual generalization through multitask finetun- ing,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Crosslingual generalization through multitask finetun- ing,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:1eb498bee6cb8a8d4fa68416153933a664dfc50dbf534b68c4e220983f724a10

Observation c4114bb0-25e7-41a2-9793-96604dbc18e1 · outbound

This paper cites Lora: Low-rank adaptation of large language models,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Lora: Low-rank adaptation of large language models,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:64220218a755111ff2983197466ffd83e624f8b8380d7a788ef5823a31150ccf

Observation 9ae2e96d-8417-4cd5-b12c-19048fb1af02 · outbound

This paper cites Scaling instruction-finetuned language models,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Scaling instruction-finetuned language models,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:cc478b9d4886cc39edfed0aa1826a4d1c7c1ec59727b39d0e83be0753f8f22d8

Observation 82606eee-2de5-4d61-bb27-1329984ec715 · outbound

This paper cites Multilingual denois- ing pre-training for neural machine translation,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Multilingual denois- ing pre-training for neural machine translation,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:47f3646dd3a45eca3533425e68e98f6b36c4f63fcdbc3ffc84c4891238600238

Observation 9795a18a-e8a6-4886-b498-f3686aa20b6c · outbound

This paper cites Ararea- soner: Evaluating reasoning-based llms for arabic nlp,.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation Ararea- soner: Evaluating reasoning-based llms for arabic nlp,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:5d423922ef70b63f91c1e05a61706fb83c5f2974b13c2decb821b2a250c88fbd

Pith citing papers

Observation 904aa2ba-8e55-437c-8f3a-8f9c91a4b9e6 · inbound

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation cites this paper.

ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation ViPo-MLLM: Visual-Pose Multimodal LLM for Gloss-Free Sign Language Translation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-12T00:53:05.419742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T00:53:05.419742Z digest=sha256:b4c303b4c510cbb5c2321deb470bce942ce2d839df0e999c563404619d478192