Pith. sign in

Paper Citation Record · LEDGER

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge

As of 18 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 0 inbound Pith citation observations for arXiv:2505.06814.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.06814 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T22:35:59.141011Z

measured 32 of 32 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

32 of 32 outbound references displayed

  • verified exact1
  • verified fuzzy19
  • unresolved12
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1c57d5a9-210a-4431-afbf-c0a191e80ee0 · outbound

This paper cites Overview of the nlpcc 2023 shared task: Chinese medical instructional video question answering.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Overview of the nlpcc 2023 shared task: Chinese medical instructional video question answering

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.758298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.022150Z digest=sha256:36dcd4101d0259b9fd598a860c4c2d2fadf00975a70866c637cb7327b4a61ed5

Observation 758a6f30-b67a-4025-be8d-c552ed6d11e5 · outbound

This paper cites Overview of the nlpcc 2024 shared task 7: Multi-lingual medical instructional video question answering.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Overview of the nlpcc 2024 shared task 7: Multi-lingual medical instructional video question answering

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.747241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.026946Z digest=sha256:33e0091577a57aafe76c102d3fd4bab218f0be3897d94e6fd9ed5f2c16f9aafa

Observation 1ec040bd-3dfe-41fb-81fa-41ad91841e34 · outbound

This paper cites A systematic review of machine learning applications in infectious disease prediction, diagnosis, and outbreak fore- casting.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A systematic review of machine learning applications in infectious disease prediction, diagnosis, and outbreak fore- casting

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.735619Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.030812Z digest=sha256:b8d2e12cabbbf09efff88b8eab1fe39acbe3743398787c19ce0f7aef3233b82b

Observation 0a546ac8-8eee-48a1-8fff-9b5cd3417b1b · outbound

This paper cites Tf-icon: Diffusion-based training-free cross-domain image composition.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Tf-icon: Diffusion-based training-free cross-domain image composition

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.724868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.034940Z digest=sha256:6db4679fe06a56a11fa4d4872c61a2465479590099bed3499ce5d254ef903da4

Observation 95db3e03-0c2f-4493-8de1-2aab886ad0c7 · outbound

This paper cites Ich-prnet: a cross-modal intracerebral haemorrhage prognostic prediction method using joint-attention in- teraction mechanism.Neural Networks, 184:107096, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Ich-prnet: a cross-modal intracerebral haemorrhage prognostic prediction method using joint-attention in- teraction mechanism.Neural Networks, 184:107096, 2025

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.713670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.038753Z digest=sha256:45bcf6093700894566d774b8559b0c38feb48798534abeef30444173048d3113

Observation 966b5fce-d1e8-48f7-94c8-36a309855057 · outbound

This paper cites Ich-scnet: Intracerebral hemorrhage segmentation and prognosis classification network using clip-guided sam mecha- nism.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Ich-scnet: Intracerebral hemorrhage segmentation and prognosis classification network using clip-guided sam mecha- nism

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.699572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.042707Z digest=sha256:103e7af3facd691aa62c1994edbafb012e4c0a1d51fda423e585547cb77526df

Observation 529b38af-485b-449e-8eb8-df14adbd592d · outbound

This paper cites Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Towards nation-wide analytical healthcare infrastructures: A privacy-preserving augmented knee rehabilitation case study

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.046732Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.046732Z digest=sha256:50b76a134484b3dd2638dc29a8fa569a03dd6e2e2f8ef9be3f6521b91f3036fc

Observation 7479754d-4183-4d75-8184-07cae77cb536 · outbound

This paper cites Learning to Unify Audio, Visual and Text for Audio-Enhanced Multilingual Visual Answer Localization.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Learning to Unify Audio, Visual and Text for Audio-Enhanced Multilingual Visual Answer Localization

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.050670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.050670Z digest=sha256:2fbc0a613f87e3249d9b085980b695e3386d349207c395c259032c9895567e55

Observation 1391dcac-a8f6-49ae-8eea-57d2cdf5a872 · outbound

This paper cites Prism: Self-pruning intrinsic selection method for training-free multi- modal data selection, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Prism: Self-pruning intrinsic selection method for training-free multi- modal data selection, 2025

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.688250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.054396Z digest=sha256:4cf3186de56aae7146fcd2ed07169e05d19038de1d603f6c8363de0efbeef8aa

Observation e3abe435-08eb-4872-b31f-5df9a368e361 · outbound

This paper cites Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Robust Watermarking Using Generative Priors Against Image Editing: From Benchmarking to Advances

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.057796Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.057796Z digest=sha256:07a8a90e4621f3aff6913e5216d2ed38b65bff5f14704994862b2dc952b1d6df

Observation d6dd0299-1b94-4dc0-8017-ff7b32d66fd3 · outbound

This paper cites Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Set You Straight: Auto-Steering Denoising Trajectories to Sidestep Unwanted Concepts

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.061770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.061770Z digest=sha256:0d9e385b6a833104805571fb725136c50836e897b77ac01bd8690eb93c8795f2

Observation fd829d8b-1740-4d6b-abbe-bace547192c5 · outbound

This paper cites LingYi: Medical Conversational Question Answering System based on Multi-modal Knowledge Graphs.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge LingYi: Medical Conversational Question Answering System based on Multi-modal Knowledge Graphs

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-08-15T22:35:59.447503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.066362Z digest=sha256:79706f979587d1b7b8cfd656699ca512fa055f75f44e4e112d664c000f840d36

Observation dfd89b6e-3d66-49a3-8568-501979759302 · outbound

This paper cites Llava steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Llava steering: Visual instruction tuning with 500x fewer parameters through modality linear representation-steering, 2025

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.676218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.070339Z digest=sha256:cc641fad4fc08134df8353cbc300fdc71190f88cc7ba6a72285111ac27ae01c5

Observation de4d2db6-26ba-49f5-a727-64c64760e57b · outbound

This paper cites Visual answer localization with cross-modal mutual knowledge transfer.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Visual answer localization with cross-modal mutual knowledge transfer

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.663265Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.073882Z digest=sha256:979457f152025ecc7c51f17f23c80d28dd9501d96bb1998db47c98a2a994a9d0

Observation 76ecb771-b6a3-4096-9fe8-2b86cdc612c4 · outbound

This paper cites Enhancing thyroid disease prediction using ma- chine learning: A comparative study of ensemble models and class balancing tech- niques.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Enhancing thyroid disease prediction using ma- chine learning: A comparative study of ensemble models and class balancing tech- niques

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.650841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.077678Z digest=sha256:d850a5d243f3c444a9b994f12dc8bc1f4979e737ed9efada9bf9c40332ef0f2c

Observation e16aaf9b-4a09-4897-a875-e1fcc5435eb5 · outbound

This paper cites A generative adversarial network-based investor sentiment in- dicator: Superior predictability for the stock market.Mathematics, 13(9):1476, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A generative adversarial network-based investor sentiment in- dicator: Superior predictability for the stock market.Mathematics, 13(9):1476, 2025

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.638327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.081388Z digest=sha256:ce3b62806ad63efdaf9c5f10b4cef2842c5a010ae82f6100710093aea168d262

Observation c8055437-5f2e-4fee-9844-4432911f1809 · outbound

This paper cites Multimodal high- order relationship inference network for fashion compatibility modeling in internet of multimedia things.IEEE Internet of Things Journal, 11(1):353–365, 2024.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Multimodal high- order relationship inference network for fashion compatibility modeling in internet of multimedia things.IEEE Internet of Things Journal, 11(1):353–365, 2024

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.626294Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.085164Z digest=sha256:ae5d24a7e91e6e958b8630f2dc3b1ea7bcfad491c6ead1ad4e4c85efe028b3cc

Observation c811c3bf-269a-451c-9600-d48be8e12205 · outbound

This paper cites De- tection of ai deepfake and fraud in online payments using gan-based models.arXiv preprint arXiv:2501.07033, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge De- tection of ai deepfake and fraud in online payments using gan-based models.arXiv preprint arXiv:2501.07033, 2025

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.088942Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.088942Z digest=sha256:b7e888f7b22eca8d4264a1ef30b2f04f24c0f98ed9f2872e02deddf705da91b2

Observation fa1f9d64-991a-4edc-a3cc-ec90b10a80b0 · outbound

This paper cites Enhancing intent understanding for ambiguous prompts through human-machine co-adaptation.arXiv preprint arXiv:2501.15167, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Enhancing intent understanding for ambiguous prompts through human-machine co-adaptation.arXiv preprint arXiv:2501.15167, 2025

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.092542Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.092542Z digest=sha256:a31dfaff5ddd08cc52f961c0f2ea8d044e7c0b57290157af91b678a5a42b8bd8

Observation e0228f52-1048-4cd5-8f3b-f16524a5ec4e · outbound

This paper cites Mace: Mass concept erasure in diffusion models.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Mace: Mass concept erasure in diffusion models

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.614545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.096530Z digest=sha256:5326a2fc3fa23d9afe6c1d14501eb95d020c4d102d598c0a9dfdb21849222bcf

Observation bebc369c-b979-4d3a-ba0e-d7bd624b5273 · outbound

This paper cites Score: Story coherence and retrieval enhancement for ai narratives.arXiv preprint arXiv:2503.23512, 2025.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Score: Story coherence and retrieval enhancement for ai narratives.arXiv preprint arXiv:2503.23512, 2025

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.100140Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.100140Z digest=sha256:3a6e40b85b31d8d1d39a5b5da8185fc862305b3436073b27ca02ef55e295fc33

Observation fc1cfef9-cc4d-4a33-8afd-ea729d6d5e98 · outbound

This paper cites Learning to locate visual answer in video corpus using question.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Learning to locate visual answer in video corpus using question

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.602358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.103699Z digest=sha256:168f63d0a576d9e516d31e1d7a2eb8e7df46f0652408e4985e90e501ec29eb3c

Observation 7293766e-9a1c-4814-8131-12dc5b30f0ce · outbound

This paper cites Improving multilingual temporal answering grounding in single video via llm-based translation and ocr enhancement.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Improving multilingual temporal answering grounding in single video via llm-based translation and ocr enhancement

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.590814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.107376Z digest=sha256:c1882676f88ef7ec9b29582664a65d89588153c8308fb5313bbd466a16323343

Observation 43e49dc4-b3b9-47b6-b4d7-ab929f706fd0 · outbound

This paper cites Improv- ing cross-modal visual answer localization in chinese medical instructional video using language prompts.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Improv- ing cross-modal visual answer localization in chinese medical instructional video using language prompts

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.577486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.111232Z digest=sha256:a24f51b3d4df85d7b0fbaa8fd8afb8b4899c66a0a74a4d2cefb51cf3da771a75

Observation c0fa08cd-de37-4a06-878c-37d0b2f5ba7f · outbound

This paper cites Mqua: Multi-level query-video augmentation for multilingual video corpus retrieval.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Mqua: Multi-level query-video augmentation for multilingual video corpus retrieval

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.114991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.114991Z digest=sha256:5470b735810378e221f15d2fd16b64ce1dc0ce720a2dede10b39286406f33854

Observation 883cf69e-05ae-46b5-87a7-4fbd6afe1f1a · outbound

This paper cites A two-stage chinese medical video retrieval framework with llm.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A two-stage chinese medical video retrieval framework with llm

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.118451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.118451Z digest=sha256:87d39ed7c1456c4dfb94af9b21fb6a758e0ba7ccb904a6f8a69ce00383c4a439

Observation e9606ac6-6f53-4326-8a77-c80581d56911 · outbound

This paper cites Mul- tilingual temporal answer grounding in video corpus with enhanced visual-textual integration.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Mul- tilingual temporal answer grounding in video corpus with enhanced visual-textual integration

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.121971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.121971Z digest=sha256:9d94e7cfec2a6e92f842060f9093acba722942b9f7fc88f41973fe953a372ee8

Observation bb847df7-720d-48d3-a765-bc12eac5f669 · outbound

This paper cites A uni- fied framework for optimizing video corpus retrieval and temporal answer ground- ing: fine-grained modality alignment and local-global optimization.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge A uni- fied framework for optimizing video corpus retrieval and temporal answer ground- ing: fine-grained modality alignment and local-global optimization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.125926Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.125926Z digest=sha256:113052abe3c4bf7382f73b75757895e353e437b2d7317020ea4c3d08e64cc7cb

Observation accf6a10-c5ee-4d6a-83d1-11f3fb446d26 · outbound

This paper cites Correlation-aware cross-modal attention net- work for fashion compatibility modeling in ugc systems.ACM Transactions on Multimedia Computing, Communications and Applications, 2024.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Correlation-aware cross-modal attention net- work for fashion compatibility modeling in ugc systems.ACM Transactions on Multimedia Computing, Communications and Applications, 2024

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.535640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.129694Z digest=sha256:e67ecbd69a620c84c9ec2a1dd64566c5004b146ce17f170eb3b06f28a7b0c36e

Observation d20edbe7-de37-424a-9a99-ef5de3e73c9e · outbound

This paper cites Language-agnostic BERT Sentence Embedding.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Language-agnostic BERT Sentence Embedding

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-15T22:35:59.133119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:35:59.133119Z digest=sha256:77fbe63f869efb357fc3fd1ac65ae16b9d17f6f27da1e4a4aa1698d007849179

Observation fcea76bc-44ba-48eb-8632-e924b1305c27 · outbound

This paper cites Vpai lab at medvidqa 2022: a two-stage cross-modal fusion method for medical instructional video classi- fication.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Vpai lab at medvidqa 2022: a two-stage cross-modal fusion method for medical instructional video classi- fication

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.524131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.137319Z digest=sha256:25eb6254b6a458524e18d7793f479be59892f5772553c541854f572554d3636f

Observation f198ac09-494c-44a9-926b-43bef7043ebf · outbound

This paper cites Category-aware multimodal attention network for fashion compatibility modeling.IEEE Transac- tions on Multimedia, 25:9120–9131, 2023.

Overview of the NLPCC 2025 Shared Task 4: Multi-modal, Multilingual, and Multi-hop Medical Instructional Video Question Answering Challenge Category-aware multimodal attention network for fashion compatibility modeling.IEEE Transac- tions on Multimedia, 25:9120–9131, 2023

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T22:35:59.512412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-15T22:35:59.141011Z digest=sha256:15df1d21e13be78c513fbbe51cd29eb1fd799e07dd53c156cddf0cf7b041c33c

Pith citing papers

No inbound Pith citation observations are available.