Pith. sign in

Paper Citation Record · LEDGER

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning

As of 21 August 2026, this Paper Citation Record lists 100 of 105 outbound references and 0 inbound Pith citation observations for arXiv:2507.13364.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.13364 v1

Coverage vector

measured 100 of 105 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T19:50:24.303188Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 105 outbound references displayed

  • verified exact8
  • verified fuzzy35
  • unresolved57
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d59a1872-489d-4553-9de4-41191b6890f5 · outbound

This paper cites Vatt: Transformers for multimodal self-supervised learning from raw video, audio and text.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Vatt: Transformers for multimodal self-supervised learning from raw video, audio and text

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.223191Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.223191Z digest=sha256:ddd75114a99523fe59a3de2a63f3b7a69cfa5b7496bf8bcecb6fbe21c5d4d4cd

Observation 58c6bae3-fdc2-4237-8b13-d67425b5116f · outbound

This paper cites Objects that sound.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Objects that sound

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.301171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.301171Z digest=sha256:8e845f31076a1bd42b7f08e65a7944b8b5f6f8dd9cf1264b611529e9e2e629ad

Observation af1c03c2-eed2-4c63-8061-3e9de9304f1e · outbound

This paper cites 3d seman- tic parsing of large-scale indoor spaces.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning 3d seman- tic parsing of large-scale indoor spaces

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.441879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.441879Z digest=sha256:d3f89ee8dd12eca1fbd1d183a71d5382b62786a3ec4ee3e4b9cdf367e65209d2

Observation 39d3b6b5-846b-49fd-87e5-30dc0c0f116b · outbound

This paper cites MAE-AST: Masked Autoencoding Audio Spectrogram Transformer.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning MAE-AST: Masked Autoencoding Audio Spectrogram Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.572549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.572549Z digest=sha256:bc99456ab1d62570b06515b4d3f7d40a16bbe1e85f3971cda36f7a4977e95c1d

Observation a21ca1f0-fa77-4a38-b1f9-6cd59471f3d0 · outbound

This paper cites Data2vec: A general frame- work for self-supervised learning in speech, vision and lan- guage.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Data2vec: A general frame- work for self-supervised learning in speech, vision and lan- guage

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.702418Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.702418Z digest=sha256:369d8b625af20ebbc36aedd1cfa2f263e11282bcd2e74bdf5289a4fbf626fd60

Observation 2bea425a-ac9b-44e4-8ad7-d70b29a3bf8d · outbound

This paper cites Generative adversarial networks based on transformer encoder and convolution block for hyperspectral image classification.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Generative adversarial networks based on transformer encoder and convolution block for hyperspectral image classification

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.854860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.854860Z digest=sha256:4eae4d3d6473e41e561dc12b3af3644de2ca7fdc741d9b65ef1c451fe17917aa

Observation 3c461085-a472-455b-901b-a4c5fd5b9a5b · outbound

This paper cites HiP: Hierarchical Perceiver.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning HiP: Hierarchical Perceiver

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:14.988580Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:14.988580Z digest=sha256:2917d711409a0b4de3870f0aef092cba4084d6f8141310e9ec998e940887cf2c

Observation 56e36f43-8424-40ea-a6a1-99d718592339 · outbound

This paper cites Hts-at: A hierarchical token-semantic audio transformer for sound classification and detection.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Hts-at: A hierarchical token-semantic audio transformer for sound classification and detection

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:15.091237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:15.091237Z digest=sha256:50f5b6d73a669182d687707a4526887cac1e32fbd79f926ff7e01a098616be07

Observation 54ea8e34-ffa0-4245-9962-ef434d3f3ae7 · outbound

This paper cites DialogSum: A Real-Life Scenario Dialogue Summarization Dataset.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning DialogSum: A Real-Life Scenario Dialogue Summarization Dataset

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:15.207070Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:15.207070Z digest=sha256:a7d79c27ea6b31f94c09b530de84f172990da43945e4f1f25cdf4502083443c3

Observation c5a18620-9c84-4893-8dc8-db6d1adb2641 · outbound

This paper cites Multi-Task Learning with Deep Neural Networks: A Survey.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Multi-Task Learning with Deep Neural Networks: A Survey

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:15.356316Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:15.356316Z digest=sha256:62635c95c083c92d3336ae93fae2ccfce66c51e73c1a8e5a6683c114c4be12f7

Observation 56770bc8-7652-4291-a236-385ac86ae989 · outbound

This paper cites One Model, Multiple Modalities: A Sparsely Activated Approach for Text, Sound, Image, Video and Code.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning One Model, Multiple Modalities: A Sparsely Activated Approach for Text, Sound, Image, Video and Code

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:26.526503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:15.468308Z digest=sha256:887dc45402b9e027774991e3f276fc50bbe447b6ab0c93a531ccf644d01f9db9

Observation 9f3f72e0-b581-414b-8320-fc49d0f25b45 · outbound

This paper cites Imagenet: A large-scale hierarchical im- age database.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Imagenet: A large-scale hierarchical im- age database

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:15.576775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:15.576775Z digest=sha256:794c1c2df8105f360f844029fad64e5ae5b42e449d14f872fb45fb71f6a8dc7b

Observation fd8e2161-2dab-474d-aafc-305379b80e06 · outbound

This paper cites BERT: Pre-training of deep bidirectional trans- formers for language understanding.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning BERT: Pre-training of deep bidirectional trans- formers for language understanding

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:15.745612Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:15.745612Z digest=sha256:0375ea68a1851bf6ac9a2c1d284561ae0618cf3a3e1235553550fa34650a7cdd

Observation 37f95f8a-075d-45ee-b344-173b91ad14b3 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:15.911782Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:15.911782Z digest=sha256:84bc621d1e58863c2d79ac6c76723e8ce6ecd2a6e191769b6fa3617b8a76ff95

Observation 55f8eeca-e4a0-4882-b8a4-0aead2363c93 · outbound

This paper cites A generaliza- tion of transformer networks to graphs.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning A generaliza- tion of transformer networks to graphs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:16.035240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:16.035240Z digest=sha256:1c01835f3f049ea2c6fe2b68362c9b8f5ef21b9a73c3d0f12317cc4830e4d655

Observation dc85f918-01e8-42ac-9e37-65f9d313ef58 · outbound

This paper cites Efficiently identifying task group- ings for multi-task learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Efficiently identifying task group- ings for multi-task learning

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:16.197878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:16.197878Z digest=sha256:0e44aec1b029a6ab10653838415df351b681d6c75f271055a4d55c2a08618cc8

Observation dae79af5-b80a-4199-8116-20bd3ce75789 · outbound

This paper cites End-to-End Audio Strikes Back: Boosting Augmentations Towards An Efficient Audio Classification Network.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning End-to-End Audio Strikes Back: Boosting Augmentations Towards An Efficient Audio Classification Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:16.324991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:16.324991Z digest=sha256:ff15d246e5bfb3a4b6a95f4e747b926bba144ea92f174b6c5a149cf8edd5adfc

Observation d9b07949-7982-4622-b307-34c03a5d24c0 · outbound

This paper cites Audio set: An ontology and human- labeled dataset for audio events.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Audio set: An ontology and human- labeled dataset for audio events

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:16.490217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:16.490217Z digest=sha256:b94d180d089e2d4b11388770c938476219cc2563c5323b6bdae4cc037ecbd99b

Observation 932aad78-8e47-4db1-aeef-9cf77236476d · outbound

This paper cites OmniMAE: Single Model Masked Pretraining on Images and Videos.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning OmniMAE: Single Model Masked Pretraining on Images and Videos

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:26.285243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:16.598120Z digest=sha256:fffeeeb0d58fb39ca6f100190a5a9dad913ebbcf275d2e02a0582ae1301abb77

Observation f09ffe82-9b83-4d02-aef9-ae2198364411 · outbound

This paper cites Omni- vore: A single model for many visual modalities.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Omni- vore: A single model for many visual modalities

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:16.711982Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:16.711982Z digest=sha256:b71e1efaf3dfe07a09645d80653e9e92a922119998df52d3c44ed8a21b632e65

Observation 475ea48a-1140-4d20-9ed5-3805e5e0fc1e · outbound

This paper cites SAMSum Corpus: A Human-annotated Dialogue Dataset for Abstractive Summarization.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning SAMSum Corpus: A Human-annotated Dialogue Dataset for Abstractive Summarization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:16.915486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:16.915486Z digest=sha256:8a5da00b78c312d667cf93cc781f69f08842bb91d464622a024a6f9dd2365c8e

Observation c5ee46f4-7d1b-482f-8cad-1e106d61c0a3 · outbound

This paper cites AST: Audio Spectrogram Transformer.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning AST: Audio Spectrogram Transformer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.039560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.039560Z digest=sha256:2102452f235b3ce79b9e4644cf1e97e0c7513a21b1c61212f08dc7e41b8cd38b

Observation e59a712b-21b1-4fb7-81b0-22fb95d4f43f · outbound

This paper cites Uavm: Towards unifying audio and visual models.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Uavm: Towards unifying audio and visual models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.202842Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.202842Z digest=sha256:1af8f62e893752c37aec7fd09c7fe7950c9e06a394cbc46cf3da4aca216e6cec

Observation 1ef8eb6e-cdb5-41bb-9fbe-0567c56ac70e · outbound

This paper cites The” something something” video database for learning and evaluating visual common sense.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning The” something something” video database for learning and evaluating visual common sense

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.322899Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.322899Z digest=sha256:d9e56d35b5f2036b620003219dcfc0185593ab09ea8d4f3108e0ddb54adff949

Observation 4684aabb-05a1-438c-b075-9b2859961c47 · outbound

This paper cites Ego4d: Around the world in 3,000 hours of egocentric video.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Ego4d: Around the world in 3,000 hours of egocentric video

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.420096Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.420096Z digest=sha256:60454bfc01a10df4e106eb063649f61a8cb5016e9cb45494b7d961e1e2c71e6c

Observation 718fe2f5-fe32-4d5c-84fc-5ac5709e6bdf · outbound

This paper cites Dynamic task prioritization for multitask learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Dynamic task prioritization for multitask learning

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.507498Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.507498Z digest=sha256:7007431248efe705970c0860b51d0e3f1db17c7c79bb50755ba05f77a696db48

Observation 87e16809-6d0e-4eda-b546-ff98c30ed626 · outbound

This paper cites MaskViT: Masked Visual Pre-Training for Video Prediction.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning MaskViT: Masked Visual Pre-Training for Video Prediction

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.617134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.617134Z digest=sha256:292287a8a4ce2906a0cc4fa386e9ce988bcf592a9b883d1e7b7a4a0297a993d4

Observation e0b9e3d5-d3ef-4db0-92b5-c89935dbf407 · outbound

This paper cites Masked autoencoders are scal- able vision learners.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Masked autoencoders are scal- able vision learners

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.712917Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.712917Z digest=sha256:1fe7888b64881612929ef98f8dfad51ae93d77a3a6df89559070a7f14050186c

Observation 0a79934a-741c-4e41-b197-0f059c137274 · outbound

This paper cites Gaussian Error Linear Units (GELUs).

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Gaussian Error Linear Units (GELUs)

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.826839Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.826839Z digest=sha256:5fbd7b86cf8526621ce447033fd76ef8e62b676d93f87d8ece0dba85ea70539f

Observation 1f89a59d-b2a3-4f03-bbb4-dad5dfd4add6 · outbound

This paper cites Spectral- former: Rethinking hyperspectral image classification with transformers.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Spectral- former: Rethinking hyperspectral image classification with transformers

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:17.932549Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:17.932549Z digest=sha256:ed43e5939ad04b7d9d58001f689d6c11bc4d13e262cb037040caa735e8f9be8a

Observation 621186a5-ba8c-416d-8ff9-bd93cdde9a52 · outbound

This paper cites Unit: Multimodal multitask learning with a unified transformer.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Unit: Multimodal multitask learning with a unified transformer

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.021377Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.021377Z digest=sha256:2cac1d7e9bb93792c4a4b2a3891d92470ba537157b305d10d2d2d64464103e58

Observation 84660675-7b4b-4f5c-82f4-da2d19e23c2c · outbound

This paper cites OGB-LSC: A Large-Scale Challenge for Machine Learning on Graphs.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning OGB-LSC: A Large-Scale Challenge for Machine Learning on Graphs

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.119583Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.119583Z digest=sha256:81dbc22ec2b09959e76c6294a35df7b2cac647704f15487bfa8f8483f046a8a7

Observation 342a32dc-600d-4380-a433-3c3e8138eded · outbound

This paper cites Perceiver IO: A General Architecture for Structured Inputs & Outputs.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Perceiver IO: A General Architecture for Structured Inputs & Outputs

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.251387Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.251387Z digest=sha256:bac28491ef48c9729bedb38780dd862e0666019ed8631170da53cb287b18a006

Observation a5d0eb73-3086-4f59-aead-67f1a0556924 · outbound

This paper cites Perceiver: General perception with iterative attention.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Perceiver: General perception with iterative attention

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.379108Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.379108Z digest=sha256:430479372ca373cd00eaa570eaa60840958fbee53e67f987bbf0a817d81c07bb

Observation 9d9f546e-fc24-4b2c-89fc-3a190f2f92bb · outbound

This paper cites A review of multimodal image matching: Methods and applications.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning A review of multimodal image matching: Methods and applications

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.514139Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.514139Z digest=sha256:01520f30d00576d55a358ca9e51a6169d607a2f40a9c8f9e76d50251bd0472f3

Observation 9cd1fe4c-4482-434f-b3e5-f327b37fc8f1 · outbound

This paper cites One Model To Learn Them All.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning One Model To Learn Them All

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.598502Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.598502Z digest=sha256:6ccb2534d582bbf8ecc391321119108b93bee88750be0bace8f9cb823d7eb17b

Observation 2c9ea7e5-0131-4774-9b13-11500b736714 · outbound

This paper cites The Kinetics Human Action Video Dataset.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning The Kinetics Human Action Video Dataset

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.711195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.711195Z digest=sha256:26d770d99df1b423adfed4e0dcbd2b6a68a3b21c2efed7920ce9663c3f47c131

Observation 30d3f7e6-cc88-430e-8730-dc40862b5df6 · outbound

This paper cites Mind the Gap! Injecting Commonsense Knowledge for Abstractive Dialogue Summarization.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Mind the Gap! Injecting Commonsense Knowledge for Abstractive Dialogue Summarization

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:25.963938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:18.839182Z digest=sha256:5dc56d94d6dc2d3f04de22c9c5ead945afce67b76fb92cf6553d088b9abdfcab

Observation 0b31472a-8eda-49b8-9476-252fb415548f · outbound

This paper cites Re- former: The efficient transformer.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Re- former: The efficient transformer

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:18.937345Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:18.937345Z digest=sha256:6d7f57d71a44ce247e40fe42035b3f7c94c069237ef3984e40090f526b3ccab3

Observation e7b145c0-c1a0-4ea9-b71c-f24d22135333 · outbound

This paper cites Hmdb: a large video database for human motion recognition.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Hmdb: a large video database for human motion recognition

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:19.051884Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:19.051884Z digest=sha256:62e2d38212a18686e592d7e304792f8b9a2cee87910a506c4d6e7d787f603bdf

Observation 1345fdee-e667-447e-8af6-ab88c4dd7d26 · outbound

This paper cites Modeling long-and short-term temporal patterns with deep neural networks.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Modeling long-and short-term temporal patterns with deep neural networks

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:19.166703Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:19.166703Z digest=sha256:2b9b9d3db674aa86def42479f31d31a5e6e11ef076e42e0428d4a40f2c6f504e

Observation 2b5de6d7-497d-47c5-9c91-c72f11041bf1 · outbound

This paper cites Stratified trans- former for 3d point cloud segmentation.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Stratified trans- former for 3d point cloud segmentation

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:19.290419Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:19.290419Z digest=sha256:4b9d54e8820b136ae2eef2aa4028803fd96667d85f36786cfd5a9a8c3e306d3f

Observation 6230888b-edf8-4947-8121-8823afc854e5 · outbound

This paper cites Regu- larization strategy for point cloud via rigidly mixed sample.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Regu- larization strategy for point cloud via rigidly mixed sample

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:32.311366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:19.428323Z digest=sha256:b75cf24ce8ad6a306b56884124e9762d990f04c1676d7d8dd16d496341f3bffa

Observation 1d5913ce-11ba-4c0d-8c9e-28202a97a75c · outbound

This paper cites Uni-perceiver v2: A generalist model for large-scale vision and vision-language tasks.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Uni-perceiver v2: A generalist model for large-scale vision and vision-language tasks

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:32.222632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:19.541910Z digest=sha256:a21e34e43fd2d7e6b0eba2e40723b9cadc40a6fe3bd37c1e6925f0922136fe95

Observation f1a474c0-6c44-45d3-9420-44286c7b83b3 · outbound

This paper cites UniFormerV2: Spatiotemporal Learning by Arming Image ViTs with Video UniFormer.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning UniFormerV2: Spatiotemporal Learning by Arming Image ViTs with Video UniFormer

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:19.646948Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:19.646948Z digest=sha256:0c99ceed4cd4d43d177b8302e798c1923eca8c6de238acd7b5af889d837109eb

Observation 941ae255-0a70-4627-8651-0196319e1ce4 · outbound

This paper cites Enhancing the locality and breaking the memory bottleneck of transformer on time series forecasting.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Enhancing the locality and breaking the memory bottleneck of transformer on time series forecasting

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:32.074110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:19.768028Z digest=sha256:08a7245a1d465958bc253419d79fcf9d4092d61dc476f72b05171349007287a1

Observation 9a8f0407-498c-4677-afa6-5800b0e5c46b · outbound

This paper cites Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Mind the Gap: Understanding the Modality Gap in Multi-modal Contrastive Representation Learning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:19.855755Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:19.855755Z digest=sha256:2c24cc6241e988d774a5a644610aeddf5d7f7d8f52b15887f43c62b94a43c406

Observation df00d131-ea42-45f1-bd8c-76ec203b2e1c · outbound

This paper cites Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Belongie, James Hays, Pietro Perona, Deva Ramanan, Piotr Doll ´ar, and C

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.892880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:19.953525Z digest=sha256:a56869cba6e9985ff5f12405201231d0895655e9cfd57f0728abd0b8304411b4

Observation 56008d9f-43a8-4667-b043-2ab6eeb4e977 · outbound

This paper cites OPT: Omni-Perception Pre-Trainer for Cross-Modal Understanding and Generation.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning OPT: Omni-Perception Pre-Trainer for Cross-Modal Understanding and Generation

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:25.761552Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:20.051811Z digest=sha256:6cf0710b99dfed7a1e9970790519803f76109aed04af1a06e219b675c07e6a6b

Observation 8f652af2-877d-4b29-90cc-df16bbcac204 · outbound

This paper cites Pyraformer: Low- complexity pyramidal attention for long-range time series modeling and forecasting.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Pyraformer: Low- complexity pyramidal attention for long-range time series modeling and forecasting

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.679062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:20.168295Z digest=sha256:34b07866e1e49878b7710d95730b9f544500c20b17e0936af0168d08954c15cc

Observation 08f78ae6-58ba-42ce-b471-ac7950235867 · outbound

This paper cites RoBERTa: A Robustly Optimized BERT Pretraining Approach.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning RoBERTa: A Robustly Optimized BERT Pretraining Approach

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:20.304795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:20.304795Z digest=sha256:9e5c691eb8857621dc14661041a71bc4abc56ec740727cb499b878fb99689703

Observation 194acbc9-711a-48bb-9adc-f519ce36393a · outbound

This paper cites Moments in time dataset: one million videos for event understanding.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Moments in time dataset: one million videos for event understanding

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.539943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:20.398454Z digest=sha256:b437723b2135950611bd8818df8bc4741800599a5a01dcb24af90f7868023c16

Observation 43b4d318-5029-4061-b742-6bc3205a9dad · outbound

This paper cites Person recognition system based on a combination of body images from visible light and thermal cameras.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Person recognition system based on a combination of body images from visible light and thermal cameras

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.423488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:20.525721Z digest=sha256:54858ed8d1ddf50cc39c42eef91bae724bb6082984d28d583ea9df5806c564f1

Observation 43fa27db-de94-4685-827b-56ae6cb30588 · outbound

This paper cites N-BEATS: Neural basis expansion analysis for interpretable time series forecasting.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning N-BEATS: Neural basis expansion analysis for interpretable time series forecasting

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:20.661670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:20.661670Z digest=sha256:fdd1289e54d456c2131104a99e4582197edd9d10d042619a480f8858bbe6aa2b

Observation 60e8e3d8-5162-4883-a951-b6a1b9b0e488 · outbound

This paper cites Cats and dogs.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Cats and dogs

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:20.767129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:20.767129Z digest=sha256:533694ff9ce83fef8a91dc4c6ab3331e49f1469e03705d47cf0a171acb260a80

Observation ed785505-9347-4868-91d0-3a3ae50cc0af · outbound

This paper cites Esc: Dataset for environmental sound clas- sification.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Esc: Dataset for environmental sound clas- sification

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.311763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:20.883668Z digest=sha256:3791335335e1a82bce87387c699eecc6335f681c9cdb39468d3724305dccd7e2

Observation 6b0d0f90-f56e-4ab7-8257-a0b3f156ec28 · outbound

This paper cites Re- thinking video vits: Sparse video tubes for joint image and video learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Re- thinking video vits: Sparse video tubes for joint image and video learning

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.184763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:20.990939Z digest=sha256:65443433a9bd806318cbb5c287c4f39b84e0f2f6873f15861e14767d3ac4409a

Observation 8a5f51c5-9667-473c-9240-5e22df4fd717 · outbound

This paper cites OmniNet: A unified architecture for multi-modal multi-task learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning OmniNet: A unified architecture for multi-modal multi-task learning

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:21.121011Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:21.121011Z digest=sha256:2ffab2535b1f1e0b655f83f1dcad7fef08ff105d1593f12c8033ecc0556ce47b

Observation 960b3122-f46a-4655-8e8c-c6780eec6014 · outbound

This paper cites Point- net++: Deep hierarchical feature learning on point sets in a metric space.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Point- net++: Deep hierarchical feature learning on point sets in a metric space

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:31.040826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:21.235366Z digest=sha256:137ec8dc3a32033c926031a50208fe9d4de55a31083192c1861eef02999cacbd

Observation 85a502b9-ace2-44b5-b02a-a1445bda232b · outbound

This paper cites Improving language understanding by gen- erative pre-training.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Improving language understanding by gen- erative pre-training

Reference 60

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:21.380262Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:21.380262Z digest=sha256:02eb4a8b7380bc7a40b50edf069f1a2278190c61cb6a88802f77aa4f96a5a4e8

Observation cebe27b6-d809-402e-aa63-1b51e2012ace · outbound

This paper cites Reliable tuberculosis de- tection using chest x-ray with deep learning, segmentation and visualization.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Reliable tuberculosis de- tection using chest x-ray with deep learning, segmentation and visualization

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.946155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:21.519423Z digest=sha256:9e4c9414afabdb753f1ca2af87726107fb884be87772cb69b717541ad96b569f

Observation 5f88d135-92c1-4203-823c-dee0e0d6a492 · outbound

This paper cites Zorro: the masked multimodal transformer.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Zorro: the masked multimodal transformer

Reference 62

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:25.551900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:21.672452Z digest=sha256:bbbc374c57a2cd6e052e34fa05f6d0917a3642e4ce8dde9b0d177b918e6c75f8

Observation 3f24eccf-3cc9-461e-a080-e9ab0c57190f · outbound

This paper cites Indoor segmentation and support inference from rgbd images.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Indoor segmentation and support inference from rgbd images

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.817984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:21.776607Z digest=sha256:0cf1805d7caf341b5a6251680005804a9e108e92bcc4e0d13195e97b687cf70a

Observation 21993166-f068-41a5-84ad-b57e3fca20bc · outbound

This paper cites Mpnet: Masked and permuted pre-training for lan- guage understanding.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Mpnet: Masked and permuted pre-training for lan- guage understanding

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.698263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:21.861378Z digest=sha256:815a0e2d1a5e10576207fce30ddb61740bd528bdd678579d09e58ac0019a69dd

Observation 9a540032-3568-429f-af77-4e6fe5f3d6a0 · outbound

This paper cites Sun rgb-d: A rgb-d scene understanding benchmark suite.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Sun rgb-d: A rgb-d scene understanding benchmark suite

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.580490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:21.997522Z digest=sha256:b0d31071977395c441aad0240d3489aed804a158a91f730b65cd6e94b5e0abd3

Observation f8a1fc2e-e727-48e9-b9d7-b1eae07b61c2 · outbound

This paper cites UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning UCF101: A Dataset of 101 Human Actions Classes From Videos in The Wild

Reference 66

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:22.073655Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:22.073655Z digest=sha256:2c77a66bddada3ccaf8336997f53a4b5778e0b82165209aeec4e2208c4aef402

Observation 3fed73df-e32b-4259-aa76-b7be8cddd991 · outbound

This paper cites OmniVec: Learning robust representations with cross modal sharing.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning OmniVec: Learning robust representations with cross modal sharing

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:22.147490Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:22.147490Z digest=sha256:27d725536147341c6632869774b39ae98067a687cb6a85f93127ae9f1b7c8521

Observation 8049192c-b68b-414c-a96b-c586b6cd28ba · outbound

This paper cites Hierarchical multi-task learning via task affin- ity groupings.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Hierarchical multi-task learning via task affin- ity groupings

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.455919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.210014Z digest=sha256:fcd69a4037811e922d5fa3248da44c62792db533879c454b8fd54ee0433c0f3b

Observation f3edae28-32b7-4164-8523-4d5b73918175 · outbound

This paper cites Benchmarking Robustness of 3D Point Cloud Recognition Against Common Corruptions.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Benchmarking Robustness of 3D Point Cloud Recognition Against Common Corruptions

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:22.270903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:22.270903Z digest=sha256:db9ed553654177256144c1821879203d5224915b2d71ffa7c29bf52617376e2b

Observation edf6e8df-359f-4bad-8d02-a2c8f3fbc02c · outbound

This paper cites Efficientnet: Rethinking model scaling for convolutional neural networks.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Efficientnet: Rethinking model scaling for convolutional neural networks

Reference 70

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:22.333424Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:22.333424Z digest=sha256:968460bbcb22b9f53435f5e0a9f56e94b04b5e29c7f093f1799fab166e199bfe

Observation c367eb61-daef-4882-96a4-71cc92cb0c7f · outbound

This paper cites Contrastive boundary learning for point cloud segmentation.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Contrastive boundary learning for point cloud segmentation

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.308893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.389608Z digest=sha256:a991eca7f13dbf3555bd50cbc8f0e9e02cc60403004c0c63f3c8b3db455de9d1

Observation 3271b91b-af12-4545-bb64-f5f2e23434bb · outbound

This paper cites Small sample hyper- spectral image classification based on the random patches network and recursive filtering.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Small sample hyper- spectral image classification based on the random patches network and recursive filtering

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:30.150417Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.462428Z digest=sha256:2ebee60d0e057d5b209f2c6f49bc82f379321d62d046cfdcdd3d16ac23b4bef1

Observation f1b25de1-73cf-4f19-b089-6357805ff965 · outbound

This paper cites Revisiting point cloud classification: A new benchmark dataset and classification model on real-world data.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Revisiting point cloud classification: A new benchmark dataset and classification model on real-world data

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:29.959279Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.537912Z digest=sha256:9c0389efce35f481a20a90899d38d1ed84a6481ec63a38e71429827f069e2b75

Observation 327fdeff-ef5c-489b-9c7f-53dee17d507c · outbound

This paper cites The inaturalist species classification and detection dataset.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning The inaturalist species classification and detection dataset

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:29.779706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.632705Z digest=sha256:34a46293eed6e09e6a9a6ca63947978a4d5d6d3573aad6862543dba64c291f2b

Observation b6988d14-15b5-46a2-98e9-f821a035c5a3 · outbound

This paper cites Attention is all you need.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Attention is all you need

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:29.627959Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.716235Z digest=sha256:3a0865e34b386d06a8495a3506dd7c63727e51f5b6f8d4a230285c816a0de925

Observation 11b94fd8-c192-4958-8fbc-4900b5be83e0 · outbound

This paper cites Internimage: Exploring large-scale vi- sion foundation models with deformable convolutions.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Internimage: Exploring large-scale vi- sion foundation models with deformable convolutions

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:29.467629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.775789Z digest=sha256:388b5913aa6cd03fdc87b64391ac0b080276d212a01b7e255c423eb08c23abba

Observation 90af2435-5042-4201-ad4c-2e1abd6c5506 · outbound

This paper cites InternVideo: General Video Foundation Models via Generative and Discriminative Learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning InternVideo: General Video Foundation Models via Generative and Discriminative Learning

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:22.840001Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:22.840001Z digest=sha256:40b42152f308ac02550c9c13ae0ae1735983b1c7b0fef8da4de555869a210d32

Observation f9b85ae5-edfd-4db3-9b20-6bd5903666be · outbound

This paper cites Masked feature pre- diction for self-supervised visual pre-training.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Masked feature pre- diction for self-supervised visual pre-training

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:29.260075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.894679Z digest=sha256:f5f97d973ed74fd8c3584dea5f6e516dbdb006478075f8dcb5c19d4534305e16

Observation 349ba8bd-f659-4ebf-8a76-276997f08069 · outbound

This paper cites Syn- cretic modality collaborative learning for visible infrared person re-identification.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Syn- cretic modality collaborative learning for visible infrared person re-identification

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:29.084751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.941403Z digest=sha256:7f473f87f2d8e48322745e8328a7db4ab2871f06014da228af4a1f0e714746da

Observation 0d42cd48-9c56-4dc2-846e-3514925e8c9a · outbound

This paper cites Controllable Abstractive Dialogue Summarization with Sketch Supervision.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Controllable Abstractive Dialogue Summarization with Sketch Supervision

Reference 80

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:25.321727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:22.994077Z digest=sha256:6378f61a2d9edd36b39036717e1c2440db476abe33b6694cbfba3b48d529b530

Observation cecf0c2d-d737-4047-a8ac-88955a12dafb · outbound

This paper cites Tinyvit: Fast pretraining distillation for small vision transformers.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Tinyvit: Fast pretraining distillation for small vision transformers

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.942665Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.044939Z digest=sha256:7cd52dd0acbedb1558183522f48e8888e5704d17a7675a7364f584d504f7f963

Observation 825288bf-78c8-4094-8cf0-1eaa9b17c048 · outbound

This paper cites Point transformer v2: Grouped vector at- tention and partition-based pooling.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Point transformer v2: Grouped vector at- tention and partition-based pooling

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.793382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.101262Z digest=sha256:350813c0df7eb236f8353c3787eb5b0571015c6f71a56736d2dc0e545a61e90e

Observation 4c15fc0b-37a7-42c0-a0ca-8f888624e3ab · outbound

This paper cites 3d shapenets: A deep representation for volumetric shapes.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning 3d shapenets: A deep representation for volumetric shapes

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.654139Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.180431Z digest=sha256:9ea68f82922c6ca6107a503ee44aca3e3faaa7ddcef079af0aeba5bbfcd69c7d

Observation d4f48735-2087-4877-8af9-04073b46d0bd · outbound

This paper cites Audiovisual SlowFast Networks for Video Recognition.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Audiovisual SlowFast Networks for Video Recognition

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:23.240129Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:23.240129Z digest=sha256:6ea502cee7c8176bbdc65cc87036ff9d09704860e607363118d41debde7226be

Observation bfb374f0-7601-4427-96df-5ee5faf71b0e · outbound

This paper cites Msr-vtt: A large video description dataset for bridging video and language.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Msr-vtt: A large video description dataset for bridging video and language

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.456028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.286212Z digest=sha256:f31896221146a0b12a4a19ee3d6067ac3b2d132e6deb0124fc4dd24fb572a8da

Observation e0c2cd51-f150-447c-a73f-a643a7c81994 · outbound

This paper cites Multimodal Learning with Transformers: A Survey.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Multimodal Learning with Transformers: A Survey

Reference 86

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:25.148397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.399732Z digest=sha256:2cbd932416b7fb7db666ab18dfd708b7793cd60abc3dfe9a166d0242cdde7452

Observation cfa2cd38-10c1-42a4-8f1e-7a24a37a3490 · outbound

This paper cites Multi-modal masked pre-training for monocular panoramic depth completion.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Multi-modal masked pre-training for monocular panoramic depth completion

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.330526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.499703Z digest=sha256:699f5d1b002b460eeb8b5a4598864c727f854033102c8cc7281611b0e94809b4

Observation f955fc72-ae65-4e44-bf53-cb3a6af81b53 · outbound

This paper cites Swin3D: A Pretrained Transformer Backbone for 3D Indoor Scene Understanding.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Swin3D: A Pretrained Transformer Backbone for 3D Indoor Scene Understanding

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:23.580652Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:23.580652Z digest=sha256:7a48b310661a4b6877341972ca79820cd0e365022888865790536235690e27d3

Observation b008ece1-7827-4814-802a-93820de787d8 · outbound

This paper cites XLNet: Generalized Autoregressive Pretraining for Language Understanding.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning XLNet: Generalized Autoregressive Pretraining for Language Understanding

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:23.682264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:23.682264Z digest=sha256:93e7ba0da51b02be354f83bd07987011419b29821dfe65d987038cee21b14dea

Observation 25e42e8f-0a25-4e26-a6fa-ea9b67d18dbd · outbound

This paper cites Deep Learning for Person Re-identification: A Survey and Outlook.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Deep Learning for Person Re-identification: A Survey and Outlook

Reference 90

Resolution
verified exact
local_arxiv, observed 2026-08-06T19:50:24.969127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.740334Z digest=sha256:fdf3ea44948808011d0672107b7331a2073565d448f68f778b1d0b8181868c81

Observation 240ad846-4d80-4e93-a896-f29ed0221efb · outbound

This paper cites Do transformers really perform badly for graph representation? In Thirty-Fifth Conference on Neural Information Process- ing Systems, 2021.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Do transformers really perform badly for graph representation? In Thirty-Fifth Conference on Neural Information Process- ing Systems, 2021

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.189988Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.790277Z digest=sha256:be81ecde3afc20b60a4745d4d497915f8e0dcfe6ac1fc41eb06a4630ee3ba920

Observation 7aceffcd-fbf6-4549-99b2-da442e37cdde · outbound

This paper cites CoCa: Contrastive Captioners are Image-Text Foundation Models.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning CoCa: Contrastive Captioners are Image-Text Foundation Models

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:23.861350Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:23.861350Z digest=sha256:8d3c822dd1473ada7faf576aa3ad5500ad4b4ea634719f76a809f5d1768422d9

Observation e1e514a6-2fda-4a17-8cac-dda7c93bbe04 · outbound

This paper cites Metaformer is actually what you need for vision.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Metaformer is actually what you need for vision

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:28.047216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.921175Z digest=sha256:6ca93f760e9badd669cf324facf36625e3532d6fa055a590ed800d3064cf3279

Observation b4baed03-004d-4915-9ad5-4c3216d2b102 · outbound

This paper cites Point-bert: Pre-training 3d point cloud transformers with masked point modeling.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Point-bert: Pre-training 3d point cloud transformers with masked point modeling

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:27.928246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:23.991722Z digest=sha256:4e52b8a730373bb603aa6c1d1a3ad4ab28ffe54fddcb6e0a5c773ac54540be73

Observation 4ffff764-3d5f-4d5f-8564-951c54b457e6 · outbound

This paper cites Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Socratic Models: Composing Zero-Shot Multimodal Reasoning with Language

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:24.066566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:24.066566Z digest=sha256:f8bba1eda2daca48822ea2b7baa1f2311c03b2f486984c0df8681b4b9c08d0e6

Observation b2ebe9a9-9854-4749-93c2-f4b57288dbc8 · outbound

This paper cites Point- cutmix: Regularization strategy for point cloud classifica- tion.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Point- cutmix: Regularization strategy for point cloud classifica- tion

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:27.750612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:24.114385Z digest=sha256:dcd33b330a01800ea43aeb9ce81d49945df1e97e2a695ed2d463bc3a34897861

Observation 521fa700-e1e6-4b00-8163-3c2a9f01e95c · outbound

This paper cites An overview of multi-task learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning An overview of multi-task learning

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:27.594862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:24.164336Z digest=sha256:1f4f5186e23f229a88c35abb33208e16c3dca54539576ed708fed271ec03a8c3

Observation 3953969f-63e7-4d49-91f8-48982b077969 · outbound

This paper cites Modality synergy complement learning with cascaded aggregation for visible-infrared person re- identification.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Modality synergy complement learning with cascaded aggregation for visible-infrared person re- identification

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:27.512561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:24.215744Z digest=sha256:b48faaa2f90442309c26950367e950fb433a829e8e1ff9fb5dfd73944a49661a

Observation af7aa99d-b892-47ac-a506-e0154c640fe8 · outbound

This paper cites Meta-Transformer: A Unified Framework for Multimodal Learning.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Meta-Transformer: A Unified Framework for Multimodal Learning

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-06T19:50:24.267156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:50:24.267156Z digest=sha256:40509868e461b167b53c331f22b45cb1c2e39187964c4199db50f2f3ea0ed486

Observation a316fe08-b859-461a-853f-a88b28f6f36c · outbound

This paper cites Places: A 10 million image database for scene recognition.

OmniVec2 -- A Novel Transformer based Network for Large Scale Multimodal and Multitask Learning Places: A 10 million image database for scene recognition

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T19:50:27.367186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-06T19:50:24.303188Z digest=sha256:e51e7e658b092b4263169580ecd6619c9d4290d1febe64053446a18b8e47616c

Pith citing papers

No inbound Pith citation observations are available.