Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:42:12.664774Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 77 of 77 outbound references and 10 inbound Pith citation observations for arXiv:2501.04322.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-10T21:42:12.664774Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:20:32.873352Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T04:27:37.114920Z
77 of 77 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 1d451084-42cc-4037-9506-b6de5e913866 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf5c3248-f3e2-4ea4-af2b-df9bd936d500 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bf881bc-0e99-4caf-bb61-9adde70df0d8 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Qwen-VL: A Versatile Vision-Language Model for Understanding, Localization, Text Reading, and Beyond
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc827713-8028-4d87-be55-0c7700e615c6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts VLMo: Unified Vision-Language Pre-Training with Mixture-of-Modality-Experts
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7c42f80-daee-48be-a582-c3d128abda0a · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Stable LM 2 1.6B Technical Report
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 528ffb5e-b3de-46df-bebd-62c3cbae8d91 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a7885e44-5241-40fd-943e-981a12235f99 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9ba9f3fd-6113-4c60-8bf7-7f067e8d82c6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 997d9897-a14b-4941-969f-f27c0cfdddba · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts ALLaVA: Harnessing GPT4V-Synthesized Data for Lite Vision-Language Models
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9530c575-dd45-4a9b-ab18-c2c1da59390f · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts ShareGPT4V: Improving Large Multi-Modal Models with Better Captions
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a901c482-064d-4f32-9757-53c9a8d624b0 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Improved Baselines with Momentum Contrastive Learning
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4470e77a-389d-44db-b612-8055e91ebebc · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Microsoft COCO Captions: Data Collection and Evaluation Server
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 86c3582a-f8f6-4936-9dc4-3d30d16c36f5 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts E.; et al
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a5ae5a6c-25ed-4757-b29b-542d4ef1dd2b · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts MobileVLM : A Fast, Strong and Open Vision Language Assistant for Mobile Devices
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b88f286f-57d3-40a1-8fdd-ad67ddade48e · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts MobileVLM V2: Faster and Stronger Baseline for Vision Language Model
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caf30f46-7b69-4c04-a4b5-4d0d4cc46265 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts BoolQ: Exploring the Surprising Difficulty of Natural Yes/No Questions
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 700667b4-e726-457f-a950-2d447c765204 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0441bc01-05c1-4d5a-994f-afa660cf4869 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76c5eea4-2ad9-40eb-b291-df0bfe4efdc6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2cacafbb-0278-4263-a8ef-92ebf4c6dd73 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37d4c636-ccab-4782-a2c4-7578d70dd9cc · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts DeepSeek-Coder: When the Large Language Model Meets Programming -- The Rise of Code Intelligence
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2f8134f-0f81-4bca-88ab-a1eb0cb86a07 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation abbbb94c-7da2-4203-939c-82dd5abcbb8c · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Measuring Massive Multitask Language Understanding
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ad1b4f3d-cac2-4f07-a91c-34249e7e90b2 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts LoRA: Low-Rank Adaptation of Large Language Models
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4e5d319-2795-4d7e-9714-730c70a1af34 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f9c7fd53-f26e-4f6c-9628-16ce04ee39a6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7adc87b0-0f60-48d8-888b-83f38c188b61 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts A.; and Manning, C
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8d0eac5-fa87-4c8c-9a9a-e77baf23af2d · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts A.; Jordan, M
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5bd11329-9335-4a22-b030-4b83e987769e · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Mixtral of Experts
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19e2f290-ecaa-496b-8802-6554e1b91693 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7e0c649d-f098-4509-a60b-38a6120a5777 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7976cf8b-2fc9-45a8-bc16-fe823f9812cc · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 885830d2-d653-44b0-a230-d18c1a682fba · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts CMMLU: Measuring massive multitask language understanding in Chinese
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ed641db-35ae-4c7d-9311-d916a6dde100 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bd1231c0-5bab-4062-a9a6-4704c070bda8 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07f54a43-9a51-491b-9b86-305f3d3c0dd2 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Evaluating Object Hallucination in Large Vision-Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1cfb2fd5-93dd-4eff-8f0a-6d59db422156 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts PaCE: Unified Multi-modal Dialogue Pre-training with Progressive and Compositional Experts
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fcfa30e-461a-456f-80cb-a501101b0056 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Mini-Gemini: Mining the Potential of Multi-modality Vision Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 659cb86b-cd8d-4705-93a9-be3602a48f78 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts MoE-LLaVA: Mixture of Experts for Large Vision-Language Models
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 877c007d-f680-4b9a-a2be-3279482ffce6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d78fc038-9e6d-4f6c-ae73-d089b9a8eebb · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Improved Baselines with Visual Instruction Tuning
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 478b0675-ae22-4741-8e1c-dbf4ba481a6e · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation ae1873cc-32ff-43e8-b359-f5317a84774f · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e243fd5d-125b-4c12-a862-c734d38cadd9 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 863ae401-a94c-4bbb-a37e-8cc954b8a695 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts MMBench: Is Your Multi-modal Model an All-around Player?
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 405653a7-573a-4a1d-b89b-fab4481abda2 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 46
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0a5a2dd0-3111-45e4-bb5d-bff582f7af49 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts DeepSeek-VL: Towards Real-World Vision-Language Understanding
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c73def13-7609-4afc-955e-af8eadc2a16c · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cdb5bf3e-7504-4e92-a1d0-eb1054360fc5 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts IconQA: A New Benchmark for Abstract Diagram Understanding and Visual Language Reasoning
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af157eb4-a97c-49ed-b00a-ced4d7fd18f6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28e832ca-d096-41d9-99da-680809a141b6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts ChartQA: A Benchmark for Question Answering about Charts with Visual and Logical Reasoning
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e9325067-30ad-4a22-9d42-bbda19a4be4d · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Don't Give Me the Details, Just the Summary! Topic-Aware Convolutional Neural Networks for Extreme Summarization
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9325ae1c-917e-4426-a23c-fe7931bb8bfe · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a3b5a508-74b2-47e8-a999-8b76bb15d4f6 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Learning Transferable Visual Models From Natural Language Supervision
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd055cde-c53e-4c06-a375-28b75faa5448 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts ImageNet-21K Pretraining for the Masses
Reference 55
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6223f476-e130-40ef-91c3-5d91b9e0a66e · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Scaling Vision with Sparse Mixture of Experts
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 019d9510-af28-4f28-a226-627b9d1d6097 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dea6e367-469f-419d-a28a-4e382e86d29c · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts LAION-400M: Open Dataset of CLIP-Filtered 400 Million Image-Text Pairs
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 42a7f07d-0e9c-47e6-9eb2-27e3fb6c1b78 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 85e20666-c225-4740-8fe5-149ab2ebc6c8 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Measuring Vision-Language STEM Skills of Neural Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 783102d4-6863-417c-8daf-09a3d388a22d · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Scaling Vision-Language Models with Sparse Mixture of Experts
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a9f6e9c-2583-42b5-b382-e50da34bf5c9 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7755448d-70d6-43a7-9964-dc949058f22b · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54600042-a95f-4061-b0ba-8bf77a3f1ccf · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts PanGu-$\pi$ Pro:Rethinking Optimization and Architecture for Tiny Language Models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 894fbb6d-f8f7-41e5-80bc-d53241a1c6e1 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Gemini: A Family of Highly Capable Multimodal Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 53edb665-371c-4ef7-9f56-ed1a9a3a3097 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 68af8662-981f-4c6b-84f2-bfef7d9c9d43 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Image as a Foreign Language: BEiT Pretraining for All Vision and Vision-Language Tasks
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 675b24f2-a236-4e59-a8b9-61864c16132e · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts PanGu-$\pi$: Enhancing Language Model Architectures via Nonlinearity Compensation
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eef4fe56-a567-401d-bc58-8049c12cbdf8 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 049351a9-5f5d-4fd3-94a3-e7519a1f4bd9 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts FewCLUE: A Chinese Few-shot Learning Evaluation Benchmark
Reference 70
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e5e26d06-e49c-4bed-9787-f7b0f29af39f · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Vision-Flan: Scaling Human-Labeled Tasks in Visual Instruction Tuning
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fed9b6c-9c23-4ad5-b1ee-45303176c4db · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts TinyGPT-V: Efficient Multimodal Large Language Model via Small Backbones
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dd87f1b0-543f-4db9-acde-2b3d6eb22148 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts Unresolved cited work
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 649883b4-5106-42e3-96fd-ab2226a01f29 · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts LLaVAR: Enhanced Visual Instruction Tuning for Text-Rich Image Understanding
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bbe55a0e-8a98-4d18-b655-69014a31488a · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts SVIT: Scaling up Visual Instruction Tuning
Reference 76
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8e64968-c8ed-4736-8c20-d820a87b9e5a · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts MiniGPT-4: Enhancing Vision-Language Understanding with Advanced Large Language Models
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b47da43e-68bb-4219-90ff-0a9e8e9856fe · outbound
Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts LLaVA-Phi: Efficient Multi-Modal Assistant with Small Language Model
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32dfa201-5a44-4629-935a-341775b2f928 · inbound
Sa2VA: Marrying SAM2 with LLaVA for Dense Grounded Understanding of Images and Videos Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4a5a7488-3a23-49d7-8382-328a111d4aa5 · inbound
Circle-RoPE: Cone-like Decoupled Rotary Positional Embedding for Large Vision-Language Models Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cf10a8ec-9ea4-4941-9797-345210f4ae62 · inbound
EvoMoE: Expert Evolution in Mixture of Experts for Multimodal Large Language Models Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 879f13ae-d0b9-4b5c-a844-38f843aaf72a · inbound
Mixed-R1: Unified Reward Perspective For Reasoning Capability in Multimodal Large Language Models Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b049d93d-4725-4e24-8888-2c87fa20c840 · inbound
CyberV: Cybernetics for Test-time Scaling in Video Understanding Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16915555-55e8-4406-84dc-2fad6b5f7017 · inbound
Dense360: Dense Understanding from Omnidirectional Panoramas Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 53
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7305e160-f785-41c6-8060-27508a49775c · inbound
DenseWorld-1M: Towards Detailed Dense Grounded Caption in the Real World Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ba8ed5f-3bb5-4fc0-a0b3-d68040702d1a · inbound
Kwai Keye-VL Technical Report Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 2020
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37685f8e-c4ee-4ad8-a34f-d2baf8251d9c · inbound
Multimodal Large Language Model-Enabled Video Translation: A Role-Oriented Survey Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dc1f5f79-6025-4147-93c1-4f3c8fc73760 · inbound
Kwai Keye-VL-2.0 Technical Report Eve: Efficient Multimodal Vision Language Models with Elastic Visual Experts
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.