Pith. sign in

Paper Citation Record · LEDGER

PresentAgent-2: Towards Generalist Multimodal Presentation Agents

As of 19 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 1 inbound Pith citation observation for arXiv:2605.11363.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.11363 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-13T02:29:42.157339Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-02T08:44:26.876152Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

40 of 40 outbound references displayed

  • verified exact24
  • verified fuzzy15
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dadcfd4b-e521-4cf0-a5f7-c6dbd12b7147 · outbound

This paper cites Paper2poster: Towards multimodal poster automation from scientific papers,.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Paper2poster: Towards multimodal poster automation from scientific papers,

Reference 1

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.469529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:3c0adca186cc31ed1daca3d1f59a920454b7670e302ad26a0c2a891011b4ca1f

Observation dca908e0-3930-44ee-bc54-25c4166f53c7 · outbound

This paper cites Presentagent: Multimodal agent for presentation video generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Presentagent: Multimodal agent for presentation video generation

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.757200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:3c11c2d70a5bf8b3ef134e60809d5231bfae768ee7735e47f37f67a388c9fbd4

Observation bf28b32b-a132-46a0-a0c1-93e7cac92259 · outbound

This paper cites Q., and Shou, M.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Q., and Shou, M

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:32:06.475438Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:014588183c8e1aba9c2abe353a886f825c9d71fdb6a303e010396c7d694f93a5

Observation cc8e9da5-6959-4f1c-9338-4917e45097e1 · outbound

This paper cites VideoAgent: Personalized Synthesis of Scientific Videos.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents VideoAgent: Personalized Synthesis of Scientific Videos

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:32:06.466432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:0ea4f76af60820dec2b8f1246d4181b5a3634ab83fa7ab30dfac7bb5a2668e6d

Observation ebd0c220-9ded-4d65-92c2-c9d60055bc7f · outbound

This paper cites Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Talk to Your Slides: High-Efficiency Slide Editing via Language-Driven Structured Data Manipulation

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:32:06.472581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:1534b45e54bf010ce362c24b2be974edcb5b3362acbf6921167855bd05b276e0

Observation 72818650-064f-4dcc-9249-215894a13ad0 · outbound

This paper cites Pptagent: Generating and evaluating presentations beyond text-to-slides.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Pptagent: Generating and evaluating presentations beyond text-to-slides

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.762408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:06811eaac308a6ff1c601c67d221ad6b945e8389f45b0355fc799b1e61c7d559

Observation ab4f3b05-932f-4b06-b539-01fd8e637cc3 · outbound

This paper cites Auto- Slides: An interactive multi-agent system for creating and customizing research presentations.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Auto- Slides: An interactive multi-agent system for creating and customizing research presentations

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.478070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:33cf21ca16b65866af05d5e5cda8a7c37d4f37feb629cc1cf9dcf6ce2786b389

Observation 371d8d01-54df-4256-8870-04a5cf958f9c · outbound

This paper cites Node-based editing for multimodal generation of text, audio, image, and video.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Node-based editing for multimodal generation of text, audio, image, and video

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.480754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:c6609c9acafb781792d132106927ae7571efc38612182f68e09075616d12e82b

Observation 3d7b7d1c-4e97-49d1-9e0f-205daa78cebf · outbound

This paper cites PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents PolyVivid: Vivid Multi-Subject Video Generation with Cross-Modal Interaction and Enhancement

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.483612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:e760a73f97c147dd04859f60f4c0d1633d784bd87d92201cd4d88b6deca998bf

Observation e91fb8d3-1b09-4292-9d6b-8aa21c96918c · outbound

This paper cites Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.441607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:1d577a02c96f769ff23330a4a90342ca4a76be773c85b04e5b85e53c4ee9786c

Observation b49932d2-ed43-4638-8890-6fc98df6d966 · outbound

This paper cites Autopresent: Designing structured visuals from scratch.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Autopresent: Designing structured visuals from scratch

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.695161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:50c47c533937e8c8a5311eb5b8f45335fbd22b7d7848f9595d62124f4f60edaa

Observation f0312a9d-97e9-4134-bf10-12baeb471f50 · outbound

This paper cites Infinity parser: Layout aware reinforcement learning for scanned document parsing.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Infinity parser: Layout aware reinforcement learning for scanned document parsing

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.415184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:fdcd0ff1f24bc48f0045e4c2b0c8e8404493b8dff0e516bef122bf0542ef5757

Observation f796951f-2d8a-4733-b4bb-edcb5d882a95 · outbound

This paper cites Doc2ppt: Automatic presen- tation slides generation from scientific documents.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Doc2ppt: Automatic presen- tation slides generation from scientific documents

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.708349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:17ddd45c8b71db5f74a8ddb4689e054f1db751ce83f0c907c0904c76a4ad4df0

Observation 42ef8323-5ca5-41b8-8b66-f24bdc886cb5 · outbound

This paper cites Slides agent: An intelligent agent for creating and analyzing presentations using large.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Slides agent: An intelligent agent for creating and analyzing presentations using large

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.690919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:bd73b52eb1d395e27097be77e4fb56d942e7b40bab75cc0ef984839bf821f60b

Observation 74429580-9ced-489c-a3c5-6da378344b90 · outbound

This paper cites SlideGen: Collabo- rative multimodal agents for scientific slide generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents SlideGen: Collabo- rative multimodal agents for scientific slide generation

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.407697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:50f7a9491c871bd293fa6a0a2b772d7ee8d8185cd1f99861c3093570b6781d8b

Observation 59329108-f988-4074-b80f-c1983545e8b7 · outbound

This paper cites Presenting a paper is an art: Self-improvement aesthetic agents for academic presentations.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Presenting a paper is an art: Self-improvement aesthetic agents for academic presentations

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.457138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:55dc9a9d4942dc6f9807ebd7462655e64baa36d723898dc3a24851572df6d0e3

Observation cf498312-75c4-42dd-92e9-6958df28398c · outbound

This paper cites Gpt4tools: Teaching large language model to use tools via self-instruction.Advances in Neural Information Processing Systems, 36:71995–72007.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Gpt4tools: Teaching large language model to use tools via self-instruction.Advances in Neural Information Processing Systems, 36:71995–72007

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.727964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:2f967366897a2d2dddeb18ace8b231d8ba3b21639601a4792e129554084e7345

Observation c0c735ad-5fb9-4cee-b0d3-93f8f5adc520 · outbound

This paper cites MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents MM-REACT: Prompting ChatGPT for Multimodal Reasoning and Action

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-14T01:17:58.977376Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:b5934cf497b0c8b15c404c1ceea86bc351e47162539bb48ee634c674bd8aac88

Observation ca5e12fd-76b9-45a0-9e7f-7bd3e36c8ec3 · outbound

This paper cites Os-genesis: Automating gui agent trajectory construction via reverse task synthesis.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Os-genesis: Automating gui agent trajectory construction via reverse task synthesis

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.680217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:2e22b7ef5c0ffca0c75c9c964c2bf97926affc682c3125cc5c699ef1e87185f1

Observation 72a6e0f9-afd5-412c-990e-12fad2d59c2d · outbound

This paper cites VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents VideoGen: A Reference-Guided Latent Diffusion Approach for High Definition Text-to-Video Generation

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.435877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:2161000b55cc541ea3b1ace01bf3953c1bbfc66557da8dd581c89eb015853e2b

Observation e788e6dc-dcf5-4cb9-b020-29fe7d053bc7 · outbound

This paper cites Phyt2v: Llm-guided iterative self- refinement for physics-grounded text-to-video generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Phyt2v: Llm-guided iterative self- refinement for physics-grounded text-to-video generation

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.738177Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:36cbf864b6a628e665c19b129afad96aeedf6ac6ffcbe9bc80c8717083d51bf9

Observation 9a6fdb57-0869-4d03-92ee-a2f67503722a · outbound

This paper cites CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents CogVideoX: Text-to-Video Diffusion Models with An Expert Transformer

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:32:06.444424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:5497033bd418b3e36ffaf8a5187e55aaa3a98d7854d49c740731646570c478a1

Observation 2f886056-ac37-44db-abbc-fb5d28659806 · outbound

This paper cites Unified multimodal understanding and generation models: Advances, challenges, and opportunities.arXiv preprint arXiv:2505.02567.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Unified multimodal understanding and generation models: Advances, challenges, and opportunities.arXiv preprint arXiv:2505.02567

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.447724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:b6e8ec98fd2f766b762a27c7f00f706634e6b647d216604a30d7cedfb4cd3270

Observation 5d738567-e7c6-46ff-91ca-6589dbf51a03 · outbound

This paper cites Qwen3.5-Omni Technical Report.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Qwen3.5-Omni Technical Report

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:32:06.418990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:1fbf4fed05f2e96f6286b3b206447b3dd8f63965682eef069ec127f8d7954a9f

Observation f6643a88-8a2c-42d7-8a54-82467eabc718 · outbound

This paper cites Motion Anything: Any to Motion Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Motion Anything: Any to Motion Generation

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.423233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:af6eb645818fd705a1735f6337e74c9082285a5153397a5d3cba81656a5122fd

Observation 497a87b8-55f1-42c9-9016-e07f2be068d3 · outbound

This paper cites InfiniMotion: Mamba Boosts Memory in Transformer for Arbitrary Long Motion Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents InfiniMotion: Mamba Boosts Memory in Transformer for Arbitrary Long Motion Generation

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.430047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:fa02f43e1fab08bc0d58d00e162a2129924b68e35dba196406d1527b36fe1e79

Observation 192b69c6-c65c-4666-9e76-c46796326fb2 · outbound

This paper cites KMM: Key Frame Mask Mamba for Extended Motion Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents KMM: Key Frame Mask Mamba for Extended Motion Generation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.463488Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:b6090490ac7de9382c1e7ac2775e35371240a7c9e798bf27276364ad8a285478

Observation 39adb5c7-7403-4651-b8b3-72cef9de1c0f · outbound

This paper cites Motion mamba: Efficient and long sequence motion generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Motion mamba: Efficient and long sequence motion generation

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.685930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:2c0fa5bdfa90b1b345653c196b9de6be47216181ce12751315c1fca9eb7a2c26

Observation 629962e1-8b72-465e-ae00-3cdf44883a88 · outbound

This paper cites Evaluating object hallucination in large vision-language models.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Evaluating object hallucination in large vision-language models

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.699752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:7734d4902cee924bd4a335ec0bd22f03ed854da6c4af63605060bfc5ef7229ce

Observation 4bcdc1dd-94f4-4e24-ac6f-01404c611947 · outbound

This paper cites Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data.Nature Communications, 16(1):7866.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Towards generalist foundation model for radiology by leveraging web-scale 2d&3d medical data.Nature Communications, 16(1):7866

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.720574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:2ba33024c9946dee25775e96f5320bd23e240b3af06cba36d89934801985c00d

Observation 7479da98-217e-44fa-a73a-b5e9b42c56b1 · outbound

This paper cites Mavis: A multi-agent framework for long-sequence video storytelling.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Mavis: A multi-agent framework for long-sequence video storytelling

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.748374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:ccb8788152906e6c1e616cb268ccb46640e9f7b5a3510e3cac720de887f480cb

Observation 073d4230-c1cc-4e16-9ade-febcb4b5aecb · outbound

This paper cites Multimodal content alignment with llm for visual presentation of papers.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Multimodal content alignment with llm for visual presentation of papers

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.673644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:77c15d53d4b544c18d1ab69b69459d66291848a999a370cb0a5db52003227247

Observation beb96f69-932c-425c-b445-7a1bdeb29323 · outbound

This paper cites PreGenie: An Agentic Framework for High-quality Visual Presentation Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents PreGenie: An Agentic Framework for High-quality Visual Presentation Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.453852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:487e4e62f7aa0865f1b9d487aebd5709fca74d29abc658b4362591fe56083be6

Observation da205164-c5ba-401f-a8c7-a3238b697d56 · outbound

This paper cites Present- coach: Dual-agent presentation coaching through exemplars and interactive feedback.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Present- coach: Dual-agent presentation coaching through exemplars and interactive feedback

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.460526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:c0d89077f4cf12fac19923a973ab935313f610878891b7f2eaeead725be073c8

Observation 9ffebeec-3f56-426a-8cb1-7dddfbadebf8 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Emerging Properties in Unified Multimodal Pretraining

Reference 35

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:32:06.426637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:0ede97d170c00a0c34697cd805def8fc55782afce90ccc0bc9eebfd5f6235c9d

Observation 8c6ce789-a5c0-4e4a-889c-b4dc7ac5ed4b · outbound

This paper cites Show-o: One Single Transformer to Unify Multimodal Understanding and Generation.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Show-o: One Single Transformer to Unify Multimodal Understanding and Generation

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-05-13T02:32:06.438819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:c218babaca02a35519c87e559a1d6cdf90b58fa127c22b74e573bbe82c89e8d7

Observation 248d91e6-138a-4823-9b10-dbb823043333 · outbound

This paper cites Showui: One vision-language-action model for gui visual agent.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Showui: One vision-language-action model for gui visual agent

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.743426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:4034db2f083dfdae6314ba7f61d2075b1ca044e5c87c5aeff407e02a013d8e5d

Observation eb0ae429-ec71-4d5e-aea3-5c5c84b69a8e · outbound

This paper cites VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents VideoDirectorGPT: Consistent Multi-scene Video Generation via LLM-Guided Planning

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.450743Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:8fcb34a77fecf036ff80cee7bf7bd456d06fb36c8bb5619b0451ff7478ce2e6a

Observation c1cd4cc8-e650-4584-975c-21b4dfa43538 · outbound

This paper cites Videostudio: Generating consistent-content and multi-scene videos.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents Videostudio: Generating consistent-content and multi-scene videos

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-13T13:02:49.668443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:2321cf909d2d0685449514eb920125ab37fa17bcc43af3c9adf64215b3830531

Observation 4de520cb-cae8-439d-bf54-36ae52f0b2c4 · outbound

This paper cites LLM-grounded Video Diffusion Models.

PresentAgent-2: Towards Generalist Multimodal Presentation Agents LLM-grounded Video Diffusion Models

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:32:06.411311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T02:29:42.157339Z digest=sha256:1cf164478c806e25ea67ab39f19e7d3f96cbf3b677e6e7f693f41c575c4391a5

Pith citing papers

Observation 06151ab6-5875-475d-8f3d-03475fa27c04 · inbound

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog cites this paper.

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog PresentAgent-2: Towards Generalist Multimodal Presentation Agents

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T08:44:26.876152Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T08:44:26.876152Z digest=sha256:8f3dde443dea44bf1ec244a9dcfbbc774c240a6b7497b1162820890d5ec52fa3