Pith. sign in

Paper Citation Record · LEDGER

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception

As of 23 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 0 inbound Pith citation observations for arXiv:2411.13314.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.13314 v4

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T16:37:25.405095Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

34 of 34 outbound references displayed

  • verified exact0
  • verified fuzzy23
  • unresolved11
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dc4ebf00-4150-4aec-87fe-542f67aadc66 · outbound

This paper cites Tacotron: Towards End-to-End Speech Synthesis.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Tacotron: Towards End-to-End Speech Synthesis

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.246134Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.246134Z digest=sha256:37a10c7fa058942f75155fc42579a701642e936ae962676ce3391f2489674f0e

Observation eb9b04be-2296-416d-b495-5f8221359ae1 · outbound

This paper cites Fastspeech: Fast, robust and controllable text to speech,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Fastspeech: Fast, robust and controllable text to speech,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.947607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.252316Z digest=sha256:1dbbb317e4674f553c70646dbd9ae5c27164b48fbde192e2e943900eafc5a1d0

Observation 7c8fb9a9-9428-4907-8a89-668d732f99d2 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.932600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.257311Z digest=sha256:282d2ee948586aada93112defb2dbc803be8913b9ca316ee4a2383368b671736

Observation fe39d2f6-fcc0-44bd-954b-b91685061290 · outbound

This paper cites Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Neural Codec Language Models are Zero-Shot Text to Speech Synthesizers

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.262204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.262204Z digest=sha256:96dde8dabf4741896953eb9bab47d3743ba18d3c7dc6cb3babc266fa59a997a7

Observation dee475ef-774e-4f1e-b786-40bf56f2d4f4 · outbound

This paper cites Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.917378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.268301Z digest=sha256:d512a2208b104470bb59b6be415d32011bb4d1774596dfdc3e5f59abe9524df8

Observation 169871b4-4acf-47c2-a6b8-b227e7d0150f · outbound

This paper cites Prompttts: Controllable text-to-speech with text descriptions,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Prompttts: Controllable text-to-speech with text descriptions,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.902242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.273408Z digest=sha256:96bb5d3bc76781f7665ab745fbd87b7c7cdf7691656d58a897aea1fdf6b6c7b4

Observation 9a51dbef-f8e1-434d-b4d1-17e8bc538ca2 · outbound

This paper cites Instructtts: Mod- elling expressive tts in discrete latent space with natural language style prompt,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Instructtts: Mod- elling expressive tts in discrete latent space with natural language style prompt,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.887551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.278929Z digest=sha256:5059bee56e6b6d736c19da338b359d3cdb0e30b31e5db9c9a6c0883329f6aec8

Observation aca669cf-095d-49cd-95a4-46bece36ba6c · outbound

This paper cites Mm-tts: Multi-modal prompt based style transfer for expres- sive text-to-speech synthesis,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Mm-tts: Multi-modal prompt based style transfer for expres- sive text-to-speech synthesis,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.872560Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.283736Z digest=sha256:f30626d10f0e0bfe8c97668acbbe7ba61487b3312d8e3190fa64583b04f3eb48

Observation 75601def-1764-431c-8831-9d72931c4282 · outbound

This paper cites Comparison of different impulse response measurement techniques,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Comparison of different impulse response measurement techniques,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.857455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.288410Z digest=sha256:e751c3c64bf639283bdee256c56ba67a395a889fac7a9dccfa1c22ae27be22da

Observation 1bb2be7d-8c94-4984-b95e-3187ed4fd516 · outbound

This paper cites Environment Aware Text-to-Speech Synthesis.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Environment Aware Text-to-Speech Synthesis

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.293133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.293133Z digest=sha256:8832a24b002b09dd6c7074ef09b20a8c59090fb36d5b463653cd8741e08c14bf

Observation 637f170c-7109-4377-82ee-998ae92e25df · outbound

This paper cites V oiceldm: Text-to-speech with environmental context,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception V oiceldm: Text-to-speech with environmental context,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.842304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.298076Z digest=sha256:d23fb6e2e2a2e01a9ad4546bfab53d909c2a6b67ecfcb39728c9420b2d2a69f9

Observation 05883d7e-61ef-4aca-80db-5b1cf7666636 · outbound

This paper cites ViT-TTS: Visual Text-to-Speech with Scalable Diffusion Transformer.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception ViT-TTS: Visual Text-to-Speech with Scalable Diffusion Transformer

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.302545Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.302545Z digest=sha256:6efcbed6fdac33c3bd6369e68954d0c2c7fb305ffd42b921d3a181c78d75beb4

Observation 75755e0f-adfb-430e-b53f-29d599cada24 · outbound

This paper cites Multi-source spatial knowledge understanding for immersive visual text-to-speech,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Multi-source spatial knowledge understanding for immersive visual text-to-speech,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.827568Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.307222Z digest=sha256:41c2bfb3ee2dc803fdec5cdf3ded99db1004b4c6958b8960e8ed9b08b1690c8f

Observation 5830bfac-aef0-4a35-acc2-faef4795b83a · outbound

This paper cites Learning transferable visual models from natural language supervision,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Learning transferable visual models from natural language supervision,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.812344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.311684Z digest=sha256:7ac6f988f3fa71970a1991f5f67d98d55ef33b940a0505e1da686958e370066b

Observation 26eab282-68aa-40dd-bb41-bbb545d206aa · outbound

This paper cites Allen, M.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Allen, M

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.797123Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.316049Z digest=sha256:45e78511610dccaf61b1bed13c53a23e367a21c9c00bbff673d7834e2a4d2271

Observation 40f962a1-9329-4381-ae5a-14b7e13fb4bc · outbound

This paper cites The festival speech synthesis system,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception The festival speech synthesis system,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.782323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.320515Z digest=sha256:6fcbc2c522320bfd411d97a471fc7f8584866f0c57b10ff457e4e85197205e91

Observation 4296abba-49b2-4dc7-9102-fc2e7cfde937 · outbound

This paper cites An hmm-based speech synthesis system applied to english,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception An hmm-based speech synthesis system applied to english,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.767141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.325063Z digest=sha256:87e26aefe9aa8a53e982d5670d2a2235d27cf2c936bfb08bd4e4f338fef1fa39

Observation 7d176ddb-9b26-4ff6-935c-4a9d5e2baec5 · outbound

This paper cites FastSpeech 2: Fast and High-Quality End-to-End Text to Speech.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception FastSpeech 2: Fast and High-Quality End-to-End Text to Speech

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.329513Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.329513Z digest=sha256:7b24d1d16664264a0482a039c404353cedb157426ed45d5d1fd25a36e70e27d5

Observation 69b0d3b4-99c5-4976-9e38-e90db2844ca6 · outbound

This paper cites Diff-TTS: A Denoising Diffusion Model for Text-to-Speech.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Diff-TTS: A Denoising Diffusion Model for Text-to-Speech

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.334264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.334264Z digest=sha256:6c9a821256cc1c3d2a00fd94650bd93f9d0edbdb9c485062f18fc1cafbe37c41

Observation c29923ea-f3d5-4bc9-9ff8-efa8f60177a6 · outbound

This paper cites Prodiff: Progressive fast diffusion model for high-quality text-to-speech,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Prodiff: Progressive fast diffusion model for high-quality text-to-speech,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.751216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.339265Z digest=sha256:606a77b39cbb7c7a076e9fe7289fb152d7e25dee8192ca57f541244c499026e6

Observation 04fff5d5-02ef-4d4a-8020-d429f7929f39 · outbound

This paper cites Clam-tts: Improving neural codec language model for zero-shot text-to-speech,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Clam-tts: Improving neural codec language model for zero-shot text-to-speech,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.733999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.343956Z digest=sha256:e437557c3acdefaf4f7023f1c578c97f4fff99156f61f7d4f3f79983309638f0

Observation b67974d1-4296-4cfe-b0c3-1bff148f1c4e · outbound

This paper cites Audiolm: a language modeling approach to audio generation,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Audiolm: a language modeling approach to audio generation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.718300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.348479Z digest=sha256:7b8b018c6fdd5165a784ddb8f9152ff84ebcb3edfb03d001fc867b8f3dbd04b4

Observation 8685cbbd-04a9-4cf4-8f11-372eed8bc125 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.353064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.353064Z digest=sha256:7d789a1346b3830bfee5e528a4a48ab2dad05177aedf35286670a062b20326db

Observation 7d6ad148-2fe3-4a92-b82b-b538d5dbd9c2 · outbound

This paper cites Image2reverb: Cross-modal reverb impulse response synthesis,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Image2reverb: Cross-modal reverb impulse response synthesis,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.702539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.357945Z digest=sha256:4b43752fbea6f62d8523de6f49bd78b112103b171b73586cd9ec2550e9d6403b

Observation 569aa7d9-a979-4b90-bf14-a7dfb4fcf717 · outbound

This paper cites Visual acoustic match- ing,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Visual acoustic match- ing,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.685836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.362533Z digest=sha256:c44f9d34b427007252b998aef38a14f49680d1ebb3d0e7eaa4c8fcd597398227

Observation e95df2cb-c3c2-44db-b18b-71e25a8d0d24 · outbound

This paper cites Self-supervised visual acoustic matching,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Self-supervised visual acoustic matching,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.670026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.367038Z digest=sha256:3449a54e66d398415e5b7d8f490e59bfc38d3f96abf5c9150900bf71b9c391bc

Observation e4a73fb9-314e-453b-a55e-b1128a46360d · outbound

This paper cites Meta-stylespeech: Multi- speaker adaptive text-to-speech generation,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Meta-stylespeech: Multi- speaker adaptive text-to-speech generation,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.371754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.371754Z digest=sha256:aae6ab338d997608000514cb6b05c6fd12788257d86df3980e7b43ef62be5e2b

Observation 57329459-da0e-4a55-9d58-18eb0eea6e71 · outbound

This paper cites Attention is all you need,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Attention is all you need,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.376435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.376435Z digest=sha256:94d1a95bb6d0c1569b39454499b63846a90d5f9894ebc2a4214127231c8e7a5a

Observation 71fe08eb-bd37-4259-a8b9-16a3fc517408 · outbound

This paper cites The lj speech dataset,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception The lj speech dataset,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.381217Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.381217Z digest=sha256:256d51b75de5766f733ec6195e5d7a1c4793a42d8f7711eb168ad9b5e0bf1853

Observation 96254a7e-cff5-49b0-b8f6-af0bdff4fb4c · outbound

This paper cites Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.623776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.386009Z digest=sha256:c578d62c00c07dd4911bf524e74cebf464f73e84e82dc782f07b1d806db00e2b

Observation 43e85903-b0a8-459b-b8c9-f919c15bc58b · outbound

This paper cites Mel-cepstral distance measure for objective speech quality assessment,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Mel-cepstral distance measure for objective speech quality assessment,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.607783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.390762Z digest=sha256:a2617dc4dd0d18959d3f4cb40fe1d52db25b795f241ad71a027b6ec8c2b39d55

Observation 46dacfa0-6931-4e79-81dd-31c6803a785f · outbound

This paper cites SC-GlowTTS: an Efficient Zero-Shot Multi-Speaker Text-To-Speech Model.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception SC-GlowTTS: an Efficient Zero-Shot Multi-Speaker Text-To-Speech Model

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-12T16:37:25.395526Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T16:37:25.395526Z digest=sha256:4bd46d9b1b766e0a1739804004685917e79733bb86eba2d143be80f5d66064a8

Observation 4883fca5-8629-48f6-ade2-2c9f514ef95d · outbound

This paper cites An overview of voice con- version and its challenges: From statistical modeling to deep learning,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception An overview of voice con- version and its challenges: From statistical modeling to deep learning,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.591490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.400582Z digest=sha256:af5754426b8c3686b1a0916d2f5a3a11df6ab941fd287d49b52c6abb107dcb23

Observation da33eb71-cba3-4e7f-a4dd-3596581e09f8 · outbound

This paper cites Grad-cam: Visual explanations from deep networks via gradient-based localization,.

I2TTS: Image-indicated Immersive Text-to-speech Synthesis with Spatial Perception Grad-cam: Visual explanations from deep networks via gradient-based localization,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T16:37:25.574200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-12T16:37:25.405095Z digest=sha256:7dd7127ea19bb360d2b44b6cfdc0fd1a2adff25f7d1535f7a2aa2bd46672e180

Pith citing papers

No inbound Pith citation observations are available.