Pith. sign in

Paper Citation Record · LEDGER

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence

As of 17 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 0 inbound Pith citation observations for arXiv:2608.05816.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.05816 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T23:04:01.886110Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

41 of 41 outbound references displayed

  • verified exact0
  • verified fuzzy32
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6c274851-dcd5-4301-ac50-615b7b1006bc · outbound

This paper cites In: Proceedings of the IEEE international conference on computer vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE international conference on computer vision

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.603025Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.554204Z digest=sha256:aa0433bf9df295523fe130c24a38530bfadc2900e34885dac8f8549af273a889

Observation 1860f9dc-9b7b-4999-8dcb-9d2fc6394f17 · outbound

This paper cites In: Proceedings of the Euro- pean conference on computer vision (ECCV).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the Euro- pean conference on computer vision (ECCV)

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.588893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.559635Z digest=sha256:66d729a08760418fe1a85989fade7eef6aa26c405b99d5e0b1f155f88093f6e9

Observation 37445cdf-c368-4921-907f-b9a63da27640 · outbound

This paper cites Advances in neural information processing systems29(2016).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Advances in neural information processing systems29(2016)

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.572902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.564650Z digest=sha256:0428fee6bd459efb70c1e69118e6eac8bce1016c512546636139bf64821005b5

Observation fe0cb4b8-2b9c-4535-bc18-b60886a618c4 · outbound

This paper cites Frontiers in Neuroscience11, 159 (2017).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Frontiers in Neuroscience11, 159 (2017)

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.558217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.569568Z digest=sha256:5ddb92dfe7a7c16dbd044a409053d67df504d8e64a2b52777eecaf7233eb812b

Observation 2226b01b-410e-4810-8a64-7fb3295ee0d7 · outbound

This paper cites Attention, Perception, & Psychophysics77(5), 1465– 1487 (2015).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Attention, Perception, & Psychophysics77(5), 1465– 1487 (2015)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.543677Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.575305Z digest=sha256:462dcecf3c33a0f99cbfd7b76ec5310bb2c41eb2dd91ea54b5bb245bacffa543

Observation 70bbde47-6c15-44e2-accc-24068ee7d583 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.527389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.580317Z digest=sha256:b2b095e633805b7510e948462171ea83e57a849748e1f3c08902f6c0080a7d32

Observation 04861d27-f7da-46b9-9d0c-60e141688c93 · outbound

This paper cites In: ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: ICASSP 2020-2020 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.510649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.586258Z digest=sha256:b0a2129cc6f6c12a5a1eea693381ae6a3a129f9a35bc5afb2b2a140901ab0ffd

Observation aa60035c-c75c-4be2-815f-8443961b437e · outbound

This paper cites In: Conference on Computer Vision and Pattern Recognition (CVPR) (2020).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Conference on Computer Vision and Pattern Recognition (CVPR) (2020)

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.492991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.591367Z digest=sha256:19a5e428f887bb0010df4137322f26d71afb7108e7bb6fa723fa48dd535675c6

Observation 24df23de-f6f3-4649-96cf-a1d483f84a1b · outbound

This paper cites In: Proceedings of the IEEE/CVF International Confer- ence on Computer Vision (ICCV) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF International Confer- ence on Computer Vision (ICCV) (2025)

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.476703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.596497Z digest=sha256:2f911b1c935623df283e857f4c8c64b55211087ad17d65abecbeb4000bc7bb69

Observation b39f2f56-35e7-4d00-a4a4-535a8ff02df8 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence (AAAI).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the AAAI Conference on Artificial Intelligence (AAAI)

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.460940Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.601464Z digest=sha256:7a851179cadddc942fbb03361d190ef570c2e51e2dffcbea06bbb74b0e34c0ca

Observation 3a424f31-5dac-467d-b69f-93376678c2f1 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) (2025)

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.444767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.613119Z digest=sha256:661de6fcf373887dc42722fdd3590590320d575df1ec4765b1f4f367e6f0d9d8

Observation 8ea478e6-0ac9-4039-9617-0d09ef77e084 · outbound

This paper cites In: Advances in Neural Information Processing Systems (NeurIPS).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Advances in Neural Information Processing Systems (NeurIPS)

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.429207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.628526Z digest=sha256:9e76b11af516d4a4fb393d86cf43c2dd565ec34813b49fc576bbea16670458a7

Observation 210ccabd-7b88-4e58-a277-8c47d2959df1 · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.645386Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.645386Z digest=sha256:4703e81a0c38f73c4cf5e47d4ad9f7f5768b17265f260be5145d3571b4442c24

Observation d322b711-de24-4baa-b2a8-46f8d5ca0d46 · outbound

This paper cites In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.401361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.656480Z digest=sha256:65990567d1284c198926d552494f9c97ef64bdc439e6c912b3fa29722d009d13

Observation a441b1f1-0c6e-4e0c-9643-8db7ab5d6bcf · outbound

This paper cites Advances in Neural Information Processing Systems33, 10077–10087 (2020).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Advances in Neural Information Processing Systems33, 10077–10087 (2020)

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.386110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.667931Z digest=sha256:e1cbc945342c91c4355327fd67d3279c3ffa56148f41631eefaf5e0ba657005e

Observation fa0c01c6-d212-4035-81ba-28fb7fbe33e1 · outbound

This paper cites In: British Machine Vision Conference (BMVC) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: British Machine Vision Conference (BMVC) (2025)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.369268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.679994Z digest=sha256:3b225763cea7be9946581eeabe16c7b79832337329b79d51fb190913c6ecce23

Observation cc5e74c0-e9c9-454e-9394-3862c0ec41f4 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition (2022) Whence the Voice? 33.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pat- tern Recognition (2022) Whence the Voice? 33

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.354037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.692047Z digest=sha256:50f40a1566325c5280cde0b746df4b31ed2ded371199820e8b9d11db5b153c15

Observation a00b0d01-9f57-469b-939c-66eb56d2f339 · outbound

This paper cites In: British Machine Vision Conference (BMVC) (2025).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: British Machine Vision Conference (BMVC) (2025)

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.337579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.704370Z digest=sha256:d2ea214c67337ebbc718cb0e605fade00c39b6bd988c9c26771914f5df3c6db0

Observation 01159ed1-3a04-4182-abd5-d43280ebe315 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.320224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.717929Z digest=sha256:2cad943205d6f3497ddb9f888b97a29ce9094821ae37f86c07fae5f957eb4615

Observation 0a9cdadb-e65e-470b-9873-9f088ea81ed5 · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.303101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.731400Z digest=sha256:1a0a1b8749c1d01e61399612e2839eab7a441870f49cc3452bcecca65b61dd17

Observation c799eb3e-e1a8-4529-98df-28647321f6d7 · outbound

This paper cites Adam: A Method for Stochastic Optimization.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Adam: A Method for Stochastic Optimization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.744967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.744967Z digest=sha256:00b2ff8db12725a0113e916513c6854e5d5d3e4603ecf69659a8b3ade9e2a420

Observation 5eb0effc-344b-4a37-bbe8-7d646a28c510 · outbound

This paper cites In: Proceedings of the IEEE/CVF international conference on computer vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF international conference on computer vision

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.759221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.759221Z digest=sha256:7c59c8be93f5bfa4993c7253b1a26fedbb92396790d8398076ad211100c7e08d

Observation 9a3a2b95-8243-4ef0-b3e9-89a7ae82b892 · outbound

This paper cites In: Proceedings of the AAAI Conference on Artificial Intelligence.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the AAAI Conference on Artificial Intelligence

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.275731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.771411Z digest=sha256:e22c4b41d6ddcf55c1b2ffae67a119b03e28ba69bd74a90f191b43a55b99a934

Observation afe5a8c0-2469-42dc-86ac-082e3f35f50f · outbound

This paper cites In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.260518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.776984Z digest=sha256:96cdae3ecc466ce09e2183b8eb76532f9a877b2e5668d871098930d73055ff46

Observation 6289fb5c-0299-4add-99be-b18439494471 · outbound

This paper cites In: European Confer- ence on Computer Vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: European Confer- ence on Computer Vision

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.243499Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.781720Z digest=sha256:8515b06fc1e968f7939ce2cf1a937c95d6d188bf3b0f9a2e19deb96da33e260c

Observation 058b684d-7c8c-48f1-acb4-3e8b47e2325f · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:04:02.227849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.786501Z digest=sha256:746b72c3e269f2453fe90c64a2c4f6bd27471f2d0ce399d9e9c69e954f209010

Observation 7d9570fc-408d-4b6a-b6aa-a5a106d818b8 · outbound

This paper cites Multi-scale Multi-instance Visual Sound Localization and Segmentation.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Multi-scale Multi-instance Visual Sound Localization and Segmentation

Reference 27

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T23:04:01.982686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.791902Z digest=sha256:dbfcdc65808ff942035f11a8621bbbf169db42424e1ab01af74cdd417abaf9a7

Observation 0b86ad80-4dc9-45c6-b617-66949b7bdde8 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Representation Learning with Contrastive Predictive Coding

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.797039Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.797039Z digest=sha256:00fecc7c0f8cf9d66def90764cd5cc98c61aad4f0cbaa73137292ba844850ec5

Observation f8669019-ec0f-4286-b572-2f2f0744da26 · outbound

This paper cites In: Proceedings of the European conference on computer vision (ECCV).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the European conference on computer vision (ECCV)

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.213675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.803709Z digest=sha256:2909ee3094cfb8dfe3ad8471975fce7a04a7fbd0822a2b073cc23c8c06bf4e13

Observation 872970d9-7b66-42c4-8e95-ee3866be1439 · outbound

This paper cites In: European conference on computer vi- sion.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: European conference on computer vi- sion

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.198824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.810853Z digest=sha256:4a0cc82b1c865d8bb5934d1c5efb4792f9b4ee93bab050a63fb57e942882f131

Observation ab100932-b508-46dd-be9d-d176605f76d5 · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:04:02.184569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.815816Z digest=sha256:fd614ff8736f513c85cbb356f4cdda73415b014bf89d14e12a0ef597910043f6

Observation 2965b4f1-3b10-43fa-87b4-e4afc5d408c5 · outbound

This paper cites In: European Conference on Computer Vision.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: European Conference on Computer Vision

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.169557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.823082Z digest=sha256:64057d72b1b90da02ae20eaa948a1205341a854d19585d57aca61d6667596a8b

Observation d95d7cd5-6837-4960-9822-0fff0adb090a · outbound

This paper cites In: Proceedings of the IEEE conference on computer vision and pattern recognition.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the IEEE conference on computer vision and pattern recognition

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.154225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.828317Z digest=sha256:702ab9a3d7075831a5be609b05958a5377fe5c2f29437430738c84da8eea7d3d

Observation 56b0f1df-d741-4d6d-bcb7-4044eb785e8f · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence42(8), 1984–1997 (2020).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence IEEE Transactions on Pattern Analysis and Machine Intelligence42(8), 1984–1997 (2020)

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.137544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.834636Z digest=sha256:6fb8f80cd57bb2d0929aa95ee0d10528b82085baaafcb2e910a8df75d3a518aa

Observation 035a9a13-d765-4d3f-976b-538cc424f664 · outbound

This paper cites Hu et al.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Hu et al

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.120779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.847282Z digest=sha256:1947dd9b8b88798d7647f4a7b3ba9fbda0c90c1e6f9b4f5c7b2c5f1d1a1ec8f8

Observation 9e0ca096-c3ef-4693-9934-eccc92b981c8 · outbound

This paper cites Journal of Speech, Language, and Hearing Research60(10), 2989–3000 (2017).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Journal of Speech, Language, and Hearing Research60(10), 2989–3000 (2017)

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.102936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.854989Z digest=sha256:3919b10b12ecec85659ac63991cb9bde77c84bb951a088dd2e487d04d14f9030

Observation 4395275b-730b-4d7f-90e9-fa21db9ffef9 · outbound

This paper cites In: Proceedings of the Computer Vision and Pattern Recognition Conference.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the Computer Vision and Pattern Recognition Conference

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.086380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.860761Z digest=sha256:4de3a88a7731ba42ff84d4cf184801c392e9fa62b27a083e4b1ad269b12a4781

Observation 0e8bbfe5-f322-4844-9222-d57d9560d1b3 · outbound

This paper cites In: International Conference on Machine Learning (2023).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: International Conference on Machine Learning (2023)

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.068961Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.866450Z digest=sha256:cdc7916852a3b429dc60ea449abb349923775c0ca709136984ef314a695be4dd

Observation 1af77587-3acd-4773-83bf-f867cd91f2cc · outbound

This paper cites an unresolved cited work.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T23:04:02.050943Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.873644Z digest=sha256:7c7b76394311f548d7841d05d2449814907622b02d55a4676f9351cdcda31c33

Observation 027d935b-c4ad-4bfa-976f-f0b81b9b0548 · outbound

This paper cites Audio-Visual Segmentation with Semantics.

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence Audio-Visual Segmentation with Semantics

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T23:04:01.879461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T23:04:01.879461Z digest=sha256:3367a9c88c81af8476b9708678820b6b02f1879907dedba14aaaf4397d9ce1fd

Observation 70c20b9a-212b-49ef-89d3-3efa484953b0 · outbound

This paper cites In: Proceedings of the European Conference on Computer Vision (ECCV) (2022).

Whence the Voice? Self-supervised Dual-source Audio-Visual Localisation via Selective Convergence In: Proceedings of the European Conference on Computer Vision (ECCV) (2022)

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T23:04:02.027050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T23:04:01.886110Z digest=sha256:13711f9d6ac22a26ed056d48ac8da4d58df2979609341497f6fd9536ced5868b

Pith citing papers

No inbound Pith citation observations are available.