Pith. sign in

Paper Citation Record · LEDGER

A Large Vision-Language Model based Environment Perception System for Visually Impaired People

As of 22 August 2026, this Paper Citation Record lists 57 of 57 outbound references and 0 inbound Pith citation observations for arXiv:2504.18027.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.18027 v1

Coverage vector

measured 57 of 57 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:31:00.288054Z

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

57 of 57 outbound references displayed

  • verified exact0
  • verified fuzzy40
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5238b3ec-0889-4319-9ec4-254728d96961 · outbound

This paper cites Ingold, The perception of the environment: Essays in livelihood, dwelling and skill.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Ingold, The perception of the environment: Essays in livelihood, dwelling and skill

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:01.025469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.070592Z digest=sha256:e05bf0635a277245bc335c66c180710e631f4f9146538a1cc0addff1e3687473

Observation ca42d922-1ed1-460b-bb85-b2b5985e29fc · outbound

This paper cites Enabling independent navigation for visually impaired people through a wearable vision-based feedback system,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Enabling independent navigation for visually impaired people through a wearable vision-based feedback system,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:01.012861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.075930Z digest=sha256:72e3a9ba2b54b8fc456cf1214ef08fd87dedce4e0d9004457e3f30749fae8da9

Observation 81703b36-5f97-486c-8024-ba29eec7ec0c · outbound

This paper cites Magnitude, temporal trends, and projections of the global prevalence of blindness and distance and near vision impairment: a systematic review and meta-analysis,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Magnitude, temporal trends, and projections of the global prevalence of blindness and distance and near vision impairment: a systematic review and meta-analysis,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:01.000856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.080180Z digest=sha256:b82e8e3866d5a208a0563653d8c696a02540ea7c349f3df94e36afe221457635

Observation 84177a55-74ac-4142-a187-32e566d797ca · outbound

This paper cites Objectively measured visual impairment and dementia prevalence in older adults in the us,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Objectively measured visual impairment and dementia prevalence in older adults in the us,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.989233Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.083756Z digest=sha256:bf0cdebc2e01895f74e546b86c472d4e1fd5906704a814439883a794d7d3b132

Observation 4239bd43-dd8e-4050-8236-e43af278c429 · outbound

This paper cites Virtual-blind-road following-based wearable navigation device for blind people,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Virtual-blind-road following-based wearable navigation device for blind people,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.976446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.087535Z digest=sha256:d969fbf2b23bf6ce0da27a81f062d5d3e0fef465541c9904662f0f47df975620

Observation 53098e2a-d256-4c56-90fd-bc8e9a837140 · outbound

This paper cites Embedded reading device for blind people: A user-centered design,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Embedded reading device for blind people: A user-centered design,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.962539Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.091100Z digest=sha256:3583382c193fe40256cb013cdccf3e88e96b2e3dba71291f39472ab070ad8fb8

Observation 0e357f4a-6a50-41a8-ab36-4d3cf8eb07d8 · outbound

This paper cites Hand-priming in object localization for assistive egocentric vision,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Hand-priming in object localization for assistive egocentric vision,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.949264Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.095164Z digest=sha256:7ff228a825259989bf59981efcd295967e72e77b8fb917f393d9d17e680b0577

Observation 7586210f-3cae-425b-a8ab-3aad3216f4e5 · outbound

This paper cites Vizwiz-fewshot: Locating objects in images taken by people with visual impairments,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Vizwiz-fewshot: Locating objects in images taken by people with visual impairments,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.936168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.098659Z digest=sha256:527bd4d7f5ea68b253e119187a101b3fd2ab891599aef918585024d208482a9e

Observation b8897f6e-bdb0-4ec7-922a-ecd80845d349 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.102041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.102041Z digest=sha256:91f4230719ee92b46bdd1df512814fc8daa8dada3b4328eb1c8cd316887107fc

Observation 8362b807-6bb6-49bb-a64b-4545ea3c38fa · outbound

This paper cites Incorporating External Knowledge into Machine Reading for Generative Question Answering.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Incorporating External Knowledge into Machine Reading for Generative Question Answering

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.105795Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.105795Z digest=sha256:111bf5ee606e3b1aff264f7d767995d0813b9b0050eb06c00decb674324a741d

Observation 7793f1ca-80db-4d2b-bdc9-aacadb2ff268 · outbound

This paper cites Check Your Facts and Try Again: Improving Large Language Models with External Knowledge and Automated Feedback.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Check Your Facts and Try Again: Improving Large Language Models with External Knowledge and Automated Feedback

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.109472Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.109472Z digest=sha256:a713c80216d9c892f136b7192ee2d0218d981555d28ccb94309bfcc08eb9bd00

Observation 6c32a6fa-7e82-4595-9138-59c04d5d0d6e · outbound

This paper cites Smart guiding glasses for visually impaired people in indoor environment,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Smart guiding glasses for visually impaired people in indoor environment,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.922921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.113421Z digest=sha256:ac02405024d33583aef55c5952ed9e011b00927c830054a7541ea75bbecf4637

Observation 476705bc-afe4-4c72-85e4-c95e66a421d9 · outbound

This paper cites An enhanced obstacle avoidance method for the visually impaired using deformable grid,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People An enhanced obstacle avoidance method for the visually impaired using deformable grid,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.910049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.117676Z digest=sha256:843251681e4cc430083af1624ff0c298b0730c37666a7c5189e10e4051407895

Observation 4131347c-bc28-441d-8b39-f54f3886c515 · outbound

This paper cites Collision detection method using image segmentation for the visually impaired,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Collision detection method using image segmentation for the visually impaired,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.897088Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.121663Z digest=sha256:a709568ab55fabfdd82b7295304c489b2cc04b72117a414b7284cba81b371d1e

Observation 2f7d754c-8d43-45c7-aac0-9c60342808ac · outbound

This paper cites Tdraw: A computer-based tactile drawing tool for blind people,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Tdraw: A computer-based tactile drawing tool for blind people,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.884312Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.125470Z digest=sha256:a02f59da149c28ccaa815fedc155fba51ae569c5e2bcce249868175a40d7e817

Observation 6c449d9b-0b3f-4423-bd65-fe56bd35dabe · outbound

This paper cites Usable gestures for blind people: Understanding preference and performance,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Usable gestures for blind people: Understanding preference and performance,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.872328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.129243Z digest=sha256:c485d087487f00276f6147d43bc0728010c72954b29da440fb7075836c9137eb

Observation a716678e-67de-4f5c-84de-4918c397b659 · outbound

This paper cites How teens with visual impairments take, edit, and share photos on social media,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People How teens with visual impairments take, edit, and share photos on social media,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.860378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.133075Z digest=sha256:2ef1f41cfd6e07ff6f33cabf28752f07bf37a3f33b4e670df51cba78a1e95168

Observation 9f894f49-d49e-4c92-9fbd-8845beb8ba1e · outbound

This paper cites Flight: A low-cost reading and writing system for economically less-privileged visually-impaired people exploiting ink-based braille system,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Flight: A low-cost reading and writing system for economically less-privileged visually-impaired people exploiting ink-based braille system,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.847723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.136838Z digest=sha256:908ad6bafc29cbbeb4cabac1c1565bdf8f5f4d344316d36d36589841404cf019

Observation 5f6859ed-659e-41b2-af1e-c7fe1ff43ee5 · outbound

This paper cites Linespace: A sensemaking platform for the blind,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Linespace: A sensemaking platform for the blind,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.834865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.140734Z digest=sha256:d795470d549c8fe38925752f7de57a70e422d3c2ad549d1b800dd1c5bab3a642

Observation 8ab3b6b7-2fae-428c-94a4-91c5f5189655 · outbound

This paper cites Tangible reels: Construction and exploration of tangible maps by visually impaired users,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Tangible reels: Construction and exploration of tangible maps by visually impaired users,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.821638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.144554Z digest=sha256:382af0974d7a0d12b4ff694932c6bf1e0c284c624fe5f2b95f3cc8926a5f8e17

Observation 2745375e-50b7-476d-bfe5-2047884835fc · outbound

This paper cites Comparing computer- based drawing methods for blind people with real-time tactile feed- back,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Comparing computer- based drawing methods for blind people with real-time tactile feed- back,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.808129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.148569Z digest=sha256:bd4985fa4a31613dfc1fbca75ba8a4ef54d5aa903cdda8edb4af256c89b2830b

Observation ee460e25-8e79-460d-9ac2-3ffa77a00f2f · outbound

This paper cites Accessible maps for the blind: comparing 3d printed models with tactile graphics,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Accessible maps for the blind: comparing 3d printed models with tactile graphics,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.795103Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.152475Z digest=sha256:2d93837d1c79c90ed1f9a8d95d28acf45f4751368c64376d20c4b2ace419d5ea

Observation 4f9b320c-d7ba-45bf-b420-4a433b9245b0 · outbound

This paper cites Taking into account sensory knowledge: The case of geo-techologies for children with visual impairments,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Taking into account sensory knowledge: The case of geo-techologies for children with visual impairments,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.782295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.156325Z digest=sha256:3b406cc9fe29934a61daf73bd2c86c59ba8867c61cd781a98f07e1911eee9193

Observation 2c4c0c9f-7df3-4901-bf1b-846fbdbee5cc · outbound

This paper cites People with visual impairment training personal object recognizers: feasibility and challenges,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People People with visual impairment training personal object recognizers: feasibility and challenges,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.769894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.160306Z digest=sha256:a8663579d1a2a949b943230ed561213b67e9666d5d4a38de26dde98dee9cb939

Observation 2d119cea-d298-4215-973f-999cd590083f · outbound

This paper cites Synthesizing stroke gestures across user populations: A case for users with visual im- pairments,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Synthesizing stroke gestures across user populations: A case for users with visual im- pairments,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.756993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.164179Z digest=sha256:d68fc3cdd3d58dff90fb6355155ca94b2fc76248ae4aa36aad0f27b046e58fe6

Observation 5a758cf3-9155-4672-9e47-5c8ec60709dd · outbound

This paper cites A face recognition application for people with visual impairments: Understanding use beyond the lab,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People A face recognition application for people with visual impairments: Understanding use beyond the lab,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.744620Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.167995Z digest=sha256:03be841470b1f43915510b1289405d50a7a3aea69fc429454b8c509fa5022ae9

Observation 96ec5311-e479-40ed-ba47-8d60998fa60b · outbound

This paper cites A multitask grocery assist system for the visually impaired: Smart glasses, gloves, and shopping carts provide auditory and tactile feedback,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People A multitask grocery assist system for the visually impaired: Smart glasses, gloves, and shopping carts provide auditory and tactile feedback,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.731875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.171921Z digest=sha256:21e7b87846fb358325d4f0b779e961a1c65153c24aaa4ddab22739106a9aabc9

Observation 6b9f7371-ff12-4baa-af1e-4f1a343d138a · outbound

This paper cites Hindsight: Enhancing spa- tial awareness by sonifying detected objects in real-time 360-degree video,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Hindsight: Enhancing spa- tial awareness by sonifying detected objects in real-time 360-degree video,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.719664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.175848Z digest=sha256:d135efb6caf478053f8918bb501f42ed63ddc2bb6a7233ea94250e7d196560a4

Observation 66c24848-801b-43a9-bba2-e0ddc698e84d · outbound

This paper cites You only look once: Unified, real-time object detection,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People You only look once: Unified, real-time object detection,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.706770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.180050Z digest=sha256:27dc371e9b762c5abef0ad2853c88381a268540ff9700243372c5d4cca047c80

Observation 782940ce-8c74-4448-bfb2-673331de9336 · outbound

This paper cites Ssd: Single shot multibox detector,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Ssd: Single shot multibox detector,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.692788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.183925Z digest=sha256:f17af456e606f8b49448dd8a5d3284591fa53889b06ec8dd61b4e08ae00f7bc1

Observation 2ee8d73a-286f-4d23-b07d-2e3cb947d673 · outbound

This paper cites Rich feature hierarchies for accurate object detection and semantic segmentation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Rich feature hierarchies for accurate object detection and semantic segmentation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.679459Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.187686Z digest=sha256:69b8f6ab2c9e4bf84a9e797336903436b0de6efccb8e60a53a372e86beea45ae

Observation db88033f-d8b6-4980-86b5-a4533595940e · outbound

This paper cites Faster r-cnn: Towards real- time object detection with region proposal networks,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Faster r-cnn: Towards real- time object detection with region proposal networks,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.665125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.191599Z digest=sha256:5eb3cd519aad3cbf328c2104ca5efa8c59204bad25bb79edb0875313343ad5d2

Observation a194552c-df25-496f-8c23-442091595cf6 · outbound

This paper cites Fully convolutional networks for semantic segmentation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Fully convolutional networks for semantic segmentation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.652065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.195421Z digest=sha256:bf99161fdf95b06d588442602ead126618a455ae0b3c6821768df5550baeadb8

Observation cd9f58ff-1615-4cc4-a81e-1e7964239d26 · outbound

This paper cites Segnet: A deep convolutional encoder-decoder architecture for scene segmentation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Segnet: A deep convolutional encoder-decoder architecture for scene segmentation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.640256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.199320Z digest=sha256:75b5539096d6bc7c22cdeb25604d83527629f5874d0a4923bb58d869d72a8491

Observation a944ebca-3c19-4ccc-9772-cbc211c6bffc · outbound

This paper cites Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Deeplab: Semantic image segmentation with deep convolutional nets, atrous convolution, and fully connected crfs,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.628702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.203093Z digest=sha256:c6efd649eaa5f2446e3bd7e56c53555647e55071e42b853c2af05640a11a5916

Observation 9150d598-d671-489d-9a03-c988fa2185f2 · outbound

This paper cites Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Rethinking semantic segmentation from a sequence-to-sequence perspective with transformers,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.616863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.206866Z digest=sha256:fa7ce05bc04b00af97ec6e6b0343481fe5294a4e0bd07a72bdb9656e97399eca

Observation f05c32a9-1f5b-4acf-8032-411eef2c3f9c · outbound

This paper cites Swin transformer: Hierarchical vision transformer using shifted windows,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Swin transformer: Hierarchical vision transformer using shifted windows,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.210929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.210929Z digest=sha256:550ed3e7843f8514a288b110d6f662471d8b21400b1bb25042dc4f0c7ce350e5

Observation f399b1a3-0bdb-41c3-b49f-62bc9b3321ba · outbound

This paper cites Attention is all you need,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Attention is all you need,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.214911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.214911Z digest=sha256:84dff5d284d7b4ad0c032bc9dba6f902438fcd3d8adef6ec9e65a7e093cc6874

Observation dc472b25-ca38-4440-bc89-769d0e8f45b2 · outbound

This paper cites Fusenet: In- corporating depth into semantic segmentation via fusion-based cnn architecture,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Fusenet: In- corporating depth into semantic segmentation via fusion-based cnn architecture,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.588242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.218978Z digest=sha256:7e7535cab8ad35bc3815c289077d4e4a67d8d899e1afd800b43a9d45bc04409f

Observation ff2e631d-caab-4573-9575-50ff1d525f6c · outbound

This paper cites RedNet: Residual Encoder-Decoder Network for indoor RGB-D Semantic Segmentation.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People RedNet: Residual Encoder-Decoder Network for indoor RGB-D Semantic Segmentation

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.223131Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.223131Z digest=sha256:8529c05d808bca9147a8553f91e22bf169928372a5d6300c10e3dd1e37a729f6

Observation 740e6dea-7417-46aa-b70a-34ceda40eef0 · outbound

This paper cites Show and tell: A neural image caption generator,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Show and tell: A neural image caption generator,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.575223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.227421Z digest=sha256:00fb96c39f41d706b818953319be3f83c9be9b300c3243c0d7a7437812edffdd

Observation c8b8e99b-6775-45f8-9922-73d45b06afb5 · outbound

This paper cites Knowing when to look: Adaptive attention via a visual sentinel for image captioning,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Knowing when to look: Adaptive attention via a visual sentinel for image captioning,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.561847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.231212Z digest=sha256:dbf012e31492a8a517cfacf8a5623c704b023e37721e2c0566909491ee1aeaee

Observation b953f83a-e732-4730-9bd1-1cee6bfa1a0e · outbound

This paper cites Show, attend and tell: Neural image caption generation with visual attention,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Show, attend and tell: Neural image caption generation with visual attention,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.235162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.235162Z digest=sha256:6859053060509f7cc399013a8b7de10fb88e577b57f53efe80f621dde412183e

Observation cc6bd0fe-de3c-42c2-8f77-dbe17d3ac78a · outbound

This paper cites Dense captioning with joint inference and visual context,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Dense captioning with joint inference and visual context,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.540912Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.239060Z digest=sha256:512cfe8a78ee0079c084c9a931d282e9a85ecf5661afca8ca3a5e0786fb725bd

Observation 2964a558-64f0-4877-926c-80e27e40cccd · outbound

This paper cites Unifying vision-and-language tasks via text generation,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Unifying vision-and-language tasks via text generation,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.243006Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.243006Z digest=sha256:c47019b1f1e33162b394c1f88df8528e330897e492e9a1585fbac9f9e6a1513c

Observation a970cb18-308b-40d8-8047-b79c31476024 · outbound

This paper cites Multimodal few-shot learning with frozen language models,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Multimodal few-shot learning with frozen language models,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.518885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.246908Z digest=sha256:b726aaa4d54dff6373ff3dde26664a23eba850369c5c81e3582acac3be9693d3

Observation defa57c1-c932-44b8-b1d4-1991733ba014 · outbound

This paper cites SimVLM: Simple Visual Language Model Pretraining with Weak Supervision.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People SimVLM: Simple Visual Language Model Pretraining with Weak Supervision

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.250761Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.250761Z digest=sha256:e7f58a23f9657c643a7627b3345d735bb77323dde3440fa0d60eac6a27b4486e

Observation 3618de6a-eedf-4b2d-b03a-ce85522cbad0 · outbound

This paper cites Visual instruction tuning,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Visual instruction tuning,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.254925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.254925Z digest=sha256:0196ea739c5f2f2d796ae56b0124c11b1121e9a6c4581aebfd7ed4fb9899752b

Observation 73657ceb-f69e-4b10-ac8f-332c7648fdbf · outbound

This paper cites BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.258268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.258268Z digest=sha256:48a20b452fa16943ee656cbab78b66ca32f098ddb1c2f7b5aed228a844c375aa

Observation ced62b91-54c3-45b1-9e08-05e7e56ae140 · outbound

This paper cites Flamingo: a visual language model for few-shot learning,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Flamingo: a visual language model for few-shot learning,

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.496661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.261705Z digest=sha256:3b7ba6a9d3fff9acdd84637106df475e5546d28f8e2add86eabe09a3d9643236

Observation f9486fcb-4b5a-4fea-b769-c90404a7b5a3 · outbound

This paper cites Detecting and Preventing Hallucinations in Large Vision Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Detecting and Preventing Hallucinations in Large Vision Language Models

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.264996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.264996Z digest=sha256:b5995c44ecaf571e39c3d4c87931805e1eccfaee1a4d089c07ca2f79b935aa43

Observation 34c754b0-afdb-4e51-8015-78cbd4d6f1ad · outbound

This paper cites Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Mitigating Hallucination in Large Multi-Modal Models via Robust Instruction Tuning

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.268434Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.268434Z digest=sha256:b79408f3d738004fcd04a51f7486302f2058b57057e5c3c0b00376b767d97249

Observation c50934a6-876e-45e6-8396-b10a0d94a28d · outbound

This paper cites Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Evaluation and Enhancement of Semantic Grounding in Large Vision-Language Models

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.272098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.272098Z digest=sha256:bcfb107d298cfa561643cafa565701de91afa4fa4204fabe052c28f173b5fbca

Observation bbbcd020-b0fb-4e83-bc27-793dbb5472e9 · outbound

This paper cites Woodpecker: Hallucination Correction for Multimodal Large Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Woodpecker: Hallucination Correction for Multimodal Large Language Models

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.276120Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.276120Z digest=sha256:cbd8a59458a04e8616c25f8ef60859c709b1d6f630bdf5572af39adabb496e9e

Observation ea7e7403-23df-49dd-ada6-ca04bb3f694c · outbound

This paper cites Qwen-vl: A versatile vision-language model for under- standing, localization, text reading, and beyond,.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Qwen-vl: A versatile vision-language model for under- standing, localization, text reading, and beyond,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T10:31:00.482844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-16T10:31:00.280139Z digest=sha256:aa0e3b0886ac9b652e511d12cd9fd10ef6638dfbf7f0a9a7be446ec71746bdcf

Observation d2610f5d-784e-4134-b3e9-d6f8c114c06b · outbound

This paper cites Evaluating Object Hallucination in Large Vision-Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People Evaluating Object Hallucination in Large Vision-Language Models

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.283944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.283944Z digest=sha256:f7bb3bea129f1ecf6d862c65187eed3c521400a3adcd1d5201382e645eee8e42

Observation 41a2c272-c971-410a-a7a6-7bb40030c138 · outbound

This paper cites MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models.

A Large Vision-Language Model based Environment Perception System for Visually Impaired People MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-16T10:31:00.288054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:31:00.288054Z digest=sha256:df544ee45dd41fe525f6f39114a3a070667c5ba32c9117e94c11ba3071e8b247

Pith citing papers

No inbound Pith citation observations are available.