Pith. sign in

REVIEW 2 cited by

An Efficient Modern Baseline for FloodNet VQA

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2205.15025 v1 pith:636GOQ5V submitted 2022-05-30 cs.CV cs.AIcs.CL

classification cs.CVcs.AIcs.CL
keywords efficientfloodnetmodernmethodsperformancesystemsystemsabstraction
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Designing efficient and reliable VQA systems remains a challenging problem, more so in the case of disaster management and response systems. In this work, we revisit fundamental combination methods like concatenation, addition and element-wise multiplication with modern image and text feature abstraction models. We design a simple and efficient system which outperforms pre-existing methods on the FloodNet dataset and achieves state-of-the-art performance. This simplified system requires significantly less training and inference time than modern VQA architectures. We also study the performance of various backbones and report their consolidated results. Code is available at https://github.com/sahilkhose/floodnet_vqa.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. FloodVision: Urban Flood Depth Estimation Using Foundation Vision-Language Models and Domain Knowledge Graph

    cs.CV 2025-09 conditional novelty 5.0 of 10

    FloodVision uses GPT-4o plus a knowledge graph of object heights to estimate urban flood depth from RGB images, achieving 8.17 cm MAE on 110 crowdsourced images.

  2. Damage Assessment after Natural Disasters with UAVs: Semantic Feature Extraction using Deep Learning

    cs.CV 2024-12 reject novelty 5.0 of 10

    A learned binary mask on semantic segmentation maps reduces UAV-to-ground data volume, but accuracy is maintained on only one of the two tested tasks.

Pith tools