REVIEW 5 cited by
Learning to Compose Neural Networks for Question Answering
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We describe a question answering model that applies to both images and structured knowledge bases. The model uses natural language strings to automatically assemble neural networks from a collection of composable modules. Parameters for these modules are learned jointly with network-assembly parameters via reinforcement learning, with only (world, question, answer) triples as supervision. Our approach, which we term a dynamic neural model network, achieves state-of-the-art results on benchmark datasets in both visual and structured domains.
Forward citations
Cited by 5 Pith papers
-
Learning to Reason Iteratively and Parallelly for Complex Visual Reasoning Scenarios
A new attention-based reasoning module combining iterative steps with parallel operation slots improves accuracy on multiple visual question answering benchmarks while staying lightweight and partially interpretable.
-
Language Model as Visual Explainer
LVX builds LLM-generated attribute trees to explain any trained image classifier without training the explainer, but its faithfulness metric is directly optimized by the method.
-
Mitigating Knowledge Conflicts in Language Model-Driven Question Answering
On memorized question-answer pairs from KMIR and NQ, bottleneck and prefix adapters trained on entity-swapped contexts let a GPT-2 reader follow the new context most of the time, though no baselines are reported.
-
Using Large Language Models for education managements in Vietnamese with low resources
A framework that fine-tunes Bloom and Vistral on a synthetic Vietnamese educational-management QA dataset, with Vistral scoring higher but with no external baseline.
-
A Comprehensive Survey on Visual Question Answering Datasets and Algorithms
A broad but dated survey of VQA datasets and algorithms that organizes the pre-2021 literature into four dataset categories and six model paradigms.
Discussion (0). Continue with ORCID to comment.