Pith. sign in

REVIEW 3 cited by

MM-Food-100K: A 100,000-Sample Multimodal Food Intelligence Dataset with Verifiable Provenance

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2508.10429 v1 pith:J7VRARQQ submitted 2025-08-14 cs.AI cs.CRcs.CV

MM-Food-100K: A 100,000-Sample Multimodal Food Intelligence Dataset with Verifiable Provenance

classification cs.AI cs.CRcs.CV
keywords mm-food-100kfoodaccessapproximatelychatgptcontributorscorpusdataset
verification ladder T0 review T1 audit T2 compute T3 formal T4 reserved
0 comments
read the original abstract

We present MM-Food-100K, a public 100,000-sample multimodal food intelligence dataset with verifiable provenance. It is a curated approximately 10% open subset of an original 1.2 million, quality-accepted corpus of food images annotated for a wide range of information (such as dish name, region of creation). The corpus was collected over six weeks from over 87,000 contributors using the Codatta contribution model, which combines community sourcing with configurable AI-assisted quality checks; each submission is linked to a wallet address in a secure off-chain ledger for traceability, with a full on-chain protocol on the roadmap. We describe the schema, pipeline, and QA, and validate utility by fine-tuning large vision-language models (ChatGPT 5, ChatGPT OSS, Qwen-Max) on image-based nutrition prediction. Fine-tuning yields consistent gains over out-of-box baselines across standard metrics; we report results primarily on the MM-Food-100K subset. We release MM-Food-100K for publicly free access and retain approximately 90% for potential commercial access with revenue sharing to contributors.

discussion (0)

Sign in with ORCID, Apple, or X to comment. Anyone can read and Pith papers without signing in.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

    cs.AI 2026-07 conditional novelty 6.0

    VLMs show a Semantic-Physical Gap on food images: strong dish naming but high MAPE on mass/nutrients and frequent unsafe advice for high-risk disease profiles.

  2. OmniFood-Bench: Evaluating VLMs for Nutrient Reasoning and Personalized Health Advice

    cs.AI 2026-07 conditional novelty 6.0

    A new benchmark shows current vision-language models can name foods but fail at estimating portion and nutrient values and often give unsafe dietary advice for chronic-disease patients.

  3. Food-R1: A Unified Multi-Task Food Vision-Language Model with Reinforcement Learning

    cs.CV 2026-06 unverdicted novelty 5.0

    Introduces CalorieBench-80K benchmark with CoT calorie reasoning and Food-R1 VLM trained via CoT cold-start then GRPO reinforcement fine-tuning, claiming consistent outperformance on food tasks.