Pith. sign in

Paper Citation Record · LEDGER

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain

As of 12 August 2026, this Paper Citation Record lists 7 of 7 outbound references and 1 inbound Pith citation observation for arXiv:2501.06277.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.06277 v1

Coverage vector

measured 7 of 7 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:08:58.135807Z

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T17:08:05.347949Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

7 of 7 outbound references displayed

  • verified exact0
  • verified fuzzy1
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e7ce40c8-0495-4e8b-8e85-7d0b86628438 · outbound

This paper cites foundation models,.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain foundation models,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:08:58.234436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:08:58.045633Z digest=sha256:6c9138135fa00cbee84e90e4af17d349860e59d97421d2ca027688f3dccfcd0b

Observation 5fd2a746-8723-4c2e-9ee2-c48677a4bb5a · outbound

This paper cites BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain BioMistral: A Collection of Open-Source Pretrained Large Language Models for Medical Domains

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-10T21:08:58.056278Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:08:58.056278Z digest=sha256:9628419b0beb783c2132e0de0cabbc7c4edba3d7182e6d68b6dc4b83c4835995

Observation ef9eabaa-0f8a-49f8-964c-5c8326d2b417 · outbound

This paper cites OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain OmniEval: An Omnidirectional and Automatic RAG Evaluation Benchmark in Financial Domain

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:08:58.062692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:08:58.062692Z digest=sha256:fa66c649268c2512b62670c36c12997796511e8235b30a6b79fa85c830af9a60

Observation a2912db0-2e2d-4bbf-9732-8689ff234535 · outbound

This paper cites Large Language Model for Participatory Urban Planning.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain Large Language Model for Participatory Urban Planning

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-10T21:08:58.135807Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:08:58.135807Z digest=sha256:92627654c0130c119686805dd0dcbb1f48933074e9f7b3c4e19b85dc1fbf15e0

Observation 86462444-67fc-478b-a60c-bc8376aa2536 · outbound

This paper cites SuperCLUE: A Comprehensive Chinese Large Language Model Benchmark.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain SuperCLUE: A Comprehensive Chinese Large Language Model Benchmark

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-10T21:08:58.129896Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:08:58.129896Z digest=sha256:6730f51f9aa9e1d5237ddc16896cf4fa055371d0a16e7e9d20b2c18dedd3ff89

Observation c067fa4c-c462-4a07-b63b-423631a37309 · outbound

This paper cites Composing Open-domain Vision with RAG for Ocean Monitoring and Conservation.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain Composing Open-domain Vision with RAG for Ocean Monitoring and Conservation

Reference 2024

Resolution
metadata mismatch
local_arxiv, observed 2026-08-10T21:08:58.219723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:08:58.050673Z digest=sha256:918187844ad749c52fd2d6832ca95732865d3e8d9545fe753c75d7085b0a4ace

Observation b2a78275-772d-4111-b523-bb9a91c5ac60 · outbound

This paper cites an unresolved cited work.

Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain Unresolved cited work

Reference 2025

Resolution
unresolved
raw_fallback, observed 2026-08-10T21:08:58.248079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:08:58.038493Z digest=sha256:004c6e64cefce602005c760d560d0a8202f4f7d5863f01f4bc209fb4aaed4593

Pith citing papers

Observation e0e42dc6-3001-4f04-830e-7d735cc3f84e · inbound

WuYu-EnvLE-Bench: A Benchmark for Evaluating Large Language Models in Environmental Law Enforcement cites this paper.

WuYu-EnvLE-Bench: A Benchmark for Evaluating Large Language Models in Environmental Law Enforcement Environmental large language model Evaluation (ELLE) dataset: A Benchmark for Evaluating Generative AI applications in Eco-environment Domain

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-01T17:08:05.347949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T17:08:05.347949Z digest=sha256:ef93d919c6e029136b466040f64db2ca1c7a44be7bb0d3ff567733e5fa84bd92