Pith. sign in

REVIEW 3 cited by

Rethinking Efficient and Effective Point-based Networks for Event Camera Classification and Regression: EventMamba

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2405.06116 v4 pith:N2FTEKX5 submitted 2024-05-09 cs.CV

classification cs.CV
keywords cloudeventtemporaleventmambaframe-basedpointpoint-basedcamera
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Event cameras draw inspiration from biological systems, boasting low latency and high dynamic range while consuming minimal power. The most current approach to processing Event Cloud often involves converting it into frame-based representations, which neglects the sparsity of events, loses fine-grained temporal information, and increases the computational burden. In contrast, Point Cloud is a popular representation for processing 3-dimensional data and serves as an alternative method to exploit local and global spatial features. Nevertheless, previous point-based methods show an unsatisfactory performance compared to the frame-based method in dealing with spatio-temporal event streams. In order to bridge the gap, we propose EventMamba, an efficient and effective framework based on Point Cloud representation by rethinking the distinction between Event Cloud and Point Cloud, emphasizing vital temporal information. The Event Cloud is subsequently fed into a hierarchical structure with staged modules to process both implicit and explicit temporal features. Specifically, we redesign the global extractor to enhance explicit temporal extraction among a long sequence of events with temporal aggregation and State Space Model (SSM) based Mamba. Our model consumes minimal computational resources in the experiments and still exhibits SOTA point-based performance on six different scales of action recognition datasets. It even outperformed all frame-based methods on both Camera Pose Relocalization (CPR) and eye-tracking regression tasks. Our code is available at: https://github.com/rhwxmx/EventMamba.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 3 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. EV-Flying: an Event-based Dataset for In-The-Wild Recognition of Flying Objects

    cs.CV 2025-06 conditional novelty 6.0 of 10

    EV-Flying is a hand-annotated event-camera dataset of birds, insects, and drones, with a PointNet++ benchmark reaching about 72% single-chunk and 92% full-track accuracy.

  2. Scalable Event Cloud Network for Event-based Classification

    cs.CV 2024-12 conditional novelty 6.0 of 10

    A frequency-aware network operating on raw-like event clouds matches or beats prior event-based models on nine benchmarks while using roughly 0.1 G MACs, far below frame and voxel baselines.

  3. Learning Normal Flow Directly From Event Neighborhoods

    cs.CV 2024-12 reject novelty 6.0 of 10

    A point-based network learns per-event normal flow from raw event camera data and, with IMU data, estimates egomotion; it transfers across datasets better than frame-based optical flow methods.

Pith tools