Pith. sign in

REVIEW 16 cited by

A Content-Driven Micro-Video Recommendation Dataset at Scale

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.15379 v1 pith:5R6UIZRG submitted 2023-09-27 cs.IR

classification cs.IR
keywords micro-videorecommendationdatasetmicrolensvideorecommenderchallengecontent-driven
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Micro-videos have recently gained immense popularity, sparking critical research in micro-video recommendation with significant implications for the entertainment, advertising, and e-commerce industries. However, the lack of large-scale public micro-video datasets poses a major challenge for developing effective recommender systems. To address this challenge, we introduce a very large micro-video recommendation dataset, named "MicroLens", consisting of one billion user-item interaction behaviors, 34 million users, and one million micro-videos. This dataset also contains various raw modality information about videos, including titles, cover images, audio, and full-length videos. MicroLens serves as a benchmark for content-driven micro-video recommendation, enabling researchers to utilize various modalities of video information for recommendation, rather than relying solely on item IDs or off-the-shelf video features extracted from a pre-trained network. Our benchmarking of multiple recommender models and video encoders on MicroLens has yielded valuable insights into the performance of micro-video recommendation. We believe that this dataset will not only benefit the recommender system community but also promote the development of the video understanding field. Our datasets and code are available at https://github.com/westlake-repl/MicroLens.

Discussion (0). Sign in to comment.

Forward citations

Cited by 16 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Bridging Language and Items for Retrieval and Recommendation: Benchmarking LLMs as Semantic Encoders

    cs.IR 2024-03 unverdicted novelty 8.0 of 10

    BLaIR is a new benchmark and 570M-review dataset showing that LLM performance rankings on recommendation tasks have little correlation with rankings on general embedding benchmarks like MTEB.

  2. Will It Go Viral? Grounding Micro-Video Popularity Prediction on the Open Web

    cs.MM 2026-05 unverdicted novelty 7.0 of 10

    WEBSHORTS dataset and SHORTS-CAST framework ground micro-video popularity prediction in structured open-web context collected at upload time and enable selective online adaptation using delayed labels.

  3. Sparse Contrastive Learning for Content-Based Cold Item Recommendation

    cs.IR 2026-04 unverdicted novelty 7.0 of 10

    SEMCo uses sparse entmax contrastive learning for purely content-based cold-start item recommendation, outperforming standard methods in ranking accuracy.

  4. UniRank: Benchmarking Ranking Models for Unified Sequential Modeling and Feature Interaction

    cs.IR 2026-07 conditional novelty 6.0 of 10

    UniRank is an open benchmark that standardizes chronological autoregressive supervision, multi-task evaluation, and capacity controls for 15 unified ranking models on five large datasets.

  5. Learning Sparse Representations of Multimodal Content for Enhanced Cold Item Recommendation

    cs.IR 2026-07 conditional novelty 6.0 of 10

    Sparse content embeddings with a pre-sparsification alpha-entmax activation outperform dense embeddings for cold-start item recommendation at lower storage cost, especially for users with multiple interests.

  6. Prediction Is Not Memory: Dual-Timescale Gated Profile Writing for Persistent User Modeling

    cs.IR 2026-07 conditional novelty 6.0 of 10

    A lightweight write-risk gate reduces harmful persistent-profile updates from 22.45% to about 14.5% on MicroLens-100K, and next-item ranking confidence is a poor substitute for write-risk scoring.

  7. Popcorn: A Configurable Benchmark for Visual Evidence in Multimodal Movie Recommendation

    cs.IR 2026-06 unverdicted novelty 6.0 of 10

    Popcorn is a new benchmark standardizing modality assembly, fusion, and evaluation of thumbnails, trailers, and full movies encoded by VLMs for multimodal movie recommendation.

  8. OmniTrend: Content-Context Modeling for Scalable Social Popularity Prediction

    cs.CV 2026-04 unverdicted novelty 6.0 of 10

    OmniTrend predicts popularity by combining separate content attractiveness and contextual exposure predictors using cross-modal and exogenous signals.

  9. RecBase: Generative Foundation Model Pretraining for Zero-Shot Recommendation

    cs.IR 2025-09 conditional novelty 6.0 of 10

    A from-scratch model that tokenizes items into hierarchical codes and predicts next-item codes reaches higher average zero-shot AUC on 8 datasets than LLM recommenders up to 7B parameters.

  10. EGRA:Toward Enhanced Behavior Graphs and Representation Alignment for Multimodal Recommendation

    cs.IR 2025-08 conditional novelty 6.0 of 10

    EGRA improves multimodal recommendation by using pretrained-model item embeddings to build the item-item graph and by dynamically weighting modality-behavior alignment per entity and per epoch.

  11. CaIRec: Calibrated Modality Imputation for Incomplete Multimodal Recommendation

    cs.IR 2026-07 conditional novelty 5.5 of 10

    Structurally calibrating imputed item modalities and then adapting them with pseudo-missing alignment and completion-aware graphs improves incomplete multimodal recommendation on Amazon benchmarks.

  12. CaIRec: Calibrated Modality Imputation for Incomplete Multimodal Recommendation

    cs.IR 2026-07 conditional novelty 5.0 of 10

    CaIRec consistently beats eleven baselines on three Amazon datasets (about 2-7% relative gains) by combining latent imputation, spectral cross-modal calibration, and recommendation-space alignment of recovered item features.

  13. Multimodal Large Language Models with Adaptive Preference Optimization for Sequential Recommendation

    cs.IR 2025-11 unverdicted novelty 5.0 of 10

    HaNoRec dynamically weights harder preference samples and applies Gaussian perturbations to output distributions to improve multimodal LLM performance on sequential recommendation tasks.

  14. Describe What You See with Multimodal Large Language Models to Enhance Video Recommendations

    cs.IR 2025-08 conditional novelty 5.0 of 10

    Replacing raw video and audio features with MLLM-generated natural-language captions improves hit rate and nDCG for two-tower and SASRec recommenders on MicroLens-100K.

  15. Beyond Noisy Signals: Dual-Level Denoising for Multi-modal Sequential Recommendation

    cs.IR 2026-07 conditional novelty 4.0 of 10

    A dual-level denoising framework that combines graph Laplacian smoothing and learnable FFT filtering improves multi-modal sequential recommendation on four benchmarks.

  16. ViLLA-MMBench: A Unified Benchmark Suite for LLM-Augmented Multimodal Movie Recommendation

    cs.IR 2025-08 unverdicted novelty 4.0 of 10

    ViLLA-MMBench is an open, YAML-configured benchmark combining audio, visual, and LLM-enriched text embeddings for movie recommendation, reporting cold-start and coverage gains from LLM augmentation.

Pith tools