REVIEW 3 cited by
AugLy: Data Augmentations for Robustness
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
Signed reviews
read the original abstract
We introduce AugLy, a data augmentation library with a focus on adversarial robustness. AugLy provides a wide array of augmentations for multiple modalities (audio, image, text, & video). These augmentations were inspired by those that real users perform on social media platforms, some of which were not already supported by existing data augmentation libraries. AugLy can be used for any purpose where data augmentations are useful, but it is particularly well-suited for evaluating robustness and systematically generating adversarial attacks. In this paper we present how AugLy works, benchmark it compared against existing libraries, and use it to evaluate the robustness of various state-of-the-art models to showcase AugLy's utility. The AugLy repository can be found at https://github.com/facebookresearch/AugLy.
Forward citations
Cited by 3 Pith papers
-
Diagnosing the Reliability of LLM-as-a-Judge via Item Response Theory
An IRT-based two-phase diagnostic framework with four metrics (CV, ρ, θratio, DW) for measuring LLM-judge intrinsic consistency and human alignment.
-
Benchmarking and Revisiting Code Generation Assessment: A Mutation-Based Approach
Benchmark scores for code-generation LLMs change substantially when the same problem is described in different words, so single-prompt benchmarks can misrank models.
-
Efficient and Accurate Image Provenance Analysis: A Scalable Pipeline for Large-scale Images
A provenance pipeline that augments retrieval with a pre-existing database of known image relationships claims to reduce analysis time from quadratic to linear while improving accuracy.
Discussion (0). Continue with ORCID to comment.