TorchAO: PyTorch-Native Training-to-Serving Model Optimization Paper • 2507.16099 • Published Jul 21, 2025 • 7
ZEBRA: Zero-Shot Example-Based Retrieval Augmentation for Commonsense Question Answering Paper • 2410.05077 • Published Oct 7, 2024 • 5
Exploring Non-Verbal Predicates in Semantic Role Labeling: Challenges and Opportunities Paper • 2307.01870 • Published Jul 4, 2023
Pap2Pat: Benchmarking Outline-Guided Long-Text Patent Generation with Patent-Paper Pairs Paper • 2410.07009 • Published Oct 9, 2024 • 1
Intriguing Properties of Large Language and Vision Models Paper • 2410.04751 • Published Oct 7, 2024 • 16
ReLiK: Retrieve and LinK, Fast and Accurate Entity Linking and Relation Extraction on an Academic Budget Paper • 2408.00103 • Published Jul 31, 2024 • 24
Stark: Social Long-Term Multi-Modal Conversation with Persona Commonsense Knowledge Paper • 2407.03958 • Published Jul 4, 2024 • 21
Distillation Contrastive Decoding: Improving LLMs Reasoning with Contrastive Decoding and Distillation Paper • 2402.14874 • Published Feb 21, 2024 • 4
A Deep Neural Network for SSVEP-based Brain-Computer Interfaces Paper • 2011.08562 • Published Nov 17, 2020 • 3
MTet: Multi-domain Translation for English and Vietnamese Paper • 2210.05610 • Published Oct 11, 2022 • 3
ViT5: Pretrained Text-to-Text Transformer for Vietnamese Language Generation Paper • 2205.06457 • Published May 13, 2022
The BigScience ROOTS Corpus: A 1.6TB Composite Multilingual Dataset Paper • 2303.03915 • Published Mar 7, 2023 • 9
SciFive: a text-to-text transformer model for biomedical literature Paper • 2106.03598 • Published May 28, 2021
BLOOM: A 176B-Parameter Open-Access Multilingual Language Model Paper • 2211.05100 • Published Nov 9, 2022 • 39
Evaluate & Evaluation on the Hub: Better Best Practices for Data and Model Measurements Paper • 2210.01970 • Published Sep 30, 2022 • 14
HuggingFace's Transformers: State-of-the-art Natural Language Processing Paper • 1910.03771 • Published Oct 9, 2019 • 26