Experimental Standards for Deep Learning in Natural Language Processing Research Paper • 2204.06251 • Published Apr 13, 2022
SnakModel: Lessons Learned from Training an Open Danish Large Language Model Paper • 2412.12956 • Published Dec 17, 2024 • 2
Evidence > Intuition: Transferability Estimation for Encoder Selection Paper • 2210.11255 • Published Oct 20, 2022
Multilingual GSM-Symbolic: What determines capability transfer across languages? Paper • 2610.03367 • Published 10 days ago • 52
SemEval-2020 Task 12: Multilingual Offensive Language Identification in Social Media (OffensEval 2020) Paper • 2006.07235 • Published Jun 12, 2020
Summon a Demon and Bind it: A Grounded Theory of LLM Red Teaming Paper • 2311.06237 • Published Nov 10, 2023 • 1
Surveying (Dis)Parities and Concerns of Compute Hungry NLP Research Paper • 2306.16900 • Published Jun 29, 2023
Efficient Methods for Natural Language Processing: A Survey Paper • 2209.00099 • Published Aug 31, 2022 • 2
garak: A Framework for Security Probing Large Language Models Paper • 2406.11036 • Published Jun 16, 2024 • 2
Prompt Refinement or Fine-tuning? Best Practices for using LLMs in Computational Social Science Tasks Paper • 2408.01346 • Published Aug 2, 2024
Introducing v0.5 of the AI Safety Benchmark from MLCommons Paper • 2404.12241 • Published Apr 18, 2024 • 13
Is a prompt and a few samples all you need? Using GPT-4 for data augmentation in low-resource classification tasks Paper • 2304.13861 • Published Apr 26, 2023