| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
AtlasNLP is a resource that maps dataset representation by country and shows that coverage, production, and task diversity are deeply uneven across the global NLP ecosystem.
One Example Is Enough to Pass Fairness Benchmarks: Rethinking Fairness Evaluation for Aligned LLMs (EMNLP 2026). Code, data and one-shot GRPO LoRA adapters.
The Language–Energy Divide: measuring energy costs of multilingual LLM inference across 122 languages. Analysis code + per-language energy/accuracy results.
Multilingual Deception Detection of GPT-generated Hotel Reviews
🎯 Accepted to COLM 2026 — One flipped example is all it takes: how one-shot GRPO breaks LLM alignment and induces systematic, generalizing bias. Code, data & models for 'It Takes One to Bias Them All.'
Wait, am I Being Fair? Characterizing deductive stereotyping in LLM reasoning and mitigating it with Fair-GCG (reasoning-time fairness steering).
Code and the VETO benchmark for 'The Wrong Kind of Right: Quantifying and Localizing Misfired Alignment in LLMs'.
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |