Awesome AI for Science
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Nouha Dziri — most-cited papers & profile · AI for Science
← authors
·
overview
Nouha Dziri
7
papers ·
17
citations
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
WildGuard: Open One-Stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
2024 · 7 citations
Faith and Fate: Limits of Transformers on Compositionality
2023 · 4 citations
Fine-Grained Human Feedback Gives Better Rewards for Language Model Training
2023 · 3 citations
The Art of Saying No: Contextual Noncompliance in Language Models
2024 · 3 citations
Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
2025
RewardBench: Evaluating Reward Models for Language Modeling
2024
TÜLU 3: Pushing Frontiers in Open Language Model Post-Training
2024
Topics
Evaluation
Training Techniques
Safety & Alignment
Vision-Language
Fine-Tuning
Code
Reinforcement Learning
cs.CL
Model Architecture
RAG