Awesome AI Agents
📄
Papers
🧭
Topics
🔥
Trending
🗺️
Map
🏆
Leaderboards
🎓
Learn
🤖
Ask AI
⋯
More
👥
Authors
📚
Reading Packs
📊
Datasets
🛠️
Tools
📰
News
📝
Blogs
✉️
Newsletter
🎯
Research Radar
🔖
Saved
+ Add Paper
☾
☀
← authors
·
overview
Loading author…
🤖
Ask AI
Himabindu Lakkaraju — most-cited papers & profile · AI Agents
← authors
·
overview
Himabindu Lakkaraju
14
papers ·
414
citations ·
30
h-index
Harvard University Press
Google Scholar ↗
Semantic Scholar ↗
OpenAlex ↗
Most-cited papers
In-context Unlearning: Language Models As Few Shot Unlearners
2023 · 212 citations
Faithfulness Vs. Plausibility: On The (un)reliability Of Explanations From Large Language Models
2024 · 107 citations
Quantifying Uncertainty In Natural Language Explanations Of Large Language Models
2023 · 36 citations
Manipulating Large Language Models To Increase Product Visibility
2024 · 27 citations
Towards a Unified Framework for Fair and Stable Graph Representation Learning
2021 · 21 citations
Learning Cost-Effective Treatment Regimes using Markov Decision Processes
2016 · 10 citations
Learning Cost-Effective and Interpretable Regimes for Treatment Recommendation
2016 · 1 citations
Towards Robust Off-Policy Evaluation via Human Inputs
2022
All Roads Lead to Rome? Exploring Representational Similarities Between Latent Spaces of Generative Image Models
2024
Explaining the Model, Protecting Your Data: Revealing and Mitigating the Data Privacy Risks of Post-Hoc Model Explanations via Membership Inference
2024
Quantifying Generalization Complexity for Large Language Models
2024
Topics
Evaluation
Prompting
Safety & Alignment
Training Techniques
Model-Based RL
Value-Based
Safe RL
Offline RL
In-Context Learning
Efficiency