·
AI & ML interests
GenAI
LLM
Multi Modal
NLP
Reinforcement Learning
Recent Activity
Organizations
upvoted a paper 5 days ago upvoted a paper 5 months ago upvoted an article about 1 year ago view article Simplifying Alignment: From RLHF to Direct Preference Optimization (DPO)
ariG23498
• • 56