| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.
You must be logged in to block users.
Contact GitHub support about this user’s behavior. Learn more about reporting abuse.
Report abuseI study how to make AI systems (LLMs, VLMs, and agents) reliable and aligned, with a focus on reward modeling, evaluation, and agentic visual reasoning.
Research Directions
Open-source & Activity
Contact
Feel free to reach out if you'd like to chat or collaborate.
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
[CVPR 2026] Official Code for "ARM-Thinker: Reinforcing Multimodal Generative Reward Models with Agentic Tool Use and Visual Reasoning"
[ICCV 2025] MM-IFEngine: Towards Multimodal Instruction Following
Python 125
Any source (PDF, video, web, audio, text) to interactive learning package with quizzes, flashcards and spaced repetition. One command, 12-section study guide.
Official Implementation of "Visual-ERM: Reward Modeling for Visual Equivalence"
Python 65
[EMNLP 2026] An official Implementation of "Beyond the Current Observation: Evaluating Multimodal Large Language Models in Controllable Non-Markov Games"
Python 41
| Back | FazBrowse Home | New Git URL |