| FazBrowse GitHub Viewer | Trending | | Home |
| Tools: [Original HTTPS Page] |
Quantum's Last Exam — an auto-generated, expert-validated discovery benchmark for AI agents in quantum science
Official implementation of the NeurIPS 2025 paper "Soft Thinking: Unlocking the Reasoning Potential of LLMs in Continuous Concept Space"
Official codebase for the paper "WorldMemArena: Evaluating Multimodal Agent Memory Through Action–World Interaction"
Official implementation of the paper "Length Value Model: Pretraining Value Model for Scalable Length Prediction and Control"
Loading…
Loading…
| Back | FazBrowse Home | New Git URL |