There was an error while loading. Please reload this page.
Are We on the Right Way for Evaluating Large Vision-Language Model?