MathVista
Math Reasoning Evaluation Platform
Evaluating mathematical reasoning in visual contexts using large language models and multimodal AI
MathVista: data, code, and evaluation for Mathematical Reasoning in Visual Contexts
253 stars
6 watching
39 forks
Language: Jupyter Notebook
last commit: almost 2 years agoai4mathlarge-language-modelslarge-multimadality-modelsmachine-learningmathematicsmathqasciencevisual-question-answering
Related projects:
| Repository | Description | Stars |
|---|---|---|
| A dataset and code framework to evaluate the ability of Large Multimodal Models (LMMs) to reason mathematically with visual contexts. | 74 | |
| A dataset and software framework for building multimodal reasoning systems to answer science questions. | 615 | |
| A command-line utility for performing arithmetic and symbolic math operations. | 178 | |
| An evaluation toolkit for large vision-language models | 1,514 | |
| A math expression evaluator library that allows users to perform calculations with precision and various units, functions, and constants. | 110 | |
| Enables visualizing pandas dataframes in Neovim using Visidata | 26 | |
| A unified framework for training large language models to understand and generate visual content | 544 | |
| A deep learning framework designed to improve visual reasoning capabilities by utilizing concepts and semantic relations. | 64 | |
| This project presents a neural network model designed to answer visual questions by combining question and image features in a residual learning framework. | 39 | |
| An open-source implementation of a deep learning model designed to improve the balance between performance and interpretability in visual reasoning tasks. | 348 | |
| Develops a framework to generate responses by composing various tools with large language models. | 1,095 | |
| Develops a multimodal Chinese language model with visual capabilities | 429 | |
| A cross-platform calculator built with QML and Material Design | 29 | |
| Creating synthetic visual reasoning instructions to improve the performance of large language models on image-related tasks | 18 | |
| Develops an end-to-end model for multiple visual perception and reasoning tasks using a single encoder, decoder, and large language model. | 1,336 |