MathVista

Math Reasoning Evaluation Platform

Evaluating mathematical reasoning in visual contexts using large language models and multimodal AI

MathVista: data, code, and evaluation for Mathematical Reasoning in Visual Contexts

GitHub

253 stars
6 watching
39 forks
Language: Jupyter Notebook
last commit: almost 2 years ago
ai4mathlarge-language-modelslarge-multimadality-modelsmachine-learningmathematicsmathqasciencevisual-question-answering

Related projects:

RepositoryDescriptionStars
mathllm/math-vA dataset and code framework to evaluate the ability of Large Multimodal Models (LMMs) to reason mathematically with visual contexts.74
lupantech/scienceqaA dataset and software framework for building multimodal reasoning systems to answer science questions.615
metadelta/mdltA command-line utility for performing arithmetic and symbolic math operations.178
open-compass/vlmevalkitAn evaluation toolkit for large vision-language models1,514
5anthosh/fcalA math expression evaluator library that allows users to perform calculations with precision and various units, functions, and constants.110
willem-j-an/visidata.nvimEnables visualizing pandas dataframes in Neovim using Visidata26
jy0205/lavitA unified framework for training large language models to understand and generate visual content544
nvlabs/relvitA deep learning framework designed to improve visual reasoning capabilities by utilizing concepts and semantic relations.64
jnhwkim/nips-mrn-vqaThis project presents a neural network model designed to answer visual questions by combining question and image features in a residual learning framework.39
davidmascharka/tbd-netsAn open-source implementation of a deep learning model designed to improve the balance between performance and interpretability in visual reasoning tasks.348
lupantech/chameleon-llmDevelops a framework to generate responses by composing various tools with large language models.1,095
airaria/visual-chinese-llama-alpacaDevelops a multimodal Chinese language model with visual capabilities429
lirios/calculatorA cross-platform calculator built with QML and Material Design29
rucaibox/comvintCreating synthetic visual reasoning instructions to improve the performance of large language models on image-related tasks18
lxtgh/omg-segDevelops an end-to-end model for multiple visual perception and reasoning tasks using a single encoder, decoder, and large language model.1,336