IdealGPT
by Hxyou
Official Code of IdealGPT
AI summary
Vision Reasoning Framework
A deep learning framework for iteratively decomposing vision and language reasoning via large language models.
- stars
- 32
- forks
- 8
- watching
- 2
Similar projects
Found by comparing what the projects do, not just their names.
Vision-Language Model Framework
Implementing a unified modal learning framework for generative vision-language models
yuxie11/r2d2157
Vision-Language Framework
A framework for large-scale cross-modal benchmarks and vision-language tasks in Chinese
fyu/dilation782
Image segmentation framework
This project provides a deep learning framework implementing dilated convolutions for semantic image segmentation
tobypde/frrn280
Image segmentation framework
A software framework for training and evaluating full-resolution residual networks for semantic image segmentation tasks
nvlabs/prismer1.3K
Vision-Language Model
A deep learning framework for training multi-modal models with vision and language capabilities.
Region-of-Interest Training
Training and deploying large language models on computer vision tasks using region-of-interest inputs
Deep learning framework
A deep learning framework built on top of Theano, providing a wide range of models and training techniques for research and development.
Vision framework
A computer vision framework for robotics applications that simplifies the creation of vision systems and generates code in multiple programming languages.
jy0205/lavit544
Visual understanding and generation framework
A unified framework for training large language models to understand and generate visual content
Optimization framework
An approach to train and optimize machine learning models in a decentralized setting by convexifying the optimization process
Hyperparameter optimizer
A reinforcement learning-based framework for optimizing hyperparameters in distributed machine learning environments.
Forgetting prevention framework
An approach to mitigating catastrophic forgetting in federated class incremental learning for vision tasks using a generative model and data-free methods
Visual unification framework
A framework for unified visual representation in image and video understanding models, enabling efficient training of large language models on multimodal data.
Optimization framework
A software framework for training neural networks to optimize dielectric metasurfaces using physics-driven generative models and global optimization algorithms.
Benchmark
An image-context reasoning benchmark designed to challenge large vision-language models and help improve their accuracy