Multimodal-GPT
Multimodal Chatbot
Trains a multimodal chatbot that combines visual and language instructions to generate responses
Multimodal-GPT
1k stars
13 watching
125 forks
Language: Python
last commit: over 3 years agoflamingogptgpt-4llamamultimodaltransformervision-and-language
Related projects:
| Repository | Description | Stars |
|---|---|---|
| A multimodal chatbot with computer vision capabilities integrated into a single model | 99 | |
| A chatbot platform that integrates with various messaging services and provides a plugin-based architecture for customization and extensibility | 1,956 | |
| Connects to ChatGPT API via MicroPython to retrieve responses and display them on an OLED screen. | 27 | |
| An evaluation platform for comparing multi-modality models on visual question-answering tasks | 478 | |
| Enables OpenAI GPT to process multimedia inputs like images and audio with text output | 184 | |
| An open-source software framework that enables large language models to process and understand point cloud data, facilitating multimodal interactions. | 670 | |
| Provides a flexible and configurable framework for training deep learning models with PyTorch. | 1,196 | |
| An AWS Lambda-based Telegram bot for interacting with ChatGPT | 320 | |
| A MATLAB application providing an interface to access OpenAI's ChatGPT API | 203 | |
| A native application allowing users to interact with the GPT chat model on various platforms. | 421 | |
| An interactive chatbot for GitHub repositories using LLMs for conversational interaction and information retrieval | 283 | |
| A GitHub application built on top of ChatGPT and Probot to enable user interactions with a conversational bot. | 379 | |
| A chatbot for openHAB using machine-learning natural language processing | 15 | |
| Develops a unified model to generate high-quality motions and text descriptions from human motion data | 1,531 | |
| A conversational language model developed to improve understanding of complex instructions and Chinese vocabulary. | 62 |