Qwen2-Audio
by QwenLM
The official repo of Qwen2-Audio chat & pretrained large audio language model proposed by Alibaba Cloud.
AI summary
Audio Response Model
An audio-language model that can analyze or respond to speech instructions based on audio input
- stars
- 1.3K
- forks
- 91
- watching
- 33
Similar projects
Found by comparing what the projects do, not just their names.
Audio Model
A multimodal audio language model developed by Alibaba Cloud that supports various tasks and languages
Audio Model
An Audio Language Model framework that uses transfer learning to generate text from audio inputs
Audio models
A collection of pre-trained audio and speech models for various applications
Audio responder
An implementation of a joint speech language model that responds directly to audio input
Audio Model
An audio and speech large language model implementation with pre-trained models, datasets, and inference options
qwenlm/qwen14.8K
Chinese language models
This repository provides large language models and chat capabilities based on pre-trained Chinese models.
Audio classifier
This project provides an implementation of a deep learning framework to classify audio signals and offers insights into the model's decision-making process using Explainable Artificial Intelligence (AI) techniques.
Audio Reconstruction
Reconstructs audio features learned by convolutional neural networks into audible sounds
qwenlm/qwen2-vl3.6K
Multimodal LM
A multimodal large language model series developed by the Qwen team to understand and process images, videos, and text.
Language model development platform
Develops and releases large language models trained on vast amounts of data for various applications, including natural language understanding, text generation, and more.
Audio API
Provides a set of audio APIs for Go programming language
Audio classifier
A system for audio classification and detection using machine learning models
XML audio definition model library
An ITU-R BS.2076 conformant XML library for audio definition model creation and modification
Speech Transcription Engine
Enables speech-to-text transcription using a pre-trained neural network model in MATLAB.
Language model toolkit
Provides pre-trained language models and tools for fine-tuning and evaluation