ftgreat / mamba-chat Public

forked from redotvideo/mamba-chat

Notifications You must be signed in to change notification settings
Fork 0
Star 0

Mamba-Chat: A chat LLM based on the state-space model architecture 🐍

Notifications

Name		Name	Last commit message	Last commit date
Latest commit History 9 Commits
data		data
scripts		scripts
trainer		trainer
LICENSE		LICENSE
README.md		README.md
chat.py		chat.py
requirements.txt		requirements.txt
train_mamba.py		train_mamba.py

Repository files navigation

Mamba-Chat 🐍

Mamba-Chat is the first chat language model based on a state-space model architecture, not a transformer.

The model is based on Albert Gu's and Tri Dao's work Mamba: Linear-Time Sequence Modeling with Selective State Spaces as well as their model implementation. This repository provides training / fine-tuning code for the model based on some modifications of the Huggingface Trainer class.

Run Mamba-Chat

We provide code that lets you run inference on mamba-chat as well as our fine-tuning code. To get started, clone this repository and install its dependencies:

git clone https://github.com/havenhq/mamba-chat.git

cd mamba-chat
pip install -r requirements.txt

You can chat with mamba by

About

Mamba-Chat: A chat LLM based on the state-space model architecture 🐍

Readme

Apache-2.0 license

Activity

0 stars

0 watching

0 forks

Report repository

Releases

No releases published

Packages

No packages published

Languages

Python 100.0%

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Repository files navigation

Mamba-Chat 🐍

Run Mamba-Chat

About

Releases

Packages

Languages

License

ftgreat/mamba-chat

Folders and files

Latest commit

History

Repository files navigation

Mamba-Chat 🐍

Run Mamba-Chat

About

Resources

License

Stars

Watchers

Forks

Releases

Packages 0

Languages

Packages