eclouder / MoE-LLM Public

Notifications You must be signed in to change notification settings
Fork 0
Star 2

Implementation of the paper: "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"

2 stars 0 forks Branches Tags Activity

Notifications

Name		Name	Last commit message	Last commit date
Latest commit History 3 Commits
model		model
LICENSE		LICENSE
README.md		README.md

Repository files navigation

MoE-LLM

Implementation of the paper: "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"

About

Implementation of the paper: "Mixture-of-Depths: Dynamically allocating compute in transformer-based language models"

Report repository

Releases

No releases published

Packages

No packages published

Languages

Python 100.0%