Andromeda: Ultra-Fast and Ultra-Intelligent SOTA Language Model 🚀🌌
Welcome to Andromeda, The Fastest, Most Creative, and Reliable Language Model Ever Built, train your own verison, conduct inference, and finetune your own verison with simple plug in and play scripts get started in 10 seconds:
Features
- 💼 Handle Ultra Long Sequences (32,000-200,000+ context lengths)
- ⚡ Ultra Fast Processing (32,000+ tokens in under 100ms)
- 🎓 Superior Reasoning Capabilities
🎯 Principles
- Efficiency: Optimize with techniques like attention flashing, rotary position encodings, and deep normalization.
- Flexibility: Adapt to various tasks and domains for wide applications.
- Scalability: Designed to scale with resources and data sizes.
- Community-Driven: Thrives on contributions from the open-source community.
💻 Install
python3.11 -m pip install --upgrade andromeda-torch
Usage
- Forward pass with random inputs
import torch
from andromeda.configs import Andromeda1Billion
model = Andromeda1Billion()
x = torch.randint(0, 256, (1, 1024)).cuda()
out = model(x) # (1, 1024, 20000)
print(out)
- Tokenized inputs
from andromeda_torch import Tokenizer
from andromeda_torch.configs import Andromeda1Billion
model = Andromeda1Billion()
tokenizer = Tokenizer()
encoded_text = tokenizer.encode("Hello world!")
out = model(encoded_text)
print(out)
📚 Training
-
Set the environment variables:
ENTITY_NAME: Your wandb project nameOUTPUT_DIR: Directory to save the weights (e.g.,./weights)MASTER_ADDR: For distributed trainingMASTER_PORTFor master port distributed trainingRANK- Number of nodes servicesWORLD_SIZENumber of gpus
-
Configure the training:
- Accelerate Config
- Enable Deepspeed 3
- Accelerate launch train_distributed_accelerate.py
For more information, refer to the Training SOP.
Todo
- Add Yarn Embeddings from zeta
📈 Benchmarks
Speed
- Andromeda utilizes one of the most reliable Attentions ever, flash attention 2.0 Triton. It consumes 50x less memory than GPT-3 and 10x less than LLAMA.
- We can speed this up even more with dynamic sparse flash attention 2.0.
License
Apache License
Metadata
Release files for andromeda-torch 0.0.9
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| andromeda_torch-0.0.9.tar.gz | 23.3 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| andromeda_torch-0.0.9-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 46.1 kB
Release files / andromeda_torch-0.0.9.tar.gz
| Download URL | andromeda_torch-0.0.9.tar.gz |
|---|---|
| Size | 23.3 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
8ff70d7a4e8768a010e753b1fad8b5fc174d121a82f8e8ef2407d4d1d02f63d5
|
|
BLAKE2b-256 checksum How to use checksums |
ac6b397a1fe98024e70c7f8e81b819fcd4d7070dae7ea373ccae6b7bd78d6341
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/1.3.2 CPython/3.11.0 Darwin/23.3.0
|
Release files / andromeda_torch-0.0.9-py3-none-any.whl
| Download URL | andromeda_torch-0.0.9-py3-none-any.whl |
|---|---|
| Size | 22.7 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
5c2ab07a6f87336ea79675dd211c0b520ea649927b73e19a53fdd78ed7ea83fb
|
|
BLAKE2b-256 checksum How to use checksums |
4a821f65e43e27842f99980a253fff118ef1328caacb6c6030fd8eac591a0570
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/1.3.2 CPython/3.11.0 Darwin/23.3.0
|