Text Generation
Transformers
Safetensors
lora
aya
tiny-aya
multilingual
code
legesher
tiny-aya-expedition
language-decoded
unsloth
arxiv:2603.11510
arxiv:2211.15533
arxiv:2510.09591
arxiv:1809.05053
arxiv:2308.16884
arxiv:2106.06937
arxiv:2210.03057
Instructions to use legesher/language-decoded-lora with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use legesher/language-decoded-lora with Transformers:
# Use a pipeline as a high-level helper from transformers import pipeline pipe = pipeline("text-generation", model="legesher/language-decoded-lora")# Load model directly from transformers import AutoModel model = AutoModel.from_pretrained("legesher/language-decoded-lora", device_map="auto") - Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- vLLM
How to use legesher/language-decoded-lora with vLLM:
Install from pip and serve model
# Install vLLM from pip: pip install vllm # Start the vLLM server: vllm serve "legesher/language-decoded-lora" # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:8000/v1/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "legesher/language-decoded-lora", "prompt": "Once upon a time,", "max_tokens": 512, "temperature": 0.5 }'Use Docker
docker model run hf.co/legesher/language-decoded-lora
- SGLang
How to use legesher/language-decoded-lora with SGLang:
Install from pip and serve model
# Install SGLang from pip: pip install sglang # Start the SGLang server: python3 -m sglang.launch_server \ --model-path "legesher/language-decoded-lora" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "legesher/language-decoded-lora", "prompt": "Once upon a time,", "max_tokens": 512, "temperature": 0.5 }'Use Docker images
docker run --gpus all \ --shm-size 32g \ -p 30000:30000 \ -v ~/.cache/huggingface:/root/.cache/huggingface \ --env "HF_TOKEN=<secret>" \ --ipc=host \ lmsysorg/sglang:latest \ python3 -m sglang.launch_server \ --model-path "legesher/language-decoded-lora" \ --host 0.0.0.0 \ --port 30000 # Call the server using curl (OpenAI-compatible API): curl -X POST "http://localhost:30000/v1/completions" \ -H "Content-Type: application/json" \ --data '{ "model": "legesher/language-decoded-lora", "prompt": "Once upon a time,", "max_tokens": 512, "temperature": 0.5 }' - Unsloth Studio
How to use legesher/language-decoded-lora with Unsloth Studio:
Install Unsloth Studio (macOS, Linux, WSL)
curl -fsSL https://unsloth.ai/install.sh | sh # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for legesher/language-decoded-lora to start chatting
Install Unsloth Studio (Windows)
irm https://unsloth.ai/install.ps1 | iex # Run unsloth studio unsloth studio -H 0.0.0.0 -p 8888 # Then open http://localhost:8888 in your browser # Search for legesher/language-decoded-lora to start chatting
Using HuggingFace Spaces for Unsloth
# No setup required # Open https://huggingface.co/spaces/unsloth/studio in your browser # Search for legesher/language-decoded-lora to start chatting
Load model with FastModel
pip install unsloth from unsloth import FastModel model, tokenizer = FastModel.from_pretrained( model_name="legesher/language-decoded-lora", max_seq_length=2048, ) - Docker Model Runner
How to use legesher/language-decoded-lora with Docker Model Runner:
docker model run hf.co/legesher/language-decoded-lora
Add Attribution & takedown section
#13
by madiedgar - opened
README.md
CHANGED
|
@@ -47,7 +47,7 @@ The hypothesis is **not** that non-English code matches or exceeds English code
|
|
| 47 |
|
| 48 |
## Base Model
|
| 49 |
|
| 50 |
-
All adapters are trained on [CohereLabs/tiny-aya-base](https://huggingface.co/CohereLabs/tiny-aya-base) (3.35B parameters). Tiny Aya was chosen because it is small (deployable on a single 16 GB T4 GPU via QLoRA),
|
| 51 |
|
| 52 |
## Adapter Inventory
|
| 53 |
|
|
@@ -222,4 +222,34 @@ Paper-grade evaluation results live on [`legesher/language-decoded-experiments`]
|
|
| 222 |
|
| 223 |
## License
|
| 224 |
|
| 225 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 47 |
|
| 48 |
## Base Model
|
| 49 |
|
| 50 |
+
All adapters are trained on [CohereLabs/tiny-aya-base](https://huggingface.co/CohereLabs/tiny-aya-base) (3.35B parameters). Tiny Aya was chosen because it is small (deployable on a single 16 GB T4 GPU via QLoRA), openly available (released under CC-BY-NC-4.0), and supports 70+ languages with explicit emphasis on lower-resourced ones — which makes the experimental ladder viable for `ur` at all.
|
| 51 |
|
| 52 |
## Adapter Inventory
|
| 53 |
|
|
|
|
| 222 |
|
| 223 |
## License
|
| 224 |
|
| 225 |
+
CC-BY-NC-4.0. The adapters inherit the license of the base model,
|
| 226 |
+
[CohereLabs/tiny-aya-base](https://huggingface.co/CohereLabs/tiny-aya-base)
|
| 227 |
+
(CC-BY-NC-4.0). The training datasets
|
| 228 |
+
([legesher/language-decoded-data](https://huggingface.co/datasets/legesher/language-decoded-data))
|
| 229 |
+
are separately licensed under Apache-2.0.
|
| 230 |
+
|
| 231 |
+
## Provenance, attribution & takedown
|
| 232 |
+
|
| 233 |
+
These adapters were fine-tuned from
|
| 234 |
+
[`CohereLabs/tiny-aya-base`](https://huggingface.co/CohereLabs/tiny-aya-base)
|
| 235 |
+
on a specific revision of the
|
| 236 |
+
[`legesher/language-decoded-data`](https://huggingface.co/datasets/legesher/language-decoded-data)
|
| 237 |
+
training conditions (see each adapter's configuration for the
|
| 238 |
+
condition and revision).
|
| 239 |
+
|
| 240 |
+
If you are the author of source code included in the training data
|
| 241 |
+
and would like attribution added or your code removed, open a
|
| 242 |
+
discussion on the dataset repository's **Community** tab or email
|
| 243 |
+
**support@legesher.com**. Removals are propagated in a new dataset
|
| 244 |
+
revision. Adapters already trained are frozen historical artifacts:
|
| 245 |
+
a dataset removal does not alter existing adapter weights, but we
|
| 246 |
+
will note affected conditions here and take reported concerns about
|
| 247 |
+
specific adapters into account.
|
| 248 |
+
|
| 249 |
+
**Usage caution.** These are research artifacts, not
|
| 250 |
+
production-ready models. Documented side effects include code
|
| 251 |
+
fragments leaking into natural-language output (strongest for Urdu
|
| 252 |
+
adapters) and matched-language regressions on specific evaluation
|
| 253 |
+
cells; the Condition 5 adapters were trained on corpora containing
|
| 254 |
+
raw translator output. Evaluate per language and per task before any
|
| 255 |
+
downstream use.
|