Logo
Explore Help
Sign In
cmedina/agenx-lora-training
Watch 1
Star 0 Fork 0
Code Issues Pull Requests Actions Packages Projects Releases Wiki Activity
164 Commits 1 Branch 0 Tags
a0f4f644b018d2034c104fcf9141097dee56873c
T
Clone
Open with VS Code Open with VSCodium Open with Intellij IDEA
Download ZIP Download TAR.GZ Download BUNDLE
Christian Medina a0f4f644b0 fix: load 4-bit model with device_map=auto (let transformers distribute)
2026-07-02 19:33:37 -04:00
training
feat: add CPU offload for optimizer states
2026-07-02 17:53:56 -04:00
.gitignore
fix: remove large dataset from git tracking, add to .gitignore
2026-06-30 15:06:13 -04:00
deploy-agenx-lora.sh
fix: use $HOME/loras by default
2026-06-30 15:51:28 -04:00
deploy-and-train.sh
refactor: simplify deploy script, add train-on-this-server.sh
2026-06-30 15:22:54 -04:00
inference.py
refactor: restructure project - scripts at root level
2026-07-01 16:39:54 -04:00
inspect_model.py
docs: add model inspection script and comment failing tests
2026-07-02 13:53:45 -04:00
prepare_dataset.py
refactor: restructure project - scripts at root level
2026-07-01 16:39:54 -04:00
test_model_loading.py
feat: add Test 11 - PEFT prepare + manual 4-bit quantization
2026-07-02 14:30:19 -04:00
train-on-this-server.sh
fix: single-process training, remove FSDP (model pre-distributed via device_map)
2026-07-02 15:53:51 -04:00
train.py
fix: load 4-bit model with device_map=auto (let transformers distribute)
2026-07-02 19:33:37 -04:00
S
Description
LoRA training infrastructure for Cyron summary generation
2.6 MiB
0 Stars 1 Watchers 0 Forks
Languages
Python 94.9%
Shell 5.1%
Powered by Gitea Version: 1.28.0+dev-324-g133a3b8567 Page: 47ms Template: 4ms
Auto
English
Bahasa Indonesia Deutsch English Español Français Gaeilge Italiano Latviešu Magyar nyelv Nederlands Polski Português de Portugal Português do Brasil Suomi Svenska Türkçe Čeština Ελληνικά Български Русский Українська فارسی മലയാളം 日本語 简体中文 繁體中文(台灣) 繁體中文(香港) 한국어
Licenses API