pablitoelpeligro commited on
Commit
3e2f205
·
verified ·
1 Parent(s): 082b448

Update README.md

Browse files
Files changed (1) hide show
  1. README.md +3 -1
README.md CHANGED
@@ -22,7 +22,7 @@ language:
22
  pipeline_tag: text-generation
23
  ---
24
 
25
- Fine-tuned Qwen2.5-0.5B GGUF for on-device personal finance chat. Maps questions to strict JSON intents (spend, merchants, categories, budgets…) for local SQLite tools. EN/IT/ES/FR. ~398 MB Q4_K_M.
26
 
27
  # AiBudget Intent 0.5B v5
28
 
@@ -31,6 +31,8 @@ Fine-tuned Qwen2.5-0.5B GGUF for on-device personal finance chat. Maps questions
31
  > Maps natural-language questions about personal spending (EN / IT / ES / FR) into a single, schema-constrained JSON intent object.
32
  > The AiBudget mobile app parses the JSON and runs **local SQLite** tools. When users choose this on-device model, **transaction data is not sent to a cloud LLM for intent parsing**.
33
 
 
 
34
  **Published artifact:** fused, quantized **GGUF only** — LoRA adapter weights are **not** in this repository.
35
 
36
  **Recommended download:** [`aibudget-intent-05b-v5-Q4_K_M.gguf`](https://huggingface.co/pablitoelpeligro/aibudget-intent-05b-GGUF/resolve/main/aibudget-intent-05b-v5-Q4_K_M.gguf) (~398 MB, SHA256 `d3d6fc0780359cd4d198e9354b9ecfc0ff9cfb0fb729f040326fa6e2f02c693a`)
 
22
  pipeline_tag: text-generation
23
  ---
24
 
25
+ LoRA fine-tune of Qwen2.5 **0.5B** (~398 MB Q4_K_M) for phones: small enough to download over cellular and run on-device, specialized for personal-finance intent JSON (EN/IT/ES/FR).
26
 
27
  # AiBudget Intent 0.5B v5
28
 
 
31
  > Maps natural-language questions about personal spending (EN / IT / ES / FR) into a single, schema-constrained JSON intent object.
32
  > The AiBudget mobile app parses the JSON and runs **local SQLite** tools. When users choose this on-device model, **transaction data is not sent to a cloud LLM for intent parsing**.
33
 
34
+ **Why a 0.5B model?** We deliberately started from a **very small** instruct model so the shipped artifact stays around **one download (~400 MB)** and can run **locally on mid-range phones** without a cloud API. LoRA SFT narrows that base model to one job—structured budget intents—not open-ended chat. Larger models can be smarter in general, but they are harder to justify for privacy-first, offline-friendly mobile installs.
35
+
36
  **Published artifact:** fused, quantized **GGUF only** — LoRA adapter weights are **not** in this repository.
37
 
38
  **Recommended download:** [`aibudget-intent-05b-v5-Q4_K_M.gguf`](https://huggingface.co/pablitoelpeligro/aibudget-intent-05b-GGUF/resolve/main/aibudget-intent-05b-v5-Q4_K_M.gguf) (~398 MB, SHA256 `d3d6fc0780359cd4d198e9354b9ecfc0ff9cfb0fb729f040326fa6e2f02c693a`)