Instructions to use TzvikaJ/QWEN_Dicta with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Notebooks
- Google Colab
- Kaggle
- Local Apps Settings
- llama.cpp
How to use TzvikaJ/QWEN_Dicta with llama.cpp:
Install (macOS, Linux)
curl -LsSf https://llama.app/install.sh | sh # Start a local OpenAI-compatible server with a web UI: llama serve -hf TzvikaJ/QWEN_Dicta:F16 # Run inference directly in the terminal: llama cli -hf TzvikaJ/QWEN_Dicta:F16
Install from WinGet (Windows)
winget install llama.cpp # Start a local OpenAI-compatible server with a web UI: llama serve -hf TzvikaJ/QWEN_Dicta:F16 # Run inference directly in the terminal: llama cli -hf TzvikaJ/QWEN_Dicta:F16
Use pre-built binary
# Download pre-built binary from: # https://github.com/ggerganov/llama.cpp/releases # Start a local OpenAI-compatible server with a web UI: ./llama-server -hf TzvikaJ/QWEN_Dicta:F16 # Run inference directly in the terminal: ./llama-cli -hf TzvikaJ/QWEN_Dicta:F16
Build from source code
git clone https://github.com/ggerganov/llama.cpp.git cd llama.cpp cmake -B build cmake --build build -j --target llama-server llama-cli # Start a local OpenAI-compatible server with a web UI: ./build/bin/llama-server -hf TzvikaJ/QWEN_Dicta:F16 # Run inference directly in the terminal: ./build/bin/llama-cli -hf TzvikaJ/QWEN_Dicta:F16
Use Docker
docker model run hf.co/TzvikaJ/QWEN_Dicta:F16
- LM Studio
- Jan
- Ollama
How to use TzvikaJ/QWEN_Dicta with Ollama:
ollama run hf.co/TzvikaJ/QWEN_Dicta:F16
- Unsloth Desktop
- Docker Model Runner
How to use TzvikaJ/QWEN_Dicta with Docker Model Runner:
docker model run hf.co/TzvikaJ/QWEN_Dicta:F16
- Lemonade
How to use TzvikaJ/QWEN_Dicta with Lemonade:
Pull the model
# Download Lemonade from https://lemonade-server.ai/ lemonade pull TzvikaJ/QWEN_Dicta:F16
Run and chat with the model
lemonade run user.QWEN_Dicta-F16
List all available models
lemonade list
- Atomic Chat
You need to agree to share your contact information to access this model
This repository is publicly accessible, but you have to accept the conditions to access its files and content.
המודלים האלה הם חלק מתוסף ה-AI של VidQuality Duplia, והגישה אליהם ניתנת ללקוחות שרכשו את התוכנה. נא להזין את מספר ההזמנה ואת כתובת המייל שבה בוצעה הרכישה. כל בקשה נבדקת ידנית ומאושרת רק כששני הפרטים מתאימים להזמנה אמיתית. בקשה לא נדחית בגלל טעות הקלדה: אם משהו לא מתאים היא נשארת ממתינה ונחזור אליך.
These weights are part of the VidQuality Duplia AI add-on and are provided to customers who purchased the product. Please enter the order number and the e-mail address used for the purchase. Requests are reviewed manually and approved only when both fields match a real order. A request is never rejected over a typo - if something does not match we leave it pending and get in touch.
Log in or Sign Up to review the conditions and access this model content.
Omni Grounding — VidQuality
עברית
מה זה. כיוונון של המודל Qwen/Qwen2.5-Omni-7B, מכווץ ל-Q4_K_M. התוכנה VidQuality Duplia משתמשת בו כדי לתאר מה באמת רואים בסרטון, וכך לאפשר חיפוש לפי משמעות במקום לפי שם הקובץ.
מה זה לא. אלה אינם הקבצים המקוריים של Alibaba. המשקולות שונו בכיוונון ולאחר מכן כווצו. המודל הבסיסי מורשה תחת Apache-2.0, וסעיף 4(b) של הרישיון מחייב שקובץ ששונה יישא הודעה בולטת על כך — והפסקה הזאת היא ההודעה.
הקובץ הנלווה mmproj-Qwen2.5-Omni-7B-f16.gguf הוא מקרן הווידאו של אותו מודל בסיס,
שהומר ל-GGUF בלי שינוי במשקולות.
הקבצים. omni-grounding-Q4_K_M.gguf הוא המודל המכוונן והמכווץ;
mmproj-Qwen2.5-Omni-7B-f16.gguf הוא המקרן, משקולות ללא שינוי.
רישיון. Apache-2.0, בירושה מהמודל הבסיסי. עותק של הרישיון נשלח עם המוצר וכלול גם במאגר הזה. ייחוס: Qwen2.5-Omni-7B, Alibaba Cloud.
למי זה מיועד. ללקוחות VidQuality Duplia. התוסף מוריד את הקבצים לתוך התוכנה; הם אינם מיועדים לשימוש עצמאי.
English
A fine-tune of Qwen/Qwen2.5-Omni-7B, quantised to Q4_K_M, used by VidQuality Duplia to describe what a video actually shows so that videos can be searched by meaning rather than by file name.
What this is, and what it is not
This is a modified work. The weights were changed by fine-tuning and then quantised; they are not Alibaba's original files. The base model is licensed Apache-2.0, and section 4(b) of that licence requires modified files to carry a prominent notice that they were changed — this section is that notice.
The companion file mmproj-Qwen2.5-Omni-7B-f16.gguf is the multimodal projector of the
same base model, converted to GGUF with the weights unchanged.
Files
| file | what it is |
|---|---|
omni-grounding-Q4_K_M.gguf |
the fine-tuned model, quantised |
mmproj-Qwen2.5-Omni-7B-f16.gguf |
the projector, unchanged weights |
Licence
Apache-2.0, inherited from the base model. A copy of the licence ships with the product and is included in this repository. Attribution: Qwen2.5-Omni-7B, Alibaba Cloud.
Who this is for
Customers of VidQuality Duplia. The add-on downloads these files into the product; they are not intended to be used on their own.
- Downloads last month
- -
4-bit
Model tree for TzvikaJ/QWEN_Dicta
Base model
Qwen/Qwen2.5-Omni-7B