Hugging Face's logo Hugging Face
  • Models
  • Datasets
  • Spaces
  • Buckets new
  • Docs
  • Enterprise
  • Pricing
    • Website
      • Tasks
      • HuggingChat
      • Collections
      • Languages
      • Organizations
    • Community
      • Blog
      • Posts
      • Daily Papers
      • Hardware
      • Learn
      • Discord
      • Forum
      • GitHub
    • Solutions
      • Team & Enterprise
      • Hugging Face PRO
      • Enterprise Support
      • Inference Providers
      • Inference Endpoints
      • Storage Buckets

  • Log In
  • Sign Up

nicolasembleton
/
Spark-X2.5-4B-onnx

Text Generation
ONNX
webgpu
spark2_5
conversational
Model card Files Files and versions
xet
Community
Spark-X2.5-4B-onnx
23.5 GB
Ctrl+K
Ctrl+K
  • 1 contributor
History: 11 commits
nicolasembleton's picture
nicolasembleton
README: int4 is the quantized variant; drop stale int8 wording
1973da1 verified 4 days ago
  • onnx
    Remove broken int8 graphs (DynamicQuantizeLinear incompatible with fp16); int4 is the quantized variant 4 days ago
  • .gitattributes
    50.8 kB
    Upload folder using huggingface_hub 4 days ago
  • README.md
    1.69 kB
    README: int4 is the quantized variant; drop stale int8 wording 4 days ago
  • chat_template.jinja
    4.65 kB
    Upload folder using huggingface_hub 4 days ago
  • export_config.json
    2.75 kB
    Upload folder using huggingface_hub 4 days ago
  • special_tokens_map.json
    613 Bytes
    Upload folder using huggingface_hub 4 days ago
  • tokenizer.json
    10.1 MB
    Upload folder using huggingface_hub 4 days ago
  • tokenizer_config.json
    20.8 kB
    Upload folder using huggingface_hub 4 days ago