AudioJev / Notice
shlv's picture
Keep only standard model assets and documentation
d506a7b verified
Raw History Blame Contribute Delete
1.14 kB
Qwen is licensed under the Qwen RESEARCH LICENSE AGREEMENT, Copyright (c) Alibaba Cloud. All Rights Reserved.
Built with Qwen.
AudioJev is a derivative of Qwen/Qwen2.5-Omni-3B. The saved Thinker weights were modified through three-stage fine-tuning of the audio and language decision path, using random-derangement paired candidate cross-entropy and aligned symmetric KL (lambda=0.5, seed=20261001, 4096 updates). The visual branch was excluded from training; the Talker is not included.
Modified weight files relative to the upstream pretrained model:
model-00001-of-00005.safetensors
model-00002-of-00005.safetensors
model-00003-of-00005.safetensors
model-00004-of-00005.safetensors
model-00005-of-00005.safetensors
These shards are byte-identical to the evaluated AudioJev checkpoint. The model/configuration and shard layout were saved by the training pipeline. The release corrects the optional total_parameters index metadata to the number of elements actually stored; tensor data and the weight map are unchanged. README.md documents this derivative. Inference software is maintained separately in AudioJev-Inference.