le-harnais / ft-agentworld-8b

World-model student; near-teacher local replacement.

  • Base model: meta-llama/Meta-Llama-3.1-8B-Instruct โ€” Built with Llama; Llama Community License applies.
  • Class: hero
  • Training data: datasets/agentworld_distill_train.jsonl (160 ex; teacher=Qwen-AgentWorld-35B-A3B)
  • Headline: token-F1 0.958 / OBSERVATION hit-rate 95% vs teacher (n=40) โ€” near teacher-quality at ~16GB

Reproduce

BASE=meta-llama/Meta-Llama-3.1-8B-Instruct OUT=refs/llm-jepa/ft-agentworld-8b tools/distill_agentworld.sh

Full recipe, datasets, and eval commands: see docs/REPRODUCE.md in the [le-harnais distribution]. Provenance & license: docs/PROVENANCE.md.

Formats in this repo

  • *.safetensors โ€” bf16 inference weights (serve with transformers or le-harnais lh-serve/candle).
  • *.Q4_K_M.gguf โ€” portable 4-bit quant (run via ollama / llama.cpp; Mac-friendly).
  • *.Q8_0.gguf โ€” higher-fidelity 8-bit quant (hero models).

Orchestration amplifies a capable generator; it does not create competence.

Downloads last month
265
Safetensors
Model size
8B params
Tensor type
BF16
ยท
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for renaudb1999/le-harnais-ft-agentworld-8b

Quantized
(889)
this model