KikoCis commited on
Commit
8d52e0c
Β·
verified Β·
1 Parent(s): 04d9e0e

card v2: banner + spec-sheet header + preservation story

Browse files
Files changed (1) hide show
  1. README.md +68 -11
README.md CHANGED
@@ -1,21 +1,78 @@
1
  ---
2
  license: mit
3
- tags: [preserved, repository-exploration, subagent, coder, agentic, qwen3, 256k, long-context]
4
- language: [en]
 
 
 
 
 
 
 
 
 
 
5
  pipeline_tag: text-generation
6
  ---
7
 
8
- # FastContext-1.0-4B-SFT (preserved original weights)
9
 
10
- **Preserved copy of Microsoft's FastContext-1.0-4B-SFT**, which Microsoft **deleted from both HuggingFace and GitHub** (verified: 404 on both) about two weeks after open-sourcing it under MIT. Re-uploaded here so the weights stay available. **Own your AI.**
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
11
 
12
- ## What it is
13
- A **repository-exploration subagent** for coding agents: invoked on-demand by your main agent, it fires **parallel read-only tool calls (READ / GLOB / GREP)** across a repo and returns **only the file paths + line ranges you need** as focused context β€” offloading file discovery so your coding agent burns far fewer tokens. Microsoft's announcement reported **~60% fewer tokens** and **+5.5% SWE-bench** *(their figures; source now deleted)*.
14
 
15
- Architecture: plain **Qwen3 dense 4B** (`Qwen3ForCausalLM`, 36 layers, **256K context**, MIT).
16
 
17
- ## Quantized GGUF
18
- Long-context-imatrix GGUFs (any llama.cpp backend): **KikoCis/FastContext-1.0-4B-longctx-imatrix-GGUF**.
19
 
20
- ## Credit
21
- Original Β© Microsoft, MIT license. This is an unmodified preservation mirror (weights unchanged).
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
 
1
  ---
2
  license: mit
3
+ library_name: transformers
4
+ tags:
5
+ - preserved
6
+ - repository-exploration
7
+ - subagent
8
+ - coder
9
+ - agentic
10
+ - qwen3
11
+ - 256k
12
+ - long-context
13
+ language:
14
+ - en
15
  pipeline_tag: text-generation
16
  ---
17
 
18
+ ![banner](banner.png)
19
 
20
+ <div style="border:2px solid currentColor; font-family:ui-monospace,'SF Mono','Cascadia Mono',Consolas,'Liberation Mono',monospace;">
21
+ <div style="border-bottom:1px solid currentColor; padding:6px 12px; font-size:11px; letter-spacing:3px; text-transform:uppercase; opacity:0.7; text-align:center;">PRESERVED ORIGINAL // REMOVED BY MICROSOFT FROM HF + GITHUB // MIT</div>
22
+ <div style="padding:14px; display:flex; flex-wrap:wrap; align-items:center; justify-content:center; gap:18px;">
23
+ <pre style="margin:0; flex:0 0 auto; font-family:ui-monospace,'SF Mono','Cascadia Mono',Consolas,monospace; font-size:9px; line-height:1.15; letter-spacing:0;">
24
+ microsoft/FastContext ──▢ 404
25
+ github.com/microsoft/FastContext ──▢ 404
26
+ β”‚
27
+ β–Ό
28
+ β”Œβ”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”
29
+ β”‚ weights preserved here β”‚
30
+ β”‚ bf16 Β· 8.0 GB Β· intact β”‚
31
+ β””β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”€β”˜
32
+ you can't un-open-source
33
+ </pre>
34
+ <div style="flex:0 1 auto; max-width:100%; text-align:center;">
35
+ <div style="font-size:23px; font-weight:800; letter-spacing:1px;">FASTCONTEXT-1.0-4B-SFT</div>
36
+ <div style="font-size:12.5px; letter-spacing:1px; opacity:0.8; margin-top:5px;"><span style="white-space:nowrap;">PRESERVED ORIGINAL WEIGHTS</span> Β· <span style="white-space:nowrap;">QWEN3 DENSE 4B</span> Β· <span style="white-space:nowrap;">256K CONTEXT</span> Β· <span style="white-space:nowrap;">BF16 Β· 8.0 GB</span></div>
37
+ </div>
38
+ </div>
39
+ <table style="display:table; table-layout:fixed; width:100%; margin:0; border-collapse:collapse; font-family:ui-monospace,'SF Mono',Consolas,monospace; font-size:12px;">
40
+ <tr>
41
+ <td style="border-top:1px solid currentColor; border-right:1px solid currentColor; padding:8px 12px;"><div style="font-size:10px; letter-spacing:1px; opacity:0.6;">WEIGHTS</div><div style="font-weight:700;">BF16 Β· UNMODIFIED</div></td>
42
+ <td style="border-top:1px solid currentColor; border-right:1px solid currentColor; padding:8px 12px;"><div style="font-size:10px; letter-spacing:1px; opacity:0.6;">ARCH</div><div style="font-weight:700;">QWEN3 DENSE Β· 36L</div></td>
43
+ <td style="border-top:1px solid currentColor; border-right:1px solid currentColor; padding:8px 12px;"><div style="font-size:10px; letter-spacing:1px; opacity:0.6;">CONTEXT</div><div style="font-weight:700;">256K NATIVE</div></td>
44
+ <td style="border-top:1px solid currentColor; padding:8px 12px;"><div style="font-size:10px; letter-spacing:1px; opacity:0.6;">LICENSE</div><div style="font-weight:700;">MIT</div></td>
45
+ </tr>
46
+ </table>
47
+ </div>
48
 
49
+ > Microsoft open-sourced FastContext under MIT, then **deleted it from both HuggingFace and GitHub** about two weeks later (verified: 404 on both, 2026-07-02). MIT means preservation is legal β€” so here it is, unmodified. **Own your AI: a model on your disk can't be sunset by a quarterly review.**
 
50
 
51
+ ## πŸ” What it is
52
 
53
+ A **repository-exploration subagent** for coding agents. Invoked on demand by your main agent, it fires **parallel read-only tool calls** (`READ` / `GLOB` / `GREP`) across a repo and returns **only the file paths + line ranges that matter**, as compact context. Your frontier coding agent stops wasting its context window (and your bill) crawling the file tree.
 
54
 
55
+ Microsoft's (now-deleted) announcement reported **~60% fewer tokens** from the main coding agent and **+5.5% on SWE-bench** β€” their figures; the source no longer exists to cite.
56
+
57
+ **Architecture**: plain `Qwen3ForCausalLM` dense 4B β€” 36 layers, 256K native context. No exotic modules; loads with standard `transformers`.
58
+
59
+ ## πŸš€ Quick start
60
+
61
+ ```python
62
+ from transformers import AutoModelForCausalLM, AutoTokenizer
63
+ m = AutoModelForCausalLM.from_pretrained("KikoCis/FastContext-1.0-4B-SFT", torch_dtype="bfloat16", device_map="auto")
64
+ tok = AutoTokenizer.from_pretrained("KikoCis/FastContext-1.0-4B-SFT")
65
+ ```
66
+
67
+ **Don't want 8 GB?** Grab the **GGUF quants** (1.96–2.5 GB, long-context imatrix, retrieval-validated 30/30 vs this bf16):
68
+ πŸ‘‰ [KikoCis/FastContext-1.0-4B-longctx-imatrix-GGUF](https://huggingface.co/KikoCis/FastContext-1.0-4B-longctx-imatrix-GGUF)
69
+
70
+ ## ⚠️ Good to know
71
+
72
+ - It's a **scout, not a solver** β€” it finds and returns evidence; pair it with a main coding agent that writes the actual fix.
73
+ - Upstream docs, harness code and issues were deleted along with the repos; usage conventions here come from the announcement and community mirrors.
74
+ - Weights are **byte-identical** to the (re-uploaded) original β€” no fine-tuning, no edits.
75
+
76
+ ## πŸ“š Credit & license
77
+
78
+ Model, weights, training: **Β© Microsoft** (MIT). This is a preservation mirror sourced via the `ShaunGves/FastContext-1.0-4B-SFT` re-upload after `microsoft/FastContext-1.0-4B-SFT` was removed. Nothing modified. Quantized companion + validation: KikoCis.