Instructions to use catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ with libraries, inference providers, notebooks, and local apps. Follow these links to get started.

Libraries

How to use catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ with MLX:

# Make sure mlx-lm is installed
# pip install --upgrade mlx-lm

# Generate text with mlx-lm
from mlx_lm import load, generate

model, tokenizer = load("catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ")

prompt = "Write a story about Einstein"
messages = [{"role": "user", "content": prompt}]
prompt = tokenizer.apply_chat_template(
    messages, add_generation_prompt=True
)

text = generate(model, tokenizer, prompt=prompt, verbose=True)

Notebooks
Google Colab
Kaggle
Local Apps Settings
LM Studio

How to use catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ with Pi:

Start the MLX server

# Install MLX LM:
uv tool install mlx-lm
# Start a local OpenAI-compatible server:
mlx_lm.server --model "catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ"

Configure the model in Pi

# Install Pi:
npm install -g @mariozechner/pi-coding-agent
# Add to ~/.pi/agent/models.json:
{
  "providers": {
    "mlx-lm": {
      "baseUrl": "http://localhost:8080/v1",
      "api": "openai-completions",
      "apiKey": "none",
      "models": [
        {
          "id": "catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ"
        }
      ]
    }
  }
}

Run Pi

# Start Pi in your project directory:
pi

Hermes Agent new

How to use catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ with Hermes Agent:

Start the MLX server

# Install MLX LM:
uv tool install mlx-lm
# Start a local OpenAI-compatible server:
mlx_lm.server --model "catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ"

Configure Hermes

# Install Hermes:
curl -fsSL https://hermes-agent.nousresearch.com/install.sh | bash
hermes setup
# Point Hermes at the local server:
hermes config set model.provider custom
hermes config set model.base_url http://127.0.0.1:8080/v1
hermes config set model.default catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ

Run Hermes

hermes

MLX LM

How to use catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ with MLX LM:

Generate or start a chat session

# Install MLX LM
uv tool install mlx-lm
# Interactive chat REPL
mlx_lm.chat --model "catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ"

Run an OpenAI-compatible server

# Install MLX LM
uv tool install mlx-lm
# Start the server
mlx_lm.server --model "catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ"
# Calling the OpenAI-compatible server with curl
curl -X POST "http://localhost:8000/v1/chat/completions" \
   -H "Content-Type: application/json" \
   --data '{
     "model": "catalystsec/Seed-OSS-36B-Instruct-4bit-DWQ",
     "messages": [
       {"role": "user", "content": "Hello"}
     ]
   }'

Seed-OSS-36B-Instruct-4bit-DWQ / chat_template.jinja

kernelpool

Add files using upload-large-folder tool

b2921ca verified 9 months ago

Raw

History Blame Contribute Delete

7.71 kB

	{# ----------‑‑‑ special token variables ‑‑‑---------- #}
	{%- set bos_token = '<seed:bos>' -%}
	{%- set eos_token = '<seed:eos>' -%}
	{%- set pad_token = '<seed:pad>' -%}
	{%- set toolcall_begin_token = '<seed:tool_call>' -%}
	{%- set toolcall_end_token = '</seed:tool_call>' -%}
	{%- set think_begin_token = '<seed:think>' -%}
	{%- set think_end_token = '</seed:think>' -%}
	{%- set budget_begin_token = '<seed:cot_budget_reflect>'-%}
	{%- set budget_end_token = '</seed:cot_budget_reflect>'-%}
	{# -------------- reflection-interval lookup -------------- #}
	{%- if not thinking_budget is defined %}
	{%- set thinking_budget = -1 -%}
	{%- endif -%}
	{%- set budget_reflections_v05 = {
	0: 0,
	512: 128,
	1024: 256,
	2048: 512,
	4096: 512,
	8192: 1024,
	16384: 1024
	} -%}
	{# Find the first gear that is greater than or equal to the thinking_budget. #}
	{%- set ns = namespace(interval = None) -%}
	{%- for k, v in budget_reflections_v05 \| dictsort -%}
	{%- if ns.interval is none and thinking_budget <= k -%}
	{%- set ns.interval = v -%}
	{%- endif -%}
	{%- endfor -%}
	{# If it exceeds the maximum gear, use the value of the last gear #}
	{%- if ns.interval is none -%}
	{%- set ns.interval = budget_reflections_v05[16384] -%}
	{%- endif -%}
	{# ---------- Preprocess the system message ---------- #}
	{%- if messages[0]["role"] == "system" %}
	{%- set system_message = messages[0]["content"] %}
	{%- set loop_messages = messages[1:] %}
	{%- else %}
	{%- set loop_messages = messages %}
	{%- endif %}
	{# ---------- Ensure tools exist ---------- #}
	{%- if not tools is defined or tools is none %}
	{%- set tools = [] %}
	{%- endif %}
	{# tools2doc.jinja #}
	{%- macro py_type(t) -%}
	{%- if t == "string" -%}str
	{%- elif t in ("number", "integer") -%}int
	{%- elif t == "boolean" -%}bool
	{%- elif t == "array" -%}list
	{%- else -%}Any{%- endif -%}
	{%- endmacro -%}
	{# ---------- Output the system block ---------- #}
	{%- if system_message is defined %}
	{{ bos_token + "system\n" + system_message }}
	{%- else %}
	{%- if tools is iterable and tools \| length > 0 %}
	{{ bos_token + "system\nYou are Doubao, a helpful AI assistant. You may call one or more functions to assist with the user query." }}
	{%- endif %}
	{%- endif %}
	{%- if use_json_tooldef is defined and use_json_tooldef %}

	{{"Tool List:\nYou are authorized to use the following tools (described in JSON Schema format). Before performing any task, you must decide how to call them based on the descriptions and parameters of these tools."}}
	{{ tools \| tojson(ensure_ascii=False) }}
	{%- else %}
	{%- for item in tools if item.type == "function" %}


	Function:
	def {{ item.function.name }}(
	{%- for name, spec in item.function.parameters.properties.items() %}
	{{- name }}: {{ py_type(spec.type) }}{% if not loop.last %},{% endif %}
	{%- endfor %}):
	"""
	{{ item.function.description \| trim }}

	{# ---------- Args ---------- #}
	{%- if item.function.parameters.properties %}
	Args:
	{%- for name, spec in item.function.parameters.properties.items() %}

	- {{ name }} ({{ py_type(spec.type) }})
	{%- if name in item.function.parameters.required %} [必填]{% else %} [选填]{% endif %}:
	{{- " " ~ (spec.description or "") }}
	{%- endfor %}
	{%- endif %}

	{# ---------- Returns ---------- #}
	{%- if item.function.returns is defined
	and item.function.returns.properties is defined
	and item.function.returns.properties %}
	Returns:
	{%- for name, spec in item.function.returns.properties.items() %}

	- {{ name }} ({{ py_type(spec.type) }}):
	{{- " " ~ (spec.description or "") }}
	{%- endfor %}
	{%- endif %}

	"""
	{%- endfor %}
	{%- endif %}
	{%- if tools is iterable and tools \| length > 0 %}

	{{"工具调用请遵循如下格式:\n<seed:tool_call>\n<function=example_function_name>\n<parameter=example_parameter_1>value_1</parameter>\n<parameter=example_parameter_2>This is the value for the second parameter\nthat can span\nmultiple lines</parameter>\n</function>\n</seed:tool_call>\n"}}
	{%- endif %}
	{# End the system block line #}
	{%- if system_message is defined or tools is iterable and tools \| length > 0 %}
	{{ eos_token }}
	{%- endif %}
	{# ---------- Thinking Budget ---------- #}
	{%- if thinking_budget is defined %}
	{%- if thinking_budget == 0 %}
	{{ bos_token+"system" }}
	{{ "You are an intelligent assistant that can answer questions in one step without the need for reasoning and thinking, that is, your thinking budget is 0. Next, please skip the thinking process and directly start answering the user's questions." }}
	{{ eos_token }}
	{%- elif not thinking_budget == -1 %}
	{{ bos_token+"system" }}
	{{ "You are an intelligent assistant with reflective ability. In the process of thinking and reasoning, you need to strictly follow the thinking budget, which is "}}{{thinking_budget}}{{". That is, you need to complete your thinking within "}}{{thinking_budget}}{{" tokens and start answering the user's questions. You will reflect on your thinking process every "}}{{ns.interval}}{{" tokens, stating how many tokens have been used and how many are left."}}
	{{ eos_token }}
	{%- endif %}
	{%- endif %}
	{# ---------- List the historical messages one by one ---------- #}
	{%- for message in loop_messages %}
	{%- if message.role == "assistant"
	and message.tool_calls is defined
	and message.tool_calls is iterable
	and message.tool_calls \| length > 0 %}
	{{ bos_token + message.role }}
	{%- if message.reasoning_content is defined and message.reasoning_content is string and message.reasoning_content \| trim \| length > 0 %}
	{{ "\n" + think_begin_token + message.reasoning_content \| trim + think_end_token }}
	{%- endif %}
	{%- if message.content is defined and message.content is string and message.content \| trim \| length > 0 %}
	{{ "\n" + message.content \| trim + "\n" }}
	{%- endif %}
	{%- for tool_call in message.tool_calls %}
	{%- if tool_call.function is defined %}{% set tool_call = tool_call.function %}{% endif %}
	{{ "\n" + toolcall_begin_token + "\n<function=" + tool_call.name + ">\n" }}
	{%- if tool_call.arguments is defined %}
	{%- for arg_name, arg_value in tool_call.arguments \| items %}
	{{ "<parameter=" + arg_name + ">" }}
	{%- set arg_value = arg_value if arg_value is string else arg_value \| string %}
	{{ arg_value+"</parameter>\n" }}
	{%- endfor %}
	{%- endif %}
	{{ "</function>\n" + toolcall_end_token }}
	{%- endfor %}
	{{ eos_token }}
	{%- elif message.role in ["user", "system"] %}
	{{ bos_token + message.role + "\n" + message.content + eos_token }}
	{%- elif message.role == "assistant" %}
	{{ bos_token + message.role }}
	{%- if message.reasoning_content is defined and message.reasoning_content is string and message.reasoning_content \| trim \| length > 0 %}
	{{ "\n" + think_begin_token + message.reasoning_content \| trim + think_end_token }}
	{%- endif %}
	{%- if message.content is defined and message.content is string and message.content \| trim \| length > 0 %}
	{{ "\n" + message.content \| trim + eos_token }}
	{%- endif %}
	{# Include the tool role #}
	{%- else %}
	{{ bos_token + message.role + "\n" + message.content + eos_token }}
	{%- endif %}
	{%- endfor %}
	{# ---------- Control the model to start continuation ---------- #}
	{%- if add_generation_prompt %}
	{{ bos_token+"assistant\n" }}
	{%- if thinking_budget == 0 %}
	{{ think_begin_token + "\n" + budget_begin_token + "The current thinking budget is 0, so I will directly start answering the question." + budget_end_token + "\n" + think_end_token }}
	{%- endif %}
	{%- endif %}