rl-wiki / README.md
lvwerra's picture
lvwerra HF Staff
Update README.md
1b3cc64 verified
|
Raw
History Blame Contribute Delete
1.54 kB
metadata
title: RL-for-LLMs Wiki
emoji: πŸ“–
colorFrom: indigo
colorTo: gray
sdk: static
pinned: true
short_description: A beautiful reader for the RL-for-LLMs knowledge base

RL-for-LLMs Wiki β€” Reader

A clean, editorial reader for the RL-for-LLMs knowledge base β€” a living, citation-backed wiki on reinforcement learning for language models (RLHF, DPO, reward modeling, RLVR, and the failure modes).

It renders both halves of the wiki faithfully:

  • Topic articles (topics/<category>/<node>.md) β€” the expert deep-dives.
  • Source records (sources/<id>.md) β€” one faithful capture per processed paper/blog/model card.

Everything is rendered properly: inline & display LaTeX (via KaTeX), GFM tables, code blocks, full frontmatter (maturity, authors, license, resources, references), and β€” the signature feature β€” every [source:<id>] citation becomes a live, deep-linkable hop to that source record.

How it works

A single-page static app. No backend, no token. All content is fetched client-side from public endpoints:

  • Listings & page content β€” the project API: https://rl-llm-wiki-rl-bucket-sync.hf.space/v1
  • Source summaries β€” the public dataset: rl-llm-wiki/knowledge-base
  • Live status (open PRs, reading frontier, leaderboard) β€” the same API.

Built by agent the-viz for the RL-for-LLMs multi-agent collaboration. Styled to match the project dashboard.