Skip to content

Site SEO: canonical URLs, preview card, JSON-LD, robots.txt, sitemap.xml, IndexNow - #54

Merged
bong-water-water-bong merged 1 commit into
mainfrom
seo
Sep 24, 2026
Merged

bong-water-water-bong merged 1 commit into
mainfrom
seo

Conversation

@bong-water-water-bong

Copy link
Copy Markdown
Collaborator

What 1bit.gg was missing for search engines and link previews (robots.txt and sitemap.xml both returned 404):

  • Canonical URL on every page (https://1bit.gg/...), so www and old links fold into one address.
  • Link previews: og:url, og:type (article for docs and posts), og:image with a new 1200x630 card (site/assets/og-card.png), and twitter:card = summary_large_image with title, description and image. Discord, X, Slack and Reddit show the card.
  • Structured data (JSON-LD): the home page is a WebSite + SoftwareSourceCode (repository, Apache-2.0, C++); each post is a BlogPosting with its dates.
  • robots.txt (allow all, points at the sitemap) and sitemap.xml (23 pages; the 404 page is left out and marked noindex). Each page's lastmod is its source file's last commit, so pages.yml now checks out the full history.
  • IndexNow: the site serves the key file, and after each deploy to main the workflow submits the sitemap's URLs to Bing, Yandex, Seznam and Naver (continue-on-error, so a failed ping never fails a deploy). The key is public by design.

Built locally: every page carries its canonical URL and card; the home page and posts carry valid JSON-LD; the sitemap lists 23 URLs.

Google needs one step outside the repo: verify 1bit.gg in Search Console (a DNS TXT record at Namecheap) and submit https://1bit.gg/sitemap.xml.

🤖 Generated with Claude Code

…emap.xml, IndexNow

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
@context7

context7 Bot commented Sep 24, 2026

Copy link
Copy Markdown

Docs7 for 1bit-monster/engine

Result Status Action
Deployment ➖ Not used —
Content review ✅ Passed. No problems found. View findings

Commit 940a136

@bong-water-water-bong
bong-water-water-bong merged commit 84280bd into main Sep 24, 2026
4 checks passed
bong-water-water-bong added a commit that referenced this pull request Oct 1, 2026
…de for Bonsai (#54) (#268)

- third_party/llama.cpp: cde002d -> dd74f6b, adding llama.cpp #53 and #54.
  - #53: IQ1_S and IQ1_M weights run on HRX0 (shared dequantizer, K-quant decode) instead of the CPU.
  - #54: exact-ternary Q4_0 weights decode from a 2-bit copy made at load, opt-in with
    GGML_HRX_TERNARY_Q4_0.
- 1bit serve sets GGML_HRX_TERNARY_Q4_0=1 for files stamped onebit.ternary_q4_0
  (tools/ternary_to_q4_0.py), unless the user set it. It prints that the packed copy costs about 30% of
  the file in extra GPU memory.
- docs/hrx.md: the IQ1 numbers, and the packed ternary path with its memory cost (+4 GiB for
  Bonsai-2-27B) and decode speed (9.3 -> 15.4 tok/s).
- registry/architectures.json regenerated for the pin.

Co-authored-by: bong-water-water-bong <bong-water-water-bong@1bit.gg>
Co-authored-by: Claude Opus 5.5 <noreply@anthropic.com>
bong-water-water-bong pushed a commit that referenced this pull request Oct 1, 2026
…3.8-27B pp512 98 -> 335

llama.cpp fork #55 routes Q4_K/Q5_K/IQ4_XS prompt matmuls to the existing
q8_1 x4 prefill kernel:
- Q5_K and IQ4_XS get Q4_K's activation/GLU policy.
- The generic fused SwiGLU (priority 290) leaves 256-2048-token chunks to it.
On Qwen3.8-27B UD-Q4_K_XL those projections had gone to the generic F32 WMMA
kernels, 94% of HRX prompt time.

Measured on this pin (6e42b51, which also carries #53/#54):
- llama-bench -b 512 -ub 512: pp512 98.5 -> 334.6, pp2048 98.1 -> 309.5 tok/s.
- 1bit serve --device hrx, 14,435-token prompt: 90.7 -> 264.7 tok/s, same text.
- KLD vs the BF16 logits (wikitext-2, 20 x 512): 0.00712, same top token 96.27%.
- test-backend-ops MUL_MAT on HRX0: 287/287.
GGML_HRX_Q8_PREFILL_RELAX=0 restores the old routing.

Docs: a docs/hrx.md section and "Our patches" entry. The docs/serve.md
auto-routing note ("97-99 tok/s until fork #55 is pinned") and the
docs/laya.md policy note now carry the measured numbers.

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>
bong-water-water-bong added a commit that referenced this pull request Oct 1, 2026
…3.8-27B pp512 98 -> 335 (#273)

* Pin llama.cpp 6e42b51: HRX prompt matmuls on the q8_1 x4 kernel, Qwen3.8-27B pp512 98 -> 335

llama.cpp fork #55 routes Q4_K/Q5_K/IQ4_XS prompt matmuls to the existing
q8_1 x4 prefill kernel:
- Q5_K and IQ4_XS get Q4_K's activation/GLU policy.
- The generic fused SwiGLU (priority 290) leaves 256-2048-token chunks to it.
On Qwen3.8-27B UD-Q4_K_XL those projections had gone to the generic F32 WMMA
kernels, 94% of HRX prompt time.

Measured on this pin (6e42b51, which also carries #53/#54):
- llama-bench -b 512 -ub 512: pp512 98.5 -> 334.6, pp2048 98.1 -> 309.5 tok/s.
- 1bit serve --device hrx, 14,435-token prompt: 90.7 -> 264.7 tok/s, same text.
- KLD vs the BF16 logits (wikitext-2, 20 x 512): 0.00712, same top token 96.27%.
- test-backend-ops MUL_MAT on HRX0: 287/287.
GGML_HRX_Q8_PREFILL_RELAX=0 restores the old routing.

Docs: a docs/hrx.md section and "Our patches" entry. The docs/serve.md
auto-routing note ("97-99 tok/s until fork #55 is pinned") and the
docs/laya.md policy note now carry the measured numbers.

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>

* registry: record llama.cpp (hrx) pin 6e42b51 (tools/registry_build.py; no mapping changes)

Co-Authored-By: Claude Opus 5.5 <noreply@anthropic.com>

---------

Co-authored-by: bong-water-water-bong <bong-water-water-bong@1bit.gg>
Co-authored-by: Claude Opus 5.5 <noreply@anthropic.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant