Skip to content
View hizrianraz's full-sized avatar
🚀
To Infinity and Beyond!
🚀
To Infinity and Beyond!

Highlights

  • Pro

Block or report hizrianraz

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
.github/profile/README.md
    __  ______
   / / / / __ \
  / /_/ / /_/ /
 / __  / _, _/
/_/ /_/_/ |_|

Hizrian Raz

Philosopher · Wanderer · Neurodivergent · Founder

Building Ainfera — an AI-Native model factory.
On the side: personal deployment packs for one box. Only Laguna is measured; receipts over hype.


Ainfera LinkedIn X Hugging Face GitHub Website


Typing SVG


About

I'm a founder who cares more about what actually runs than what trends.

  • I build Ainfera — an AI-Native model factory (company work lives under @ainfera-ai).
  • I publish personal, unaffiliated measurement packs so one DGX Spark stays reproducible.
  • I optimize for evidence, clarity, and long-horizon craft — not launch theater.

Packs here are not company IP, not finetunes, and not invented numbers.


Skills

Domain What I actually do
Model systems Serve paths, quant honesty, agentic runbooks, eval receipts
Inference craft llama.cpp / GGUF pins, FP8 deployment previews, single-node MoE fit research
Agent loops Tool-use harnesses, format/routing smoke, hermes-style checks
Evidence discipline Digests, freeze clocks, SAQS claim binder, smoke ≠ headline
Product eng TypeScript surfaces, Python toolchains, shell pack automation
Founder ops Spec → measure → ship, ADHD-aware execution, conviction over consensus

Python TypeScript Shell PyTorch Hugging Face llama.cpp GGUF CUDA vLLM Docker Linux macOS Git GitHub VS Code


Focus right now

Personal deployment stack on one DGX Spark (GB10 · 128 GB). Only Laguna currently has family measurement authority.

Scripts · pins · cards · eval receipts.
Not finetunes. Not company IP. Not stretch claims.

# Pack Role (honest) Status
01 Laguna-S-2.1 · HF Repo-maintenance deployment target · official Q4_K_M Historical measured path · source candidate · freeze pending
02 Qwen3-Coder-Next · HF Interactive coding · official FP8 Preview · exact runtime model-load pending
03 DeepSeek-V4-Flash REAP25 · HF Research pointer · REAP25/Pulsar reference Docs-only HOLD / NO_HERO

Laguna-S Spark Qwen3-Coder-Next DeepSeek REAP25

HF models

Size · pins · windows

Window (WIB)

List target 2026-08-03 12:00 · freeze target 2026-08-02 18:00
Release stays blocked until the August 2 freeze attestation passes across source, Hub, runtime safety, and claims.

Hero CTAs

Laguna Aug 3 20:00 only after explicit freeze clearance · Qwen null until a dated :8001 remeasurement and explicit gate · DeepSeek null (NO_HERO)

Laguna day-0 measured artifact

poolside/Laguna-S-2.1-GGUF · laguna-s-2.1-Q4_K_M.gguf
sha256 a8b55c75714ea73fd90ec85de5defdc0b8d88ca0ad2108343cdd8fc22f7583e4
engine pin 04b2b72 (poolsideai/llama.cpp, branch laguna) · measure tip bf82eab
format/routing smoke 40/40 · Hermes 27/27 validated-not-executed · ~21.47 t/s gen128
not long-horizon agent reliability proof · DFlash/NVFP4 not day-0 flagship

Official DeepSeek DSpark footprint reference

deepseek-ai/DeepSeek-V4-Flash-DSpark at revision 62af8fffb2f7030cac4de2f0169f5b8d1101b646: weight files total 166,886,535,336 bytes (155.425198 GiB). Repository-file accounting only; not a runtime measurement.

Upstream / quality refs (not day-0 serve authority)

poolside/Laguna-S-2.1-NVFP4 · Qwen/Qwen3-Coder-Next-FP8 · twaggs88 REAP25 DSpark GGUF · bases Laguna-S-2.1 · Qwen3-Coder-Next · DeepSeek-V4-Flash


Operating system

measure  →  pin  →  publish receipts  →  refuse the stretch claim
Principle In practice
Conviction over consensus One box. Real digests. Reproducible scripts.
Verify the premise Smoke ≠ headline. Verifier ≠ gate clearance.
Honest labels Experimental stays experimental until dated measure lands.
Integrity is structural Claim binder ships with the pack — not a blog afterthought.

Public claim binder for every Spark pack:
SPARK_AGENTIC_QUANT_STANDARD.md (SAQS)

Honesty locks — always on
  • diy_gguf = false — Laguna GGUF mirror hosts official Poolside bytes only
  • public_promo_before_launch = false
  • smoke ≠ agent headline · evidence-bind ≠ gate
  • Qwen day-0 = official-FP8 Preview, not a throughput, quality, retention, or tool-success claim
  • DeepSeek day-0 = docs-only, unmeasured HOLD / NO_HERO scaffold
  • DSpark ≠ DGX Spark · peer GGUF ≠ the official DSpark artifact; listing is not a fit claim
  • Packs are not Ainfera product surfaces, company eval, sales SKUs, or open finetunes

Reproduce a pack

1  read README + INSTALL.yaml
2  pull the named upstream  (or Laguna official GGUF + SHA256SUMS)
3  run Laguna only after its freeze + target receipt gates; run Qwen only as an explicit Preview; keep DeepSeek disabled
4  diff results/            — smoke ≠ headline

Surface

Static links only — no third-party stats widgets (those CDN pins break and show empty images).

Public repos Ainfera org LinkedIn HF X

Track Home
Personal packs github.com/hizrianraz
Company ainfera.ai · @ainfera-ai
LinkedIn linkedin.com/in/hizrian-raz
Models huggingface.co/hizrianraz
Profile binder SPARK_AGENTIC_QUANT_STANDARD.md

Elsewhere

Company ainfera.ai · @ainfera-ai
LinkedIn linkedin.com/in/hizrian-raz
Models huggingface.co/hizrianraz
Pack questions open a discussion on the relevant HF repo

Thanks for reading this far.

Personal · unaffiliated measurement surface · not Ainfera product IP Polished 2026-07-30 · Laguna measure authority bf82eab

Pinned Loading

  1. Laguna-S-2.1-Spark-Agentic Laguna-S-2.1-Spark-Agentic Public

    Personal measured Laguna-S-2.1 pack · Spark-only full MoE · Q4 headline · promo 2026-08-03 12:00 WIB

    Python 1

  2. Qwen3-Coder-Next-Spark-Agentic Qwen3-Coder-Next-Spark-Agentic Public

    Spark stand-behind agentic pack — Qwen3-Coder-Next official Q4_K_M (diy_gguf=false). Pull/serve scripts + smoke harness. Companion to Laguna S+XS Aug 3 set.

    Python

  3. DeepSeek-V4-Flash-Spark-Agentic DeepSeek-V4-Flash-Spark-Agentic Public

    Card-only hold scaffold for DeepSeek-V4-Flash on DGX Spark (personal, unmeasured)

    Python