mr-r0b0t - r0b0tlab
- 52 followers
- United States of America
- @mr_r0b0t
Popular repositories Loading
-
hermes-concurrent-agents
hermes-concurrent-agents PublicDeploy concurrent Hermes Agent workers on unified-memory GPUs (GB10, DGX Spark) for maximum total tok/s. Profile-isolated, kanban-coordinated, crash-recovering.
-
llm-wiki_obsidian_hermes_r0b0tlabbra1n
llm-wiki_obsidian_hermes_r0b0tlabbra1n PublicFilesystem-first LLM-Wiki + Obsidian + Hermes Agent memory system. Markdown source of truth, SQLite FTS5 search, secret scanning, tier-based memory. Built for local LLM setups.
-
qwen36-35b-a3b-nvfp4-gb10-native-mtp
qwen36-35b-a3b-nvfp4-gb10-native-mtp PublicGB10 NVFP4 native MTP reproducibility pack for Qwen3.6-35B-A3B
-
minimax-m27-nvfp4-gb10-benchmark
minimax-m27-nvfp4-gb10-benchmark PublicMiniMax M2.7 NVFP4 dual-GB10 Blackwell benchmark: vLLM FlashInfer-CUTLASS, public data, HTML canvas report, and Docker runtime.
HTML 13
-
deepseek-v4-flash-nvfp4-gb10-benchmark
deepseek-v4-flash-nvfp4-gb10-benchmark PublicDeepSeek-V4-Flash native Blackwell FP8 benchmark on dual DGX Spark GB10 (TP=2, MTP, RoCE). c=1 38.4 t/s, c=16 144.6 t/s aggregate.
Python 10
-
nvidia-qwen-3.6-27B-sm121-nvfp4
nvidia-qwen-3.6-27B-sm121-nvfp4 PublicNVIDIA Qwen3.6-27B NVFP4 on SM121 (GB10) — vLLM v0.24.0 with native NVFP4 KV cache via FlashInfer FA2 JIT. 67% more KV capacity than FP8.
Python 7
Repositories
- qwen3.8-max-dossier Public
Qwen3.8-Max launch-day dossier — a self-portrait HTML artifact built by the model it documents. Official Qwen sources only.
- nvidia-qwen-3.6-27B-sm121-nvfp4 Public
NVIDIA Qwen3.6-27B NVFP4 on SM121 (GB10) — vLLM v0.24.0 with native NVFP4 KV cache via FlashInfer FA2 JIT. 67% more KV capacity than FP8.
- hermes-alibaba-token-plan Public Forked from oliver-mee/hermes-alibaba-token-plan
Hermes Agent model-provider plugin for Alibaba Cloud Token Plan (Global + China)
- dgx-spark-playbooks Public Forked from NVIDIA/dgx-spark-playbooks
Collection of step-by-step playbooks for setting up AI/ML workloads on NVIDIA DGX Spark devices with Blackwell architecture.
- tldraw-skill Public
Production-grade Hermes Agent skill for tldraw 5.2.5+ development, migration, sync, automation, and evaluation
- hermes-concurrent-agents Public
Deploy concurrent Hermes Agent workers on unified-memory GPUs (GB10, DGX Spark) for maximum total tok/s. Profile-isolated, kanban-coordinated, crash-recovering.
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…