Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Avifenesh
/
memra-bench
like
3
GGUF
benchmark
cuda
speculative-decoding
nvfp4
License:
mit
Model card
Files
Files and versions
xet
Community
Copy to bucket
new
main
memra-bench
10.9 GB
Ctrl+K
Ctrl+K
1 contributor
History:
44 commits
Avifenesh
CONFIGS: release-train tail extended v0.101.0 -> v0.103.0 (v0.102.0 ornith vision, v0.103.0 serve-doors train)
a55df82
verified
1 day ago
drafts
Upload drafts/kf4/qwen36-35b-a3b-iq4xs/owngen-ranks-32768.txt with huggingface_hub
about 1 month ago
prompts
Upload prompts/p3-agentic-long-v2.txt with huggingface_hub
about 1 month ago
trims
Upload trims/frspec-corpus-32768.gguf with huggingface_hub
about 1 month ago
.gitattributes
2.66 kB
Upload drafts/kf4/qwen36-35b-a3b-iq4xs/draft-owntrim-nvfp4head-q4blk.gguf with huggingface_hub
about 1 month ago
CONFIGS.md
9.63 kB
CONFIGS: release-train tail extended v0.101.0 -> v0.103.0 (v0.102.0 ornith vision, v0.103.0 serve-doors train)
1 day ago
README.md
8.5 kB
board update: v0.101.0 headline cell (DSpark route + DFlash2 drafter, RTX PRO 6000, c=1 — 3.57 tok/round @ 13.7 ms/round ~= 260 tok/s decode-class greedy; bench context, opt-in route, cited from the release receipts)
1 day ago
mtp-Qwen3.6-27B-Q4_K_M-frspec-balanced32768.gguf
1.13 GB
xet
Upload folder using huggingface_hub
about 2 months ago
mtp-Qwen3.6-27B-Q4_K_M-frspec-code75-32768.gguf
1.13 GB
xet
Upload folder using huggingface_hub
about 2 months ago
mtp-Qwen3.6-27B-Q4_K_M-frspec32768.gguf
1.13 GB
xet
Upload folder using huggingface_hub
about 2 months ago