Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Nurgalyev Shakhizat
shakhizat
13
10
20
Follow
mamyrbek's profile picture
johnnynv's profile picture
GO1984's profile picture
4 followers
·
65 following
AI & ML interests
None yet
Recent Activity
new
activity
14 days ago
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4:
ValueError: Model config must specify `dflash_config.mask_token_id`
liked
a model
25 days ago
RedHatAI/Kimi-K3-NVFP4
new
activity
25 days ago
RedHatAI/GLM-5.2-speculator.dspark:
Failed to run on 8xB200 - AssertionError: Tried to load weights of size torch.Size([6144, 30720])to a parameter of size torch.Size([6144, 18432])
View all activity
Organizations
shakhizat
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
nvidia/NVIDIA-Nemotron-3.5-Lightning-30B-A3B-NVFP4
14 days ago
ValueError: Model config must specify `dflash_config.mask_token_id`
4
#3 opened 14 days ago by
shakhizat
New activity in
RedHatAI/GLM-5.2-speculator.dspark
25 days ago
Failed to run on 8xB200 - AssertionError: Tried to load weights of size torch.Size([6144, 30720])to a parameter of size torch.Size([6144, 18432])
9
#1 opened about 1 month ago by
shakhizat
New activity in
unsloth/Qwen3.6-35B-A3B-NVFP4
about 1 month ago
ValueError: moe_backend='flashinfer_b12x' is not supported for FP8 MoE.
3
#5 opened about 1 month ago by
shakhizat
New activity in
nvidia/Qwen3.6-27B-NVFP4
about 2 months ago
[need help] vllm 0.24.0 on dgx spark warn Your GPU does not have native support for FP4
➕
1
3
#16 opened about 2 months ago by
imshenshen
CUDA error: an illegal memory access was encountered using vllm(sm120)
1
#7 opened about 2 months ago by
shakhizat
Does it support MTP?
➕
👍
3
4
#6 opened about 2 months ago by
sharon8811
New activity in
nvidia/Qwen3.6-35B-A3B-NVFP4
3 months ago
Benchmark Report: Qwen3.6-35B-A3B-NVFP4 on NVIDIA DGX Spark, Jetson Thor, Blackwell 6000 Pro
👍
2
#3 opened 3 months ago by
shakhizat
on the DGX spark
👍
5
11
#1 opened 3 months ago by
shakhizat
New activity in
nvidia/NVIDIA-Nemotron-3-Super-120B-A12B-NVFP4
5 months ago
Run on DGX Spark
19
#14 opened 5 months ago by
LimeemiL
RTX Pro 6000 support
3
#7 opened 6 months ago by
justinjja
New activity in
nvidia/NVIDIA-Nemotron-3-Nano-30B-A3B-BF16
8 months ago
Tool calling with reasoning parsing broken
11
#3 opened 8 months ago by
nephepritou
New activity in
onnx-community/gemma-3n-E2B-it-ONNX
12 months ago
How to convert fine tuned gemma3n to ONNX format
#5 opened 12 months ago by
shakhizat
New activity in
unsloth/Qwen3-235B-A22B-GGUF
over 1 year ago
ValueError: Cannot use chat template functions because tokenizer.chat_template is not set and no template argument was passed!
1
#4 opened over 1 year ago by
shakhizat