Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
luchangli03
luchangli03
5
3
Follow
Steve-Guo's profile picture
1 follower
·
4 following
AI & ML interests
None yet
Recent Activity
new
activity
21 days ago
inference-optimization/Kimi-K3-0.40B:
get error when served by sglang
new
activity
2 months ago
mmangkad/GLM-5.2-NVFP4:
why the config.json is different from official GLM 5.2?
new
activity
2 months ago
lukealonso/GLM-5.2-NVFP4:
launch failed with EAGLE
View all activity
Organizations
luchangli03
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
inference-optimization/Kimi-K3-0.40B
21 days ago
get error when served by sglang
#4 opened 21 days ago by
luchangli03
New activity in
mmangkad/GLM-5.2-NVFP4
2 months ago
why the config.json is different from official GLM 5.2?
1
#1 opened 2 months ago by
luchangli03
New activity in
lukealonso/GLM-5.2-NVFP4
2 months ago
launch failed with EAGLE
➕
1
2
#3 opened 2 months ago by
luchangli03
liked
a model
6 months ago
lightseekorg/kimi-k2.5-eagle3
3B
•
Updated
Mar 16
•
141k
•
15
New activity in
AQ-MedAI/Kimi-K25-eagle3
6 months ago
Can you reduce the kv head num of this model? "num_key_value_heads": 64, which requies a lots of kv cache
2
#1 opened 6 months ago by
luchangli03
liked
a model
6 months ago
AQ-MedAI/Kimi-K25-eagle3
1B
•
Updated
Jun 25
•
4.65k
•
10
New activity in
jerryzh168/Kimi-K2-Thinking-FP8
6 months ago
Can you provide the code that convert the int4 weight to fp8? thanks
#2 opened 6 months ago by
luchangli03
New activity in
chutesai/DeepSeek-V3.1-Terminus-NextN
9 months ago
what's the difference between this nextn and self contained MTP model in DeepSeek-V3.1-Terminus?
#1 opened 9 months ago by
luchangli03
liked
a Space
over 1 year ago
Running
4k
The Ultra-Scale Playbook
🌌
4k
The ultimate guide to training LLM on large GPU Clusters