Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
33.7
TFLOPS
Avi Fenesh
Avifenesh
29
2
23
Follow
buttetsu's profile picture
leiqiang's profile picture
Venomed's profile picture
6 followers
·
11 following
https://inference.tiyuvta.ai/
avi_fenesh
avifenesh
avi-fenesh
AI & ML interests
None yet
Recent Activity
liked
a model
about 7 hours ago
incoai/GLM-5.3-Flash-DFlash2
replied
to
their
post
about 8 hours ago
Wrote this up because it still feels backwards. Acceptance dropped, tok/s went up. Full draft head: 66.7% / 117.1 tok/s. Trimmed: 63.6% / 121.7. Later remeasure +5.1% on a Pro 6000, +6.4% on a 5090 laptop. Full head still accepts more. Loses anyway. Not my idea. FR-Spec. I just wired it for the native MTP head. https://huggingface.co/blog/Avifenesh/masked-mtp-drafts https://huggingface.co/Avifenesh/Qwen3.8-27B-NVFP4-MTP-GGUF
updated
a model
about 12 hours ago
tiyuvta/Qwen3.8-Flash-Next-NVFP4
View all activity
Organizations
Avifenesh
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
published
an
article
16 days ago
view article
Article
Masked MTP drafts
Avifenesh
•
16 days ago
•
1