Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
Raushan Turganbay
RaushanTurganbay
233
64
45
Follow
shuyuej's profile picture
Blessing988's profile picture
Fishtiks's profile picture
116 followers
·
41 following
zucchini-nlp
AI & ML interests
Generation and Multimodality
Recent Activity
liked
a model
about 20 hours ago
moonshotai/Kimi-K3
upvoted
an
article
6 days ago
Kimi K3 Model Overview: 2.8T Parameters, MXFP4 Quantization, and What the Open Weights Mean for the Community
updated
a model
11 days ago
hf-internal-testing/kimi-k25-for-integration-test
View all activity
Organizations
RaushanTurganbay
's activity
All
Models
Datasets
Spaces
Buckets
Papers
Collections
Community
Posts
Upvotes
Likes
Articles
New activity in
allenai/Olmo-3-7B-Instruct
11 days ago
Update tokenizer_config.json
#14 opened 11 days ago by
RaushanTurganbay
New activity in
zai-org/GLM-4.1V-9B-Thinking
11 days ago
Fix config model type
#22 opened 11 days ago by
RaushanTurganbay
New activity in
transformers-community/group-beam-search
21 days ago
Optimize group beam search by sharing the initial prefill across beams
3
#6 opened 22 days ago by
lavrenko
New activity in
transformers-community/group-beam-search
22 days ago
Possible performance issue: group beam search appears to prefill identical prefixes once per beam
3
#5 opened 28 days ago by
lavrenko
New activity in
transformers-community/sink_cache
28 days ago
"sink_cache" fails on Transformers 5.12 due to outdated "Cache.__init__" API
1
#1 opened 28 days ago by
lavrenko
New activity in
transformers-community/constrained-beam-search
28 days ago
constrained-beam-search produces degenerate output with use_cache=True
3
#2 opened 30 days ago by
lavrenko
New activity in
transformers-community/dola
28 days ago
DoLa custom generation produces gibberish with use_cache=True
1
#1 opened 30 days ago by
lavrenko
Fix Transformers v5 cached decoding in DoLa
2
#2 opened 29 days ago by
lavrenko
New activity in
transformers-community/constrained-beam-search
28 days ago
Support Transformers v5 cache handling
2
#3 opened 29 days ago by
lavrenko
New activity in
transformers-community/group-beam-search
about 1 month ago
Custom group-beam-search changes deterministic output when cache is enabled
3
#3 opened about 1 month ago by
lavrenko
Support Transformers v5 cache handling
2
#4 opened about 1 month ago by
lavrenko
New activity in
llava-hf/llava-1.5-7b-hf
2 months ago
[Question] Why does LLaVA evaluation assert batch_size == 1 for benchmark inference?
1
#62 opened 2 months ago by
austinkomachi
New activity in
OpenGVLab/InternVL2-8B
4 months ago
Compatibility with v5
#23 opened 4 months ago by
RaushanTurganbay
New activity in
OpenGVLab/InternVL2-1B
4 months ago
Compatibility with v5
#10 opened 4 months ago by
RaushanTurganbay
New activity in
OpenGVLab/InternVL2-2B
4 months ago
Compatibility with v5
#7 opened 4 months ago by
RaushanTurganbay
New activity in
OpenGVLab/InternViT-300M-448px-V2_5
4 months ago
Compatibility with v5
#5 opened 4 months ago by
RaushanTurganbay
New activity in
OpenGVLab/InternViT-300M-448px
4 months ago
Compatibility with v5
#6 opened 4 months ago by
RaushanTurganbay
New activity in
PerceptronAI/Isaac-0.1
4 months ago
Compatibility with v5
#6 opened 4 months ago by
RaushanTurganbay
New activity in
PerceptronAI/Isaac-0.2-1B
4 months ago
Compatibility with v5
#4 opened 4 months ago by
RaushanTurganbay
New activity in
PerceptronAI/Isaac-0.1-Base
4 months ago
Compatibility with v5
#4 opened 4 months ago by
RaushanTurganbay
Load more