Hugging Face
Models
Datasets
Spaces
Buckets
new
Docs
Enterprise
Pricing
Website
Tasks
HuggingChat
Collections
Languages
Organizations
Community
Blog
Posts
Daily Papers
Hardware
Learn
Discord
Forum
GitHub
Solutions
Team & Enterprise
Hugging Face PRO
Enterprise Support
Inference Providers
Inference Endpoints
Storage Buckets
Log In
Sign Up
🤝
Open to Collab
166.4
TFLOPS
Mike Ravkine
PRO
mike-ravkine
49
27
685
Follow
Keron's profile picture
LeoHyperlink's profile picture
shtefcs's profile picture
88 followers
·
72 following
the-crypt-keeper
AI & ML interests
LLM Research / Development / Evaluation
Recent Activity
liked
a model
1 day ago
ai9stars/G9v3-39A5B
liked
a model
2 days ago
MiniMaxAI/MiniMax-H3
posted
an
update
3 days ago
https://huggingface.co/deepseek-ai/DeepSeek-V4-Flash-0731 is a fantastic model, the only hint of weakness so far is in the tough instruction-following test. For the GPU-poor, my sm89-compatible fork of vLLM is at https://github.com/the-crypt-keeper/vLLM-sm89/tree/sm89-ds4-work I am running L40S but it should work on regular L40 and 4090D as well. I have not yet tested the W2 quantization for performance loss, thats next up, so you'll need 192GB to run it.
View all activity
Organizations
mike-ravkine
's models
6
Sort: Recently updated
mike-ravkine/GLM-4.7-REAP-50-FP8-Dynamic
Text Generation
•
185B
•
Updated
Jan 15
•
7
mike-ravkine/Solar-Open-100B-FP8-Dynamic
103B
•
Updated
Jan 3
•
10
mike-ravkine/Fimbulvetr-11B-v2.1-16K-exl2-6bpw
Updated
Dec 19, 2024
•
5
mike-ravkine/Meta-Llama-3-8B-Instruct-ct2-int8
Updated
Nov 2, 2024
•
4
mike-ravkine/WizardCoder-15B-V1.0-GGUF
16B
•
Updated
Dec 29, 2023
•
60
•
1
mike-ravkine/BlueHeeler-12M
Text Generation
•
Updated
Jun 22, 2023
•
8