About
Got tired of guessing whether a model fits my card, so I made slopsome.com — a VRAM fit-calculator + real tokens/sec for local and API models. Pick a model, your GPU, and quant, and it tells you if it fits / needs offload / multi-GPU / won't fit, plus rough speed. Just added EXL3 (bits-per-weight) and an advanced panel for KV-cache quant + draft tokens and concurrency, since that's where VRAM quietly blows up. submitted by /u/Defiant_Rough5325 [link] [comments]
Comments (1)
how do you plan to monetize this long term
Related Products
Typist
Transcribe audio to text in seconds.
rapym
TKCORE AI
All-in-one AI platform featuring TkCore-V5.5-Pro & top LLMs
Textsight.ai
TextSight.ai detects AI-like writing, gives a 0–100 Authenticity / Humanization
ai image prompts
Copy-paste AI image prompts that actually work.
ChatFlow
AI chatbot for your website — live in minutes
ComingUp