Will that LLM run on your GPU? VRAM fit, tokens/sec, specs and a leaderboard for local + API models — right inside Discord.
Slopsome brings LLM & GPU stats straight into Discord — no setup, no API key.
Ask whether a model will run on a given GPU and get a real answer: VRAM fit (in-VRAM / with offload / needs more), plus prefill (compute-bound) and decode (bandwidth-bound) tokens/sec — for any quant: GGUF k-quants & i-quants, AWQ, GPTQ, EXL2/EXL3.
Commands
/fit — will a model run on a GPU? VRAM + tok/s, any quant & context/model — specs, benchmarks, per-quant VRAM, pricing for one model/gpu — VRAM, bandwidth, TFLOPS and reference tok/s for a GPU/leaderboard — top models by composite score/search — find a model by name / maker / family/help — list everythingCovers local open-weight models and hosted API models. Data comes from slopsome.com — a free, no-login search engine for LLM & GPU stats. Invite the bot and type /help.
0
리뷰 0개
리뷰는 등록된 사용자만 남길 수 있습니다. 모든 리뷰는 Top.gg의 사이트 중재자가 관리합니다. 게시하기 전 저희의 지침을 반드시 확인해 주세요.
별점 5점
0
별점 4점
0
별점 3점
0
별점 2점
0
별점 1점
0