One’s Vibe

Project detail · Developer Tools

Will It Fit? - Opinionated llama.cpp VRAM Estimator

Estimate VRAM requirements for running models with llama.cpp

hypfer.github.io
Will It Fit? - Opinionated llama.cpp VRAM Estimator — screenshot
Try live demo Live checked15 hours ago

Try this first: Input model size and quantization format to estimate VRAM

Tech Profile

Built by
Solo maker
AI help
Primarily AI-assisted

Site health

🛠 4 site-health suggestions await the maker — claim this project (sign in with X) to view.

Source

Hacker News

hypfer
According to this shitty vibecoded thing "I" built https://hypfer.github.io/will-it-fit-llama-cpp/ (and I guess according to math too), FP16 K/V would give me something like 90k context at the same model quant, which doesn't really fit my usage. But maybe someone else has experience to share there

For the maker: hang the plaque, claim the project

This project is unclaimed — sign in, hang the plaque on your site, and it's yours.

Comments

Sign in to comment.

  • No comments yet. Be the first.

Keep exploring

Similar projects

Sign in to report a problem with this project.