One-click LLM deployments on Private GPU Collection Every model deployable on HexGrid Cloud with one click. Dedicated GPU, private API endpoint, OpenAI-compatible. Visit https://hexgrid.cloud • 10 items • Updated Jun 7
Best Open-Source Coding LLMs for Private Deployment Collection Code generation, debugging, review, and test writing. All deployable privately on dedicated GPUs at hexgrid.cloud • 3 items • Updated Jun 7
Production-Ready Quantized Chat LLMs — 4-bit & 8-bit Collection FP8, AWQ-4Bit and W8A8 quantized versions of popular models. Lower VRAM, same production quality. Deploy at hexgrid.cloud in one click. • 9 items • Updated Jun 7
Open Source RAG Stack — Embed + Rerank + Generate Collection The complete open-source RAG pipeline. Best of the embedding models, one reranker, one chat model. All deployable on dedicated GPUs at hexgrid.cloud • 6 items • Updated Jun 7