Filtering by: Local LLM, total 2 post(s)Clear filter
DeepSeek V4.1 Flash Deployment Guide
Tutorials and Guides2026-09-299823

DeepSeek V4.1 Flash Deployment Guide

Deploy DeepSeek V4.1 Flash locally with llama.cpp, API integration, quantization and Flash Attention optimization.

DeepSeek V4.1 Flashllama.cppLocal LLMFlash AttentionQuantization
Read more
llama.cpp Tutorial: Run Local LLMs with OpenAI API
Tutorials and Guides2026-08-136294

llama.cpp Tutorial: Run Local LLMs with OpenAI API

Learn llama.cpp deployment, GGUF models, GPU acceleration, quantization and OpenAI-compatible local API setup.

llama.cppLocal LLMGGUFLLM DeploymentQuantization
Read more