Una GPU VPS usa e getta per i tuoi LLM: LiteLLM, Ollama, vLLM e storage cifrato
Noleggia una GPU VPS, trasformala in una API LLM compatibile OpenAI (LiteLLM, Ollama, vLLM), poi distruggi la VM: modelli e dati vivono su un volume LUKS cifrato.
Noleggia una GPU VPS, trasformala in una API LLM compatibile OpenAI (LiteLLM, Ollama, vLLM), poi distruggi la VM: modelli e dati vivono su un volume LUKS cifrato.
Rent a GPU VPS, turn it into an OpenAI-compatible LLM API (LiteLLM, Ollama, vLLM), then destroy the VM — your models and data survive on an encrypted LUKS volume.
How I came to understand language models a little better by trying to squeeze one into a few megabytes.
Come ho capito un po' meglio i modelli linguistici provando a farne stare uno in pochi megabyte.
A plain-language LLM glossary — but not alphabetical: in dependency order, so by the time you reach Mixture of Experts you already have all the pieces. With analogies for sysadmins.