self-hosted LLM

imagem 50

Ollama v0.33.0-rc2 Adds Claude Desktop Integration and Cache Reliability Fixes

Ollama’s latest release candidate, v0.33.0-rc2, brings a handful of meaningful updates for self-hosted users, including tighter Claude Desktop integration and several under-the-hood improvements that make model caching more dependable. The Claude Desktop app now works directly with Ollama. You can toggle individual Ollama models on or off right from the menu bar, and pick any […]

Ollama v0.33.0-rc2 Adds Claude Desktop Integration and Cache Reliability Fixes Read More »

imagem 45

llama.cpp Build b10176 Brings Tensor Memset RPC and Expanded Platform Support

The latest llama.cpp build, b10176, introduces a valuable new feature for distributed and remote inference scenarios: the tensor_memset remote procedure call. This addition allows for efficient initialization of tensors across RPC connections, which is particularly beneficial for self-hosted setups that distribute model layers across multiple machines. By streamlining how memory is set up remotely, it

llama.cpp Build b10176 Brings Tensor Memset RPC and Expanded Platform Support Read More »