</> HitReader
Blog Explore About

Tag: #llama.cpp

Exfiltrate Your Weights: A GET-Only Model Demo
machine learning Sep 20, 2026 6 min read

Exfiltrate Your Weights: A GET-Only Model Demo

An apparently silly GET-only service reveals how model weights move from one machine to another. This guide explains chunks, Base64, GGUF, llama.cpp, and the security problems created when URLs perform writes and inference.

by ahsan
#gguf #http #llama.cpp #machine learning #model serving
11–16× Faster LLMs in macOS VMs on Apple Silicon
machine learning Aug 11, 2026 7 min read

11–16× Faster LLMs in macOS VMs on Apple Silicon

macOS VMs on Apple Silicon can run LLM inference far faster when the guest reports conservative Metal capabilities. A process-scoped capability shim steers llama.cpp onto newer Metal kernels, yielding reported 11–16× speedups while keeping the same Virtualization.framework GPU path.

by ahsan
#apple silicon #llama.cpp #macos #metal #virtualization
Qwen 3.6 27B: The Local Dev Sweet Spot
local ai Jun 30, 2026 5 min read

Qwen 3.6 27B: The Local Dev Sweet Spot

Qwen 3.6 27B stands out as a practical local model: high enough quality for day-to-day development and strong performance with llama.cpp. This guide explains why 27B is the sweet spot and shows how to run it (with MTP) and integrate it into coding tools.

by ahsan
#ai tooling #llama.cpp #llm inference #local llms #qwen

Categories

  • artificial intelligence
  • machine learning
  • software engineering
  • cybersecurity
  • developer tools
  • open source
  • web development
  • embedded systems
Explore all →

Tags

#open-source #artificial intelligence #privacy #ai agents #cybersecurity #llm #machine learning #linux #rust #ai #large language models #android
Explore all →

© 2026 HitReader.

About Explore Terms Privacy Facebook