Jonathon's AI Wiki

kv-cache

2 items with this tag.

  • Jul 24, 2026

    Hermes on Apple Silicon — Local Model, Backend & Quant Guide (Mac/MLX)

    • hermes
    • local-llm
    • apple-silicon
    • mlx
    • llama-cpp
    • ollama
    • lm-studio
    • qwen
    • gemma
    • quantization
    • kv-cache
    • tool-calling
    • mac
    • reddit-sourced
    • r-hermesagent
    • qwen3-6
    • chat-template
    • mtp
  • Jun 13, 2026

    Headroom — Context Compression Layer for AI Agents

    • context-compression
    • token-optimization
    • proxy
    • mcp
    • claude-code
    • cross-agent-memory
    • open-source
    • headroom
    • apache-2
    • cost-reduction
    • kv-cache

Created with Quartz v5.0.0 © 2026

  • ✦ Explore the graph in 3D