Ollama 2
- Local-Splitter: Cutting Cloud LLM Costs by Putting a Small Model in Front Apr 15, 2026 Cloud LLM tokens are expensive. Not in the “my AWS bill is high” sense — in the “I’m burning $0.015 per 1K output tokens and my coding agent …
- Ollama-Forge for Security Research: Local Models, Refusal Ablation, and Reproducible Pipelines Feb 16, 2026 Introduction Security work often involves prompts and data you cannot send to commercial APIs: malware descriptions, exploit drafts, …