Ollama
2026
- Local-Splitter: Cutting Cloud LLM Costs by Putting a Small Model in Front Apr 15 Cloud LLM tokens are expensive. Not in the “my AWS bill is high” sense — in the “I’m burning $0.015 per 1K …
- Ollama-Forge for Security Research: Local Models, Refusal Ablation, and Reproducible Pipelines Feb 16 Introduction Security work often involves prompts and data you cannot send to commercial APIs: malware descriptions, exploit drafts, …