← Blog

Ollama

5 articles about "Ollama".

Content Marketing自動化OllamaThreadsSolo DevBuildInPublicLocal LLM

Content Cascade Engine: Write One Blog Post, Auto-Generate 5 Social Posts

I built a Content Cascade system that scans for new blog posts every morning at 7 AM, uses a local Ollama model to split them into 3-5 Threads posts. Zero API cost, zero manual work. One article becomes six pieces of content. Full architecture, prompt design, and quality data inside.

· 14 min read
AI CostGeminiClaudeOllamaLocal LLMFree TierBenchmark

Ollama vs Gemini vs Claude: 90-Day Cost Test

Gemini's free tier wins on speed and long context, Ollama on unlimited private offline runs, Claude Pro on quality. 90 days of real cost-per-request data.

· 10 min read
NVIDIAGPULocal LLMCloud APICost AnalysisOllamaInferenceROI

Local LLM vs Cloud API Cost: GPU Break-Even

On-prem LLM vs cloud API cost: at ~3,150 requests/month, budget APIs beat a used RTX 3060 Ti ($13.33/mo). API bills scale per token; local cost stays flat.

· 9 min read
NVIDIAGPUAI AgentOrchestrationOpenClawOllamaArchitecture

Multi-Agent AI Orchestration on One 8GB GPU

How to run 4 AI agents on one 8GB RTX 3060 Ti without conflicts: a 3-layer architecture of isolated workspaces, staggered systemd timers, sequential inference.

· 11 min read
NVIDIAGPUOllamaLocal LLMAI AgentInference

Running a 4-Agent AI Fleet on a Single NVIDIA RTX 3060 Ti

We run 4 autonomous AI agents on a single NVIDIA RTX 3060 Ti with 8GB VRAM. 13.2 tok/s inference, 105 daily tasks, 99.9% uptime. Here's the complete hardware setup, performance tuning, and lessons learned from 30 days of production.

· 10 min read