Build a private Gmail agent in Python using local LLMs to summarize, classify, and draft emails without cloud exposure. Gmail API and Ollama setup.
AI
Hands-on guides to LLMs, agents, prompt engineering, and the AI tools Botmonster runs every day for real work, not demos.
Evaluating AGENTS.md: are repository context files actually helpful?
When do repository context files like CLAUDE.md help AI agents? An analysis of documentation quality, structure, and when context files become noise.
Run FLUX 2 locally in 2026: VRAM by GPU + ComfyUI setup
Run FLUX 2 locally in ComfyUI. VRAM by GPU from 8GB to 24GB, GGUF builds, the variant that fits your card, cost versus cloud, and the files to grab.
Why small language models (SLMs) are better for edge devices
Small Language Models like Phi-4 and Gemma 3 run offline on Raspberry Pi and other edge devices. Why sub-4B parameters beat cloud inference calls.
SDXL 2.0 LoRA: 50-300 MB adapters on 12 GB VRAM
Fine-tune Stable Diffusion XL 2.0 with LoRA adapters: dataset prep with vision-language models, training with kohya_ss, custom styles on 12GB VRAM.
Setup a private local RAG knowledge base
Build a private RAG with Qdrant, BGE-M3 embeddings, and Ollama. Index your documents locally and answer questions with zero data leaving your machine.
Botmonster Tech




