AI Agents on VPS
34 articles in this section
This section covers running AI tools on your own server — without relying on someone else's API and with full control over your data. It starts with what AI agents are and why you'd run them on a VDS, installing Ollama to run local LLMs, the Open WebUI interface on top of it, and how engines like vLLM, llama.cpp, and Text Generation WebUI compare on different hardware.
From there it moves to building your own systems: agents built with LangChain, multi-agent setups through CrewAI and AutoGen, no-code automation in n8n, the visual builder Flowise, and RAG for answering questions over your own documents using vector databases like Chroma, Qdrant, or pgvector inside PostgreSQL. It also covers connecting external APIs — OpenAI, Claude, and Gemini — along with their cost and how they compare.
There are practical scenarios too: a Telegram bot built on LangChain and Ollama, a PHP-and-JavaScript chat widget for a website, speech transcription with Whisper, image generation with Stable Diffusion and ComfyUI, sizing VRAM for a given model, and protecting agents from prompt injection.
- Local LLMs: Ollama, vLLM, llama.cpp
- Agents and RAG: LangChain, CrewAI, vector databases
- No-code automation: n8n and Flowise
- Monitoring and securing AI agents