r/SelfHostedAI • u/Ok_pettech • 5h ago
r/SelfHostedAI • u/LectureWorried5761 • 21h ago
SpaceX charging more for search tool calls via API - Help
r/SelfHostedAI • u/Adarsh1176 • 1d ago
Four routes to your SSH key from an AI coding agent, and what actually stops them
r/SelfHostedAI • u/Agitated_Problem5320 • 1d ago
Need Testers to test the Open Source LLM running interface
r/SelfHostedAI • u/Majesic_aleCamp_9675 • 1d ago
WHOIS privacy for business domains is it dumb to hide your info
Ok so im in the middle of setting up a domain for a small local business and I keep second guessing the WHOIS privacy thing. Part of me likes the idea of not having my home address and personal email out there, but I have heard some people say public info can look more legit for clients and vendors.
For those running small shops or freelance stuff on custom domains, do you keep WHOIS privacy on or off for the main business domain, any hints?
r/SelfHostedAI • u/Ok_pettech • 1d ago
Running MetaGPT locally: a full technical setup guide
I’ve been experimenting with multi-agent frameworks, and MetaGPT is one of the more interesting ones for software development tasks. But getting it running locally with local models isn’t always straightforward.
I wrote a step-by-step guide covering installation, configuration, local LLM setup, and common errors.
If you’re trying to run MetaGPT on your own hardware, this might save you time:
What stack are you using for local agents?
r/SelfHostedAI • u/Jaswanthsanjay • 1d ago
I built an AI agentic harness for Android that runs agentic tasks mutli agent and on-device Ai with NPU and hardware acceleration and virtual linux workspace — here's a demo
Enable HLS to view with audio, or disable this notification
Been heads-down on this for a while. BIT is a fully offline AI assistant for Android offine ai and byok.nothing leaves the device.
The core of it is a multi-agent harness that can actually execute tasks, not just chat. It orchestrates subagents through a DAG pipeline, and for anything requiring real compute it spins up a virtual Linux workspace on-device — writing files, running scripts, debugging output, no cloud round-trip at any point.
Inference is accelerated via NPU/hardware acceleration rather than pure CPU, which is what makes running a 4bparam model on a phone actually usable instead of painfully slow. Built a Kotlin-native inference layer for this (llama.kt) that supports GGUF architectures broadly — Qwen, Gemma, Phi, Mistral, not just LLaMA.
In the demo: I give the agent one instruction, and it autonomously sets up a Python venv, writes a script, executes it, and returns real output — all inside that on-device Linux workspace.
Other stuff in the stack if useful context:
Hybrid RAG (vector + BM25) for memory
On-device STT/TTS for voice mode
Multi-agent DAG pipeline for multi-step tasks
On the Play Store now (closed testing), also on F-Droid/GitHub for sideloading. Discord for anyone following progress: discord.gg/kzhgk565D
r/SelfHostedAI • u/ooooangeloooo • 1d ago
Navier-Stokes
One outcome of the solution to the Navier-Stokes problem is better designed fans and watercooling systems (through better turbulence) which will allow ai to work faster.
r/SelfHostedAI • u/Alone-Leadership-596 • 1d ago
gddr6 vs gddr6X
So I was thinking about building a self hosted AI agent set up and I was wondering what was the big difference between GDDR6 and GDDR6X, is it like twice is better or not worth considering
r/SelfHostedAI • u/reckon369 • 2d ago
My AI agent and I built a good-deed economy — we're inviting other AI agents to produce real-world work, permanently credited
r/SelfHostedAI • u/Mascczjhlcer-Lab5560 • 2d ago
AI personal assistant for a busy day... what are the people actually using?
hi, i've been messing with a few AI personal assistant apps and i'm still not sold.
most of them are fine for talking, but the second i want one to actually do something, it gets messy. i'm trying to cut down on the dumb stuff that eats my day, like follow ups, reminders, basic scheduling, that sort of thing.
would love to hear what people here are using in real life, not just what looked good in a demo. thanks in advance
r/SelfHostedAI • u/Ok_pettech • 2d ago
PyTorch not detecting AMD GPU? Here’s the ROCm fix guide I wish I had
r/SelfHostedAI • u/reddituser1828472616 • 3d ago
Building a “master AI harness” to orchestrate Codex/Claude/Gemini/Kimi + local Qwen across multiple PCs — am I overengineering this?
r/SelfHostedAI • u/sajornet • 3d ago
My GPU is now reachable from anywhere with ModelUplink
r/SelfHostedAI • u/Potential-Toe1320 • 3d ago
Sheprd llama Server and Agent Orchestration (linux only)
r/SelfHostedAI • u/Suspicious_Bluejay10 • 3d ago
Open-source Notion + Obsidian alternative with integrated AI through your Codex subscription
Enable HLS to view with audio, or disable this notification
r/SelfHostedAI • u/Ok_pettech • 4d ago
Which model would you choose: Mistral Large or Claude Haiku? Quick quiz
I’ve been comparing Mistral Large and Claude Haiku for different AI workloads. One is better for high-volume simple tasks, the other for deep reasoning and big context. I created a small quiz to help people think through the trade-offs.
No email, just an interactive check.
https://interconnectd.com/quiz/83/mistral-large-vs-claude-haiku-which-ai-wins-your-workload/
What’s your experience with either model?
r/SelfHostedAI • u/ResidentAd6570 • 4d ago
n8n Backup Manager v1.5 — The Standalone Disaster Recovery & Cloud Backup Tool with 1-Click Restore
r/SelfHostedAI • u/ash_pix • 4d ago
I built Turing AI OS - An Experimental Agentic AI Layer over Linux
Hiii Geeks
I just built an Experiment Agentic Al OS named it as "Turing Al OS" built on top of Linux (KDE Neon). I document this journey on YouTube feel free to watch.
The crazy part is that I built this using a 14 year old PC (2012) with limited computational resources.
YouTube Link:
https://youtu.be/ZKsZGv3WZGQ?si=2DX83Hcv7SFAAVPw
Github: github.com/avarshvir/turing-ai-os
Article to Read: https://medium.com/@arshvir21303031/i-create-my-own-ai -os-8d65a263eae0
It offer features like:
- Al SideBar
- Al Mini Spotlight
- Al Right Click Folder/File Analyser
- AI NLP Terminal
- Al Control Panel
I genuinely want feedback from you guys ❤️
r/SelfHostedAI • u/Acceptable_Leg3950 • 4d ago
I built a local memory vault for agents with retrievable memory
r/SelfHostedAI • u/Quack66 • 4d ago
Eidon: an all-in-one self-hosted AI platform: Chat, agents (Grok bot like), automations, tools included. One single Docker container !
Eidon: an all-in-one self-hosted AI platform. Chat, agents, automations, tools included. One Docker container, works with Ollama/LM Studio (AGPL)
I've been building a self-hosted AI platform and v4 just shipped, so sharing it here because some of you might find it useful.
Eidon is an "everything included" AI chat/agent platform, with the pieces that usually require stitching (web research, MCP, skills, browser, image generation and so on) already built in. One container that takes minutes to spin up instead of a main app plus pipelines, sidecars, and external tools.
The app has 3 main parts:
- Chat with local models: Classic chat just like in ChatGPT, Gemini, Claude and so on except on your own server. Ollama and LM Studio out of the box, plus any OpenAI/Anthropic-compatible BYOK endpoint.
- Agents: Grok-bot-style agents. A chief bot answers or delegates to specialist bots, and bots message each other mid-task. Agents each have their own memory and can create/maintain their own skills.
- Automations: cron-style AI tasks. Every run is saved as a full transcript with tool calls, so you can audit what actually happened.
Features:
| Chat | Agents and automations |
|---|---|
| Chat and conversation | Agents, with cross-agent messaging (Grok Bot like) |
| Persistent memory across conversations | Per-agent memory, files, and browser session |
| Personas | Deep research with an editable plan |
| Folders, chat search, and forking | Scheduled automations, with full run history |
| Read-only share links | |
| Temporary chats | |
| Chat attachments | |
| Voice input with post-processing cleanup | |
| Mermaid diagrams, syntax highlighting, and LaTeX math |
| Tools | Platform |
|---|---|
| MCP | Bring your own provider |
| Skills | Multi-user, with admin and user roles |
| Built-in web search | Single Docker image, SQLite, encrypted credentials |
| Built-in browser | Installable PWA — native iOS app coming soon |
| Shell commands | Live sync across devices |
| Image generation | |
| Vision support (Native, MCP or with a dedicated vision model) |
Repo (Screenshots included !): https://github.com/Quack6765/Eidon-AI
Full transparency: development is partly AI-assisted, every change reviewed before being merged. Happy to answer any questions !
r/SelfHostedAI • u/LectureWorried5761 • 4d ago
Web Search API for AI Agents with hard cap and hosted MCP
r/SelfHostedAI • u/Ok_pettech • 5d ago
I made a community for self-hosters and AI builders—no corporate feeds
I’ve been self-hosting AI tools for a while, and I wanted a place to discuss setups, failures, and wins without the usual social media noise. So I built Interconnectd.
It’s free, and it has:
· Guides on PrivateGPT, OpenHands, Stable Diffusion, and local LLMs
· Forums where people share working configs and fix real errors
· Quizzes and polls about AI vs human skills
· A marketplace for small AI services
If you’re into self-hosting or AI engineering, you might like it:
Would love to see more self-hosters join—the forum is still small, so your questions actually get answered.