
  <rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
    <channel>
      <title>KI im Mittelstand – News &amp; Technologie-Übersicht</title>
      <link>https://www.ki-mittelstand.eu/blog</link>
      <description>Faktenbasierte News, Technologien und Praxislösungen zu Künstlicher Intelligenz für mittelständische Unternehmen. Ohne Hype, mit klaren Quellen.</description>
      <language>de-de</language>
      <managingEditor>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</managingEditor>
      <webMaster>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</webMaster>
      <lastBuildDate>Mon, 20 Jul 2026 00:00:00 GMT</lastBuildDate>
      <atom:link href="https://www.ki-mittelstand.eu/tags/llm/feed.xml" rel="self" type="application/rss+xml"/>
      
  <item>
    <guid>https://www.ki-mittelstand.eu/blog/agree21-protokolle-llm-analyse</guid>
    <title>agree21 Protokolle LLM: Kundenabwanderung 15% früher erkennen</title>
    <link>https://www.ki-mittelstand.eu/blog/agree21-protokolle-llm-analyse</link>
    <description>Entdecken Sie, wie LLMs agree21-Protokolle analysieren, um Kundenabwanderung bis zu 15% früher zu erkennen. Sparen Sie Compliance-Kosten und verbessern Sie die Datenqualität mit KI.</description>
    <pubDate>Mon, 20 Jul 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>finance</category><category>agree21</category><category>llm</category><category>protokollanalyse</category><category>kundenbindung</category><category>compliance</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/anythingllm-enterprise-fuer-fertigung-450k-weniger-ausschuss</guid>
    <title>AnythingLLM Enterprise: Team-Chatbot selbst hosten</title>
    <link>https://www.ki-mittelstand.eu/blog/anythingllm-enterprise-fuer-fertigung-450k-weniger-ausschuss</link>
    <description>AnythingLLM Enterprise als lokaler Multi-User Dokumenten-Chatbot: DSGVO-konformes Setup mit Docker und Ollama für deutsche Mittelständler.</description>
    <pubDate>Thu, 25 Jun 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ki-technologie</category><category>mittelstand</category><category>deutschland</category><category>fertigung</category><category>qualitaetskontrolle</category><category>chatbot</category><category>llm</category><category>anythingllm-enterprise</category><category>team-chatbot-self-hosted</category><category>multi-user-rag</category><category>ausschussreduzierung</category><category>lokale-ki</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/anythingllm-fuer-fertigung-250000-einsparung-durch-team-ki-2</guid>
    <title>AnythingLLM im Mittelstand: Team-KI lokal betreiben</title>
    <link>https://www.ki-mittelstand.eu/blog/anythingllm-fuer-fertigung-250000-einsparung-durch-team-ki-2</link>
    <description>Mit AnythingLLM firmeninterne Dokumente durchsuchen und analysieren. Bis zu €250.000 Einsparung für Fertigungsbetriebe durch schnellere Fehlerfindung und Wissensmanagement. DSGVO-konforme Self-Hosted-Lösung.</description>
    <pubDate>Thu, 02 Apr 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ki-technologie</category><category>mittelstand</category><category>deutschland</category><category>fertigung</category><category>qualitaetskontrolle</category><category>chatbot</category><category>llm</category><category>dokumenten-ki-lokal</category><category>anythingllm-installation</category><category>team-chatbot-self-hosted</category><category>dsgvo</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/deepseek-offline-deutschland-praxisleitfaden-für-deutsche-km</guid>
    <title>Air-Gapped LLM: Llama 3.3 ohne Internet betreiben</title>
    <link>https://www.ki-mittelstand.eu/blog/deepseek-offline-deutschland-praxisleitfaden-für-deutsche-km</link>
    <description>Llama 3.3 komplett offline auf isoliertem Server: Setup in 4 Stunden, Hardware ab €3.200, €0 API-Kosten. Für KRITIS und Produktion.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>air-gapped</category><category>offline</category><category>llm</category><category>self-hosted</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/dspy-framework-llm-programmierung</guid>
    <title>DSPy: LLM-Pipelines ohne Prompt-Engineering</title>
    <link>https://www.ki-mittelstand.eu/blog/dspy-framework-llm-programmierung</link>
    <description>DSPy von Stanford ersetzt Prompt-Engineering durch deklarative Module. 60-80% weniger Entwicklungszeit, Einstieg in 2 Tagen.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>dspy</category><category>llm</category><category>framework</category><category>programmierung</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/enterprise-ki-chatbot-dsgvo-compliance</guid>
    <title>KI-Gateway: LLM-Kosten pro Abteilung tracken</title>
    <link>https://www.ki-mittelstand.eu/blog/enterprise-ki-chatbot-dsgvo-compliance</link>
    <description>KI-Gateway senkt API-Kosten um 28–40% durch zentrales Caching und Routing. Multi-Tenant LLM-Zugang mit Rate Limits und Kostentracking.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ki-gateway</category><category>llm</category><category>multi-tenant</category><category>api-management</category><category>infrastruktur</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/gewerbekunden-kredit-ki-chatbot-fuer-fertigung-spart-250000</guid>
    <title>Gewerbekunden-Kredit: KI-Chatbot für Fertigung spart €250.000 2026</title>
    <link>https://www.ki-mittelstand.eu/blog/gewerbekunden-kredit-ki-chatbot-fuer-fertigung-spart-250000</link>
    <description>Fertigungsunternehmen in Deutschland können mit einem KI-gestützten Kredit-Chatbot für Gewerbekunden über €250.000 pro Jahr einsparen. Dieser Artikel erklärt die Technologie und den ROI.</description>
    <pubDate>Mon, 01 Jun 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ki-technologie</category><category>mittelstand</category><category>deutschland</category><category>fertigung</category><category>qualitaetskontrolle</category><category>computer-vision</category><category>ausschuss</category><category>chatbot</category><category>llm</category><category>kundenservice</category><category>dsgvo</category><category>kredit-chatbot-bank</category><category>gewerbekunden-ki-beratung</category><category>banking-ai-assistent</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/gpt4o-on-premise-ist-das-moeglich</guid>
    <title>GPT-4o on-premise: Ist das möglich?</title>
    <link>https://www.ki-mittelstand.eu/blog/gpt4o-on-premise-ist-das-moeglich</link>
    <description>Kann man GPT-4o on-premise betreiben? Die Antwort ist nein — aber es gibt Alternativen, die GPT-4o-Niveau für weniger als €10.000 Hardware erreichen.</description>
    <pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>gpt4o</category><category>openai</category><category>proprietary</category><category>on-premise</category><category>llm</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/ki-ohne-cloud-on-premise-deutschland-2026-self-hosted-prakti</guid>
    <title>Self-Hosted LLM: Kosten-Guide für Ollama &amp; vLLM 2026</title>
    <link>https://www.ki-mittelstand.eu/blog/ki-ohne-cloud-on-premise-deutschland-2026-self-hosted-prakti</link>
    <description>Vergleichen Sie die Self-Hosting-Kosten von Open-Source-LLMs. Unser Praxis-Guide zeigt TCO für Ollama vs. vLLM für 100+ Mitarbeiter ab 500€/Monat.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>on-premise</category><category>self-hosted</category><category>llm</category><category>vergleich</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/ki-server-modell-update-deployen</guid>
    <title>KI-Server-Modell-Update: Neue LLM-Versionen deployen</title>
    <link>https://www.ki-mittelstand.eu/blog/ki-server-modell-update-deployen</link>
    <description>KI-Server-Modell-Update: Neue LLM-Versionen auf dem eigenen Server deployen — Zero-Downtime-Strategien und Rollback-Plan für produktive Pipelines.</description>
    <pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>update</category><category>llm</category><category>versioning</category><category>ki-server</category><category>devops</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/ki-wissensmanagement-fertigung-firmenwissen-in-2-sek-durchsu</guid>
    <title>KI-Wissensmanagement Fertigung: Firmenwissen in 2 Sek. durchsuchen</title>
    <link>https://www.ki-mittelstand.eu/blog/ki-wissensmanagement-fertigung-firmenwissen-in-2-sek-durchsu</link>
    <description>KI-Wissensmanagement für die Fertigung: Firmenwissen lokal in unter 2 Sekunden durchsuchen, mit über 95% Trefferquote per Chat-Schnittstelle und RAG.</description>
    <pubDate>Mon, 30 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ki-technologie</category><category>mittelstand</category><category>deutschland</category><category>fertigung</category><category>qualitaetskontrolle</category><category>computer-vision</category><category>ausschuss</category><category>chatbot</category><category>llm</category><category>kundenservice</category><category>dsgvo</category><category>wissensmanagement-ki</category><category>enterprise-search-lokal</category><category>dokumentensuche-ai</category><category>rag</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/langfuse-self-hosted-llm-monitoring-und-tracing-2026-praktis</guid>
    <title>LangGraph Agenten: Workflows mit 38.000 € Einsparung</title>
    <link>https://www.ki-mittelstand.eu/blog/langfuse-self-hosted-llm-monitoring-und-tracing-2026-praktis</link>
    <description>LangGraph orchestriert mehrstufige KI-Agenten mit Verzweigungen und Tool-Aufrufen. Ein Mittelständler spart 38.000 €/Jahr bei der Angebotsbearbeitung.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>langgraph</category><category>agenten</category><category>workflows</category><category>llm</category><category>automatisierung</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/llama-3-deutsch-installation-anleitung-2026-llm-gateway-prak</guid>
    <title>Llama 3.3 70B: Installation auf Deutsch (Ubuntu/Docker)</title>
    <link>https://www.ki-mittelstand.eu/blog/llama-3-deutsch-installation-anleitung-2026-llm-gateway-prak</link>
    <description>Richten Sie Llama 3.3 70B für deutsche Anwendungsfälle ein. In dieser Anleitung zeigen wir in 5 Schritten die Installation via Docker auf Ubuntu.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>llama</category><category>fine-tuning</category><category>deutsch</category><category>llm</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/llmlite-vs-ollama-lokale-enterprise-ki</guid>
    <title>LiteLLM Proxy: 30-40 % API-Kosten sparen</title>
    <link>https://www.ki-mittelstand.eu/blog/llmlite-vs-ollama-lokale-enterprise-ki</link>
    <description>LiteLLM Proxy bündelt OpenAI, Anthropic und Ollama unter einer API. Setup in unter 1 Stunde, 30-40 % weniger API-Kosten durch Routing.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>litellm</category><category>proxy</category><category>llm</category><category>self-hosted</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/meistgenutzte-llms-2026-europa-hosting</guid>
    <title>Die zehn meistgenutzten LLMs — welche in Europa laufen</title>
    <link>https://www.ki-mittelstand.eu/blog/meistgenutzte-llms-2026-europa-hosting</link>
    <description>Sieben der zehn meistgenutzten Modelle kommen aus China, mit offenen Gewichten. Was das für deutsche Unternehmen heißt — Lizenz- und VRAM-Tabelle.</description>
    <pubDate>Tue, 04 Aug 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>open-weights</category><category>llm</category><category>hosting</category><category>dsgvo</category><category>gpu</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/moe-vram-bedarf-aktive-parameter</guid>
    <title>MoE erklärt: warum 30B nicht 30B VRAM braucht</title>
    <link>https://www.ki-mittelstand.eu/blog/moe-vram-bedarf-aktive-parameter</link>
    <description>Ein 30B-MoE belegt die vollen 19 GB VRAM — sparsam ist nur die Rechenlast pro Token. Die Speicherrechnung Posten für Posten, mit Modelltabelle.</description>
    <pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>moe</category><category>llm</category><category>gpu</category><category>vram</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/ollama-modelfile-erstellen-deutsch-anleitung</guid>
    <title>Ollama Modelfile: Eigenes KI-Modell erstellen</title>
    <link>https://www.ki-mittelstand.eu/blog/ollama-modelfile-erstellen-deutsch-anleitung</link>
    <description>Ollama Modelfile erstellen und eigene KI-Modelle konfigurieren: System-Prompts, Parameter und Praxisbeispiele Schritt für Schritt.</description>
    <pubDate>Sun, 08 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ollama</category><category>mittelstand</category><category>deutschland</category><category>self-hosted-ki</category><category>llm</category><category>ki-technologie</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/ollama-ubuntu-installieren-anleitung-2025-self-hosted-prakti</guid>
    <title>Ollama Ubuntu installieren: LLM lokal 15 Min</title>
    <link>https://www.ki-mittelstand.eu/blog/ollama-ubuntu-installieren-anleitung-2025-self-hosted-prakti</link>
    <description>Ollama auf Ubuntu installieren: Lokales LLM in 15 Minuten. Llama 3.1 auf eigenem Server, €0 API-Kosten, volle DSGVO-Kontrolle.</description>
    <pubDate>Mon, 09 Mar 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ollama</category><category>ubuntu</category><category>self-hosted</category><category>llm</category><category>installation</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/qwen-eigenen-server-chinas-open-weight-modell-on-premise</guid>
    <title>Qwen auf eigenem Server: Chinas Open-Weight-Modell on-premise betreiben</title>
    <link>https://www.ki-mittelstand.eu/blog/qwen-eigenen-server-chinas-open-weight-modell-on-premise</link>
    <description>Qwen von Alibaba als Open-Weight-Modell auf dem eigenen KI-Server betreiben: Warum deutsche Unternehmen jetzt auf chinesische LLMs setzen – DSGVO-konform, komplett on-premise.</description>
    <pubDate>Sat, 18 Jul 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>qwen</category><category>open-weight</category><category>llm</category><category>self-hosted</category><category>ki-server</category><category>mittelstand</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/rag-fuer-fertigung-arbeitsanweisungen-per-chatbot-150000-meh</guid>
    <title>RAG für Fertigung: Arbeitsanweisungen per Chatbot – 150.000€ mehr Output 2026</title>
    <link>https://www.ki-mittelstand.eu/blog/rag-fuer-fertigung-arbeitsanweisungen-per-chatbot-150000-meh</link>
    <description>RAG für Fertigung: Finden Sie Arbeitsanweisungen in Sekunden. Spart 150.000€ mehr Output pro Jahr für mittelständische Betriebe. DSGVO-konform mit Audit-Trail.</description>
    <pubDate>Sun, 31 May 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ki-technologie</category><category>mittelstand</category><category>deutschland</category><category>fertigung</category><category>qualitaetskontrolle</category><category>computer-vision</category><category>ausschuss</category><category>chatbot</category><category>llm</category><category>kundenservice</category><category>dsgvo</category><category>rag-banking</category><category>banken-chatbot</category><category>arbeitsanweisungen-ki</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/soofi-s-selbst-hosten-gpu</guid>
    <title>Soofi S selbst hosten: welche GPU reicht wirklich?</title>
    <link>https://www.ki-mittelstand.eu/blog/soofi-s-selbst-hosten-gpu</link>
    <description>Soofi S hat 3,2 Mrd. aktive Parameter, braucht aber VRAM für 31,6 Mrd. Was das für die GPU-Wahl heißt — mit offener Rechnung und Stand der Tools.</description>
    <pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>soofi</category><category>llm</category><category>gpu</category><category>self-hosted</category><category>mittelstand</category><category>deutschland</category>
  </item>

  <item>
    <guid>https://www.ki-mittelstand.eu/blog/soofi-s-vs-llama-qwen-deutsch</guid>
    <title>Soofi S vs Llama 3 vs Qwen für deutschen Text</title>
    <link>https://www.ki-mittelstand.eu/blog/soofi-s-vs-llama-qwen-deutsch</link>
    <description>Soofi S, Llama 3.3 und Qwen3.5 im Vergleich für deutsche Texte: Lizenz, VRAM, Verfügbarkeit — und ein Testaufbau mit den eigenen Dokumenten.</description>
    <pubDate>Sun, 02 Aug 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>llm</category><category>open-source</category><category>self-hosted</category><category>soofi</category><category>mittelstand</category><category>deutschland</category>
  </item>

    </channel>
  </rss>
