
  <rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom">
    <channel>
      <title>KI im Mittelstand – News &amp; Technologie-Übersicht</title>
      <link>https://www.ki-mittelstand.eu/blog</link>
      <description>Faktenbasierte News, Technologien und Praxislösungen zu Künstlicher Intelligenz für mittelständische Unternehmen. Ohne Hype, mit klaren Quellen.</description>
      <language>de-de</language>
      <managingEditor>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</managingEditor>
      <webMaster>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</webMaster>
      <lastBuildDate>Wed, 29 Jul 2026 00:00:00 GMT</lastBuildDate>
      <atom:link href="https://www.ki-mittelstand.eu/tags/llama-cpp/feed.xml" rel="self" type="application/rss+xml"/>
      
  <item>
    <guid>https://www.ki-mittelstand.eu/blog/ollama-cluster-mehrere-rechner-sharding</guid>
    <title>Ollama-Cluster über mehrere Rechner: Sharding</title>
    <link>https://www.ki-mittelstand.eu/blog/ollama-cluster-mehrere-rechner-sharding</link>
    <description>Ein LLM über zwei Rechner verteilen: Ollama kann es nicht, llama.cpp schon. Befehle, echte Benchmarks — und warum meist eine größere GPU gewinnt.</description>
    <pubDate>Wed, 29 Jul 2026 00:00:00 GMT</pubDate>
    <author>phillip.pham@pexon-consulting.de (KI Mittelstand Team)</author>
    <category>ollama</category><category>llama-cpp</category><category>self-hosted</category><category>gpu</category><category>mittelstand</category><category>deutschland</category>
  </item>

    </channel>
  </rss>
