<?xml version="1.0" encoding="utf-8" standalone="yes"?>
<rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom" xmlns:content="http://purl.org/rss/1.0/modules/content/">
  <channel>
    <title>Ai on C&amp;CZ</title>
    <link>https://cncz.science.ru.nl/tags/ai/</link>
    <description>Recent content in Ai on C&amp;CZ</description>
    <generator>Hugo -- gohugo.io</generator>
    <language>en-us</language>
    <managingEditor>postmaster</managingEditor>
    <lastBuildDate>Wed, 03 Jun 2026 14:55:08 +0200</lastBuildDate><atom:link href="https://cncz.science.ru.nl/tags/ai/index.xml" rel="self" type="application/rss+xml" />
    <item>
      <title>Update on on-premise LLM service</title>
      <link>https://cncz.science.ru.nl/news/2026-06-04-local_llm_update/</link>
      <pubDate>Wed, 03 Jun 2026 14:55:08 +0200</pubDate>
      <author>Miek Gieben</author>
      <guid>https://cncz.science.ru.nl/news/2026-06-04-local_llm_update/</guid>
      <description>&lt;p&gt;In a &lt;a href=&#34;https://cncz.science.ru.nl/news/2026-02-05_local_llm_chat/&#34;&gt;previous news article&lt;/a&gt; we announced the availability of a local LLM
with no connection to big tech. We haven&amp;rsquo;t really advertised it further, but we can say it has proven popular.&lt;/p&gt;
&lt;p&gt;Part of the reason to launch was to discover bottlenecks, well, we&amp;rsquo;ve discovered bottlenecks.&lt;/p&gt;
&lt;p&gt;Currently there are ~200 registrations and ~10 people online concurrently. Heavy users may have noticed
that the service can be slow at times.&lt;/p&gt;
&lt;h2 id=&#34;plans&#34;&gt;Plans&lt;/h2&gt;
&lt;p&gt;In order to make more efficient use of our server, we will:&lt;/p&gt;</description>
    </item>
    
    <item>
      <title>New on-premise LLM service, available for all Science Students and Staff</title>
      <link>https://cncz.science.ru.nl/news/2026-02-05_local_llm_chat/</link>
      <pubDate>Thu, 05 Feb 2026 09:12:56 +0100</pubDate>
      <author>Miek Gieben</author>
      <guid>https://cncz.science.ru.nl/news/2026-02-05_local_llm_chat/</guid>
      <description>&lt;p&gt;At the start of this year we installed a high-performance server (512 GB memory, two high-end GPUs). This machine now hosts local large-language models &lt;a href=&#34;https://en.wikipedia.org/wiki/Large_language_model&#34; target=&#34;_blank&#34;&gt;LLMs&lt;/a&gt;, meaning the AI runs entirely on‑premise and your data never leaves our network.&lt;/p&gt;
&lt;h2 id=&#34;what-you-can-do-today&#34;&gt;What you can do today&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;Access the web UI at &lt;a href=&#34;https://chat.science.ru.nl&#34; target=&#34;_blank&#34;&gt;chat.science.ru.nl&lt;/a&gt;. Log in with your Science RU username (e.g., jdoe, not &lt;a href=&#34;mailto:jdoe@science.ru.nl&#34;&gt;jdoe@science.ru.nl&lt;/a&gt;).&lt;/li&gt;
&lt;li&gt;The service is powered by &lt;a href=&#34;https://ollama.com/&#34; target=&#34;_blank&#34;&gt;Ollama&lt;/a&gt; and &lt;a href=&#34;https://openwebui.com/&#34; target=&#34;_blank&#34;&gt;Open
WebUI&lt;/a&gt;, providing a clean, responsive chat interface.&lt;/li&gt;
&lt;li&gt;Early tests show reliable &lt;em&gt;speech‑to‑text&lt;/em&gt; (including Dutch) and the ability to &lt;em&gt;review source code&lt;/em&gt;. A user experience that feels very close to the larger commercial models, even though we are currently using open‑source models a few generations older.&lt;/li&gt;
&lt;/ul&gt;
&lt;h2 id=&#34;current-status---experimental&#34;&gt;Current status - experimental&lt;/h2&gt;
&lt;ul&gt;
&lt;li&gt;We don&amp;rsquo;t know how many concurrent users the system can support and how much persistent storage will be required.&lt;/li&gt;
&lt;li&gt;No automated backups are in place; please treat the data you store there as temporary.&lt;/li&gt;
&lt;/ul&gt;
&lt;ul&gt;
&lt;li&gt;Uptime is not guaranteed. The service may need to be restarted or the server rebooted when we apply configuration changes or updates.&lt;/li&gt;
&lt;/ul&gt;
&lt;ul&gt;
&lt;li&gt;The setup will evolve as we gather feedback, so your input is valuable.&lt;/li&gt;
&lt;/ul&gt;
&lt;h2 id=&#34;which-models-are-available&#34;&gt;Which models are available?&lt;/h2&gt;
&lt;p&gt;Browse the full catalog on &lt;a href=&#34;https://ollama.com/search&#34; target=&#34;_blank&#34;&gt;ollama.com/search&lt;/a&gt;. At the moment we run two modes:
&lt;strong&gt;&lt;em&gt;gpt‑oss&lt;/em&gt;&lt;/strong&gt; and &lt;strong&gt;&lt;em&gt;deepseek-r1&lt;/em&gt;&lt;/strong&gt; , but others can be added on request. We ran &lt;strong&gt;&lt;em&gt;deepseek-r1:671b&lt;/em&gt;&lt;/strong&gt;, but
this model is too large for the machine we have.&lt;/p&gt;</description>
    </item>
    
  </channel>
</rss>
