<?xml version="1.0" encoding="utf-8" standalone="yes"?><rss version="2.0" xmlns:atom="http://www.w3.org/2005/Atom"><channel><title>Technical Documentation :: Documentation for AI Services</title><link>https://docs.ai.gwdg.de/en/technical/index.html</link><description>Info We will soon be adding additional documentation here, alongside the SAIA Platform.</description><generator>Hugo</generator><language>en</language><atom:link href="https://docs.ai.gwdg.de/en/technical/index.xml" rel="self" type="application/rss+xml"/><item><title>SAIA Platform</title><link>https://docs.ai.gwdg.de/en/technical/saia-platform/index.html</link><pubDate>Mon, 01 Jan 0001 00:00:00 +0000</pubDate><guid>https://docs.ai.gwdg.de/en/technical/saia-platform/index.html</guid><description>SAIA is the Scalable Artificial Intelligence (AI) Accelerator that hosts our AI services. Such services include Chat AI and CoCo AI, with more to be added soon. SAIA API (application programming interface) keys can be requested and used to access the services from within your code.
API keys are not necessary to use the Chat AI web interface.
The SAIA API is suitable for interactive inference scenarios. If you have a large amount (eg. thousands of LLM queries) of requests that you can process asynchronously, the batch paradigm of our HPC cluster is the better choice. Your batch will be completed more predictably, in less time, and with lower cost. Check out how to get started with our HPC cluster and then running LLMs to learn how you can setup up a batch inference job on the cluster. vLLM is another popular choice for LLM inference.</description></item></channel></rss>