Skip to content

Aqueduct Release Notes

2026-09-17

Our Qwen 3.5 Model will be retired on 30 September .

Replacement models

DeepSeek-V4-Flash is now supported by the RTX 6000s GPUs. This allows us to expand its deployment on that node and use it as the replacement for Qwen 3.5.

Moving the current DeepSeek model off the H200 node will also free up capacity there. We are considering Qwen3.8-27B for that node. As our most powerful node, the H200 could host four instances and help us scale our small-model offering.

Possible future releases

We are monitoring two candidates for a later upgrade on the RTX 6000 node:

Either model could be a follow-up option, depending on compatibility and testing. There is no confirmed upgrade date yet.

Further, we are also still working with the ASC team to resolve the issues with the AMD node. Once it is operational, DeepSeek-V4.1-Flash or GLM-5.3 are possible deployment candidates, subject to compatibility and testing.


(Deutsche Version)

Unser Modell Qwen 3.5 wird am 30. September außer Betrieb genommen.

Nachfolgemodelle

DeepSeek-V4-Flash läuft inzwischen auf den RTX-6000-GPUs. Damit können wir den Einsatz auf diesem Knoten ausweiten und das Modell als Ersatz für Qwen 3.5 anbieten.

Durch den Umzug des aktuellen DeepSeek-Modells werden auf dem H200-Knoten Kapazitäten frei. Für diesen Knoten ziehen wir Qwen3.8-27B in Betracht. Als unser leistungsstärkster Knoten könnte der H200 vier Instanzen des Modells betreiben und damit die Kapazität unseres Angebots an kleineren Modellen weiter erhöhen.

Mögliche zukünftige Modelle

Wir beobachten zwei Kandidaten für ein späteres Upgrade auf dem RTX-6000-Knoten:

Beide Modelle kommen als Nachfolger infrage, sofern sie mit unserer Hardware kompatibel sind und die Tests erfolgreich verlaufen. Ein Termin für das Upgrade steht noch nicht fest.

Außerdem arbeiten wir weiterhin mit dem ASC-Team daran, die Probleme mit dem AMD-Knoten zu beheben. Sobald dieser einsatzbereit ist, kommen DeepSeek-V4.1-Flash oder GLM-5.3 für den Einsatz dort infrage, ebenfalls abhängig von Kompatibilität und Testergebnissen.

2026-08-07

DeepSeek V4 Flash is now available in Aqueduct as deepseek-v4-flash-284b.

It shows strong benchmark performance, especially for its size. That combination makes it a very strong candidate right now.

One thing to note: DeepSeek does not support image inputs. This means we will lose that modality for our large model. Image input will still be possible via our Qwen 3.6 model. Feel free to provide us with feedback if this breaks any of your current use cases.

Model details:

Rollout plan:

  • DeepSeek V4 Flash now runs in parallel on our H200s while Qwen 3.5 continues on the RTX 6000 Pros
  • Currently, there are upstream issues in vLLM w.r.t. the deployment of DeepSeek V4 on RTX Pro 6000 Blackwell.
  • Qwen 3.5 will eventually be fully deprecated once DeepSeek runs on RTX Pro 6000 Blackwell. DeepSeek will then inherit the "main" model alias.
  • We will update you with a concrete timeline as soon as we have a clear picture, but not earlier than 21st of August.

(deutsche Version)

DeepSeek V4 Flash ist jetzt in Aqueduct als deepseek-v4-flash-284b verfügbar.

Es zeigt eine starke Benchmark-Performance, besonders im Verhältnis zu seiner Größe. Diese Kombination macht es aktuell zu einer sehr starken Wahl.

Ein Hinweis dazu: DeepSeek unterstützt keine Bild-Inputs. Das bedeutet, dass wir diese Modalität bei unserem großen Modell verlieren. Bild-Input wird weiterhin über unser Qwen 3.6-Modell möglich sein. Gebt uns gerne Feedback, falls dies eure aktuellen Use Cases beeinträchtigt.

Modell-Details:

Rollout-Plan:

  • DeepSeek V4 Flash läuft nun parallel auf unseren H200s, während Qwen 3.5 weiterhin auf den RTX 6000 Pros läuft
  • Aktuell gibt es Upstream-Probleme in vLLM bezüglich des Deployments von DeepSeek V4 auf RTX Pro 6000 Blackwell.
  • Qwen 3.5 wird schließlich vollständig deprecated, sobald DeepSeek auf RTX Pro 6000 Blackwell läuft. DeepSeek übernimmt dann den "main"-Modell-Alias.
  • Wir informieren euch mit einem konkreten Zeitplan, sobald wir ein klares Bild haben, jedoch nicht vor dem 21. August.