Files
2026-09-29 22:00:35 +00:00

41 lines
1.8 KiB
Markdown

---
title: "noris AI Platform"
category: systems
tags: [ai, llm, noris, gpu, embeddings]
created: "2026-09-27"
modified: "2026-09-29"
---
# noris AI Platform (ai.noris.de)
> Interne AI-Plattform der noris Network AG. Bereitstellung von LLMs, Embeddings und Image Generation.
## Endpoints
- **Chat:** `https://ai.noris.de/v1/chat/completions`
- **Embeddings:** `https://ai.noris.de/v1/embeddings`
- **Images:** ⚠️ `/v1/images/generations` wird vom Bifrost Gateway **NICHT** unterstützt. Image-Gen-Modelle (qwen-image-2-1) werden über `/v1/chat/completions` angesprochen — das Bild kommt als base64-PNG im `content`-Array zurück (Typ `image_url`, `data:image/png;base64,...`). Siehe `references/vllm-image-generation.md` im Skill `serving-llms-vllm`.
## Modelle
| Typ | Modell-ID | Hinweise |
|-----|-----------|----------|
| Flagship LLM | `glm-5-2` | Primary, OpenRouter-kompatibel |
| General | `gemma-4-31b-it` | Vision-fähig, genutzt von LLM Vision |
| Large MoE | `gpt-oss-120b` | |
| Mid-range | `qwen3.6-27b` | |
| Mid-range | `qwen3.8-27b` | |
| Fast | `ds-v4-flash` | Low-latency, Paperless OCR |
| Embedding | `harrier` | Vektorembeddings |
| Image Gen | `qwen-image-2-1` | Via `/v1/chat/completions` (NOT images/generations). Base64-PNG im content-Array. ~30s/ Bild. |
## Verbraucher
- **Hermes Agent** — Primärmodell `glm-5-2` via OpenRouter
- **HA LLM Vision** — `gemma-4-31b-it` für Bildanalyse (Frigate Events)
- **Paperless** — `ds-v4-flash` für OCR/Kategorisierung
- **Personal Coach Bot** — `glm-5-2` via noris direkt
- **Dynamic Coach** — `glm-5-2` via noris direkt, `qwen-image-2-1` für Visualisierungen
## Related
- [[systems/frigate]] — nutzt noris AI für Event-Klassifizierung
- [[systems/homeassistant]] — LLM Vision Integration
- [[systems/paperless]] — OCR via ds-v4-flash