STACKIT RAG Template
stackitcloud/rag-template
A template for building AI chatbots and document management systems using Retrieval-Augmented Generation (RAG), vector search, and FastAPI, designed for deployment on Kubernetes.
ToolAI.io · Chaîne GitHub
Découvrez les nouveaux projets AI open source ajoutés sur GitHub, avec les informations du dépôt, les instructions de configuration, les licences et les ressources associées.
Données du dépôt public
stackitcloud/rag-template
A template for building AI chatbots and document management systems using Retrieval-Augmented Generation (RAG), vector search, and FastAPI, designed for deployment on Kubernetes.
taichuy/1flowbase
An open-source AI gateway that allows local agent clients to publish fusion-style multi-model workflows as OpenAI and Claude-compatible virtual models, providing full observability into traces, tokens, latency, and costs.
sno-ai/llmix
A production LLM call layer for AI agents and tools that wraps existing provider SDKs with config-driven model presets, caching, resilience patterns, and key rotation across Python, TypeScript, and Rust.
askimo-ai/askimo
A native desktop AI client for chat, local RAG, multi-step AI workflows (Plans), and agent skills, supporting multiple cloud and local LLM providers while keeping user files strictly on the machine.
deepseek-ai/DeepSeek-V3
An open-source project focused on large language model research and inference.
deepseek-ai/DeepSeek-R1
An open-source project focused on reasoning model research and evaluation.
QwenLM/Qwen3
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud
QwenLM/Qwen3-Coder
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team
deepseek-ai/DeepSeek-Coder
DeepSeek Coder: Let the Code Write Itself
NVIDIA/TensorRT-LLM
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way
NVIDIA/garak
the LLM vulnerability scanner
openai/evals
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks
Page 6 / 7 · 75 projets