ToolAI.io · GitHub चैनल

GitHub पर ओपन-सोर्स AI प्रोजेक्ट्स

LLM, एजेंट, MCP, RAG और AI डेवलपमेंट प्रोजेक्ट्स की तथ्यों पर आधारित इंडेक्स, जिसमें हर प्रविष्टि के लिए लाइसेंस, सेटअप, डाउनलोड, स्क्रीनशॉट और संबंधित संसाधन शामिल हैं।

सार्वजनिक रिपॉज़िटरी का डेटा

प्रोजेक्ट इंडेक्स

41 प्रोजेक्ट्स
Hugging Face Transformers का स्क्रीनशॉट
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

Hugging Face Transformers

huggingface/transformers

A model-definition framework for state-of-the-art machine learning models across text, vision, audio, and multimodal domains, supporting both inference and training.

★ 163.3K⑂ 34.1KPython
Apache-2.0Q98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89.5K⑂ 21KPython
Apache-2.0Q98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

sglang

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

★ 32.2K⑂ 8.1KPython
Apache-2.0Q98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

OpenVINO

openvinotoolkit/openvino

Open-source toolkit for optimizing and deploying AI inference across edge-to-cloud environments.

★ 10.6K⑂ 3.3KC++
Apache-2.0Q98
Xorbits Inference (Xinference) का स्क्रीनशॉट
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

Xorbits Inference (Xinference)

xorbitsai/inference

A powerful and versatile library designed to serve language, speech recognition, and multimodal models. It allows users to swap GPT for any LLM by changing a single line of code and run models on cloud, on-prem, or locally via a unified, production-ready inference API.

★ 9.5K⑂ 853Python
Apache-2.0Q98
Plano: AI-Native Proxy Server and Data Plane for Agentic Apps का स्क्रीनशॉट
एजेंट और मल्टी-एजेंट

Plano: AI-Native Proxy Server and Data Plane for Agentic Apps

katanemo/plano

Plano is an AI-native proxy server and data plane built in Rust that centralizes LLM routing, agent orchestration, observability, and guardrails, allowing developers to focus on core agent logic rather than infrastructure plumbing.

★ 7K⑂ 484Rust
Apache-2.0Q98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

Mooncake: KVCache-Centric Infrastructure for Distributed LLM Serving

kvcache-ai/mooncake

Mooncake is a C++ infrastructure project for large-scale LLM inference and training. It separates prefill, decode, and storage resources while providing high-performance transfer, distributed KV-cache storage, and fault-tolerant expert-parallel communication.

★ 6.1K⑂ 1KC++
Apache-2.0Q98
इन्फरेंस, डिप्लॉयमेंट और रनटाइम

GPUStack

gpustack/gpustack

GPUStack is an open-source GPU cluster manager for serving AI models and provisioning on-demand, SSH-accessible GPU instances. It orchestrates inference engines including vLLM, SGLang, and TensorRT-LLM across on-premises, Kubernetes, and cloud environments.

★ 5.4K⑂ 604Python
Apache-2.0Q98
Lemonade: Local AI Server for GPU and NPU Inference का स्क्रीनशॉट
MCP और टूल कॉलिंग

Lemonade: Local AI Server for GPU and NPU Inference

lemonade-sdk/lemonade

Lemonade is an open-source local AI server that enables users to run optimized Large Language Models (LLMs), speech, and image generation models directly on their own GPUs and NPUs, providing a free and private alternative to cloud APIs.

★ 5.2K⑂ 436C++
Apache-2.0Q98
CSGHub: Open-Source LLM Asset Management Platform का स्क्रीनशॉट
एजेंट और मल्टी-एजेंट

CSGHub: Open-Source LLM Asset Management Platform

opencsgs/csghub

CSGHub is an open-source, on-premise platform for managing the full lifecycle of Large Language Model assets, including models, datasets, spaces, and code, offering functionality comparable to a private Hugging Face.

★ 4.2K⑂ 525Vue
Apache-2.0Q98

पृष्ठ 1 / 4 · 41 प्रोजेक्ट

हाल ही में अपडेट किए गए

cherry-studioCherryHQ/cherry-studio★ 51.5K siyuansiyuan-note/siyuan★ 46.2K career-opscareer-ops-hq/career-ops★ 70.2K tensorflowtensorflow/tensorflow★ 198.8K streamlitstreamlit/streamlit★ 45.7K pytorchpytorch/pytorch★ 102.8K

सर्वाधिक स्टार वाले

ECCaffaan-m/ECC★ 248.8K Hermes Agentnousresearch/hermes-agent★ 227.1K tensorflowtensorflow/tensorflow★ 198.8K AutoGPTSignificant-Gravitas/AutoGPT★ 187.1K ollamaollama/ollama★ 180.2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.9K