ToolAI.io · GitHub چینل

GitHub پر ٹرینڈنگ AI پروجیکٹس

LLM، ایجنٹ، MCP، RAG اور کوڈنگ کے موضوعات میں فعال اوپن سورس AI ریپوزٹریز کو اسٹارز اور تصدیق شدہ پروجیکٹ معلومات کے ساتھ ٹریک کریں۔

عوامی ریپوزٹری کا ڈیٹا

پروجیکٹ انڈیکس

41 پروجیکٹس
hal0 کا اسکرین شاٹ
انفرنس، ڈپلائمنٹ اور رَن ٹائم

hal0

hal0ai/hal0

An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.

★ 68⑂ 7Python
Apache-2.0Q89
انفرنس، ڈپلائمنٹ اور رَن ٹائم

xLLM: High-Performance Inference Engine for Diverse AI Accelerators

xllm-ai/xllm

xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.

★ 1.5K⑂ 279C++
Apache-2.0Q98
انفرنس، ڈپلائمنٹ اور رَن ٹائم

vLLM Ascend

vllm-project/vllm-ascend

A community-maintained hardware plugin that enables vLLM to run model inference and serving workloads on supported Ascend NPUs.

★ 2.7K⑂ 2.1KC++
Apache-2.0Q98
ایجنٹس اور ملٹی ایجنٹ

dstack

dstackai/dstack

A vendor-agnostic control plane for provisioning and orchestrating training, inference, development, and agentic workloads across GPU clouds, Kubernetes, and on-premises infrastructure.

★ 2.2K⑂ 249Python
MPL-2.0Q98
memra کا اسکرین شاٹ
انفرنس، ڈپلائمنٹ اور رَن ٹائم

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 314⑂ 35Rust
MITQ92
انفرنس، ڈپلائمنٹ اور رَن ٹائم

any-llm: A Unified Python Interface for LLM Providers

mozilla-ai/any-llm

any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.

★ 2.2K⑂ 211Python
Apache-2.0Q98
tiktoken — Fast tokenization کا اسکرین شاٹ
انفرنس، ڈپلائمنٹ اور رَن ٹائم

tiktoken — Fast tokenization

openai/tiktoken

tiktoken is a fast BPE tokeniser for use with OpenAI's models

★ 19K⑂ 1.6KRust
MITQ86
Xorbits Inference (Xinference) کا اسکرین شاٹ
انفرنس، ڈپلائمنٹ اور رَن ٹائم

Xorbits Inference (Xinference)

xorbitsai/inference

A powerful and versatile library designed to serve language, speech recognition, and multimodal models. It allows users to swap GPT for any LLM by changing a single line of code and run models on cloud, on-prem, or locally via a unified, production-ready inference API.

★ 9.5K⑂ 853Python
Apache-2.0Q98
Plano: AI-Native Proxy Server and Data Plane for Agentic Apps کا اسکرین شاٹ
ایجنٹس اور ملٹی ایجنٹ

Plano: AI-Native Proxy Server and Data Plane for Agentic Apps

katanemo/plano

Plano is an AI-native proxy server and data plane built in Rust that centralizes LLM routing, agent orchestration, observability, and guardrails, allowing developers to focus on core agent logic rather than infrastructure plumbing.

★ 7K⑂ 484Rust
Apache-2.0Q98
Hugging Face Transformers کا اسکرین شاٹ
انفرنس، ڈپلائمنٹ اور رَن ٹائم

Hugging Face Transformers

huggingface/transformers

A model-definition framework for state-of-the-art machine learning models across text, vision, audio, and multimodal domains, supporting both inference and training.

★ 163.3K⑂ 34.1KPython
Apache-2.0Q98
CSGHub: Open-Source LLM Asset Management Platform کا اسکرین شاٹ
ایجنٹس اور ملٹی ایجنٹ

CSGHub: Open-Source LLM Asset Management Platform

opencsgs/csghub

CSGHub is an open-source, on-premise platform for managing the full lifecycle of Large Language Model assets, including models, datasets, spaces, and code, offering functionality comparable to a private Hugging Face.

★ 4.2K⑂ 525Vue
Apache-2.0Q98
Lemonade: Local AI Server for GPU and NPU Inference کا اسکرین شاٹ
MCP اور ٹول کالنگ

Lemonade: Local AI Server for GPU and NPU Inference

lemonade-sdk/lemonade

Lemonade is an open-source local AI server that enables users to run optimized Large Language Models (LLMs), speech, and image generation models directly on their own GPUs and NPUs, providing a free and private alternative to cloud APIs.

★ 5.2K⑂ 436C++
Apache-2.0Q98

صفحہ 2 / 4 · 41 پروجیکٹس

حال ہی میں اپ ڈیٹ کیا گیا

cherry-studioCherryHQ/cherry-studio★ 51.5K siyuansiyuan-note/siyuan★ 46.2K career-opscareer-ops-hq/career-ops★ 70.2K tensorflowtensorflow/tensorflow★ 198.8K streamlitstreamlit/streamlit★ 45.7K pytorchpytorch/pytorch★ 102.8K

سب سے زیادہ اسٹارز والے

ECCaffaan-m/ECC★ 248.8K Hermes Agentnousresearch/hermes-agent★ 227.1K tensorflowtensorflow/tensorflow★ 198.8K AutoGPTSignificant-Gravitas/AutoGPT★ 187.1K ollamaollama/ollama★ 180.2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174.9K