ToolAI.io · Kanal GitHub

Direktori proyek AI di GitHub

Jelajahi direktori lengkap ToolAI yang berisi proyek AI sumber terbuka di GitHub, yang diatur berdasarkan topik, bahasa, dan lisensi.

Data repositori publik

Indeks proyek

24 Proyek
Agen & Multi-Agen

dstack

dstackai/dstack

A vendor-agnostic control plane for provisioning and orchestrating training, inference, development, and agentic workloads across GPU clouds, Kubernetes, and on-premises infrastructure.

★ 2,2K⑂ 249Python
MPL-2.0Q98
Inferensi, Deployment & Runtime

any-llm: A Unified Python Interface for LLM Providers

mozilla-ai/any-llm

any-llm is a Python SDK for communicating with multiple LLM providers through one interface. It supports switching providers and models with minimal code changes while using official provider SDKs when available.

★ 2,2K⑂ 211Python
Apache-2.0Q98
Inferensi, Deployment & Runtime

xLLM: High-Performance Inference Engine for Diverse AI Accelerators

xllm-ai/xllm

xLLM is a C++ inference framework for LLM, VLM, DiT and REC models. It targets high-throughput, low-latency deployment on several AI accelerator families and separates service-layer scheduling and availability from engine-layer computation.

★ 1,5K⑂ 279C++
Apache-2.0Q98
Tangkapan layar memra
Inferensi, Deployment & Runtime

memra

avifenesh/memra

A from-scratch LLM inference engine built in Rust and CUDA, specifically optimized for RTX 5090 (sm_120a) and H100 (sm_90a) architectures, delivering exactness-gated performance without relying on frameworks like ggml.

★ 314⑂ 35Rust
MITQ92
Tangkapan layar hal0
Inferensi, Deployment & Runtime

hal0

hal0ai/hal0

An open-source, self-hosted home AI inference platform designed to turn a Linux box into an OpenAI-compatible inference appliance, with native optimization for AMD Strix Halo hardware.

★ 67⑂ 7Python
Apache-2.0Q89
Tangkapan layar tiktoken — Fast tokenization
Inferensi, Deployment & Runtime

tiktoken — Fast tokenization

openai/tiktoken

tiktoken is a fast BPE tokeniser for use with OpenAI's models

★ 19K⑂ 1,6KRust
MITQ86
Tangkapan layar Qwen3-Coder — Code-focused language models
LLM & Model Fondasi

Qwen3-Coder — Code-focused language models

QwenLM/Qwen3-Coder

Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team

★ 16,8K⑂ 1,2KPython
Lisensi tidak terdeteksiQ82
Tangkapan layar TensorRT-LLM — Optimized LLM inference
Inferensi, Deployment & Runtime

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82

Halaman 2 / 2 · 24 proyek

Baru diperbarui

qwen-code — Command-line coding agentQwenLM/qwen-code★ 27,1K SBproxysoapbucket/sbproxy★ 49 XERJxerj-org/xerj★ 1,4K PwrAgentpwrdrvr/pwragent★ 29 OpenGenicloudgeni-ai/opengeni★ 56 NemoClaw — Agent runtime and deployment toolingNVIDIA/NemoClaw★ 22,2K

Paling banyak diberi bintang

Hermes Agentnousresearch/hermes-agent★ 227,1K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,3K skills — Reusable agent skills and workflowsanthropics/skills★ 170,1K Hugging Face Transformershuggingface/transformers★ 163,3K Firecrawlfirecrawl/firecrawl★ 161,1K LangChainlangchain-ai/langchain★ 143,6K