ToolAI.io · GitHub-kanaal

Trending AI-projecten op GitHub

Volg actieve open-source AI-repositories rond LLM, agents, MCP, RAG en coderen, met sterren en geverifieerde projectgegevens.

Openbare repositorygegevens

Projectindex

41 Projecten
Computer vision

ultralytics

ultralytics/ultralytics

Training and inference tools for object detection, segmentation and pose estimation models.

★ 61,3K⑂ 11,7KPython
AGPL-3.0Q90
Inferentie, implementatie en runtime

ray

ray-project/ray

A distributed computing runtime for machine learning training, tuning and model serving.

★ 43,7K⑂ 8KPython
Apache-2.0Q85
Inferentie, implementatie en runtime

ollama

ollama/ollama

A tool for running and managing large language models locally.

★ 180,2K⑂ 17,7KGo
MITQ90
Inferentie, implementatie en runtime

LocalAI

mudler/LocalAI

A local inference engine for self-hosting models and AI services.

★ 48,9K⑂ 4,4KGo
MITQ85
Inferentie, implementatie en runtime

jan

janhq/jan

A local AI chat application for running models on a personal computer.

★ 44,3K⑂ 3KTypeScript
Licentie niet gedetecteerdQ80
Computer vision

yolov5

ultralytics/yolov5

A PyTorch-based object detection project with training, inference and export workflows.

★ 58K⑂ 17,5KPython
AGPL-3.0Q90
Inferentie, implementatie en runtime

llmgateway

theopenco/llmgateway

Route, manage, and analyze your LLM requests across multiple providers with a unified API interface.

★ 1,6K⑂ 172TypeScript
NOASSERTIONQ98
Inferentie, implementatie en runtime

sglang

sgl-project/sglang

SGLang is a high-performance serving framework for large language models and multimodal models.

★ 32,2K⑂ 8,1KPython
Apache-2.0Q98
Fine-tuning, training en data

DeepSpeed

deepspeedai/deepspeed

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

★ 43K⑂ 4,9KPython
Apache-2.0Q94
Inferentie, implementatie en runtime

vllm

vllm-project/vllm

A high-throughput and memory-efficient inference and serving engine for LLMs

★ 89,5K⑂ 21KPython
Apache-2.0Q98
Schermafbeelding van TensorRT-LLM — Optimized LLM inference
Inferentie, implementatie en runtime

TensorRT-LLM — Optimized LLM inference

NVIDIA/TensorRT-LLM

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way

★ 14,4K⑂ 2,7KC++
NOASSERTIONQ82

Pagina 1 / 4 · 41 projecten

Recent bijgewerkt

cherry-studioCherryHQ/cherry-studio★ 51,5K siyuansiyuan-note/siyuan★ 46,2K career-opscareer-ops-hq/career-ops★ 70,2K tensorflowtensorflow/tensorflow★ 198,8K streamlitstreamlit/streamlit★ 45,7K pytorchpytorch/pytorch★ 102,8K

Meeste sterren

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K