ToolAI.io · Kanal GitHub

Direktori proyek AI di GitHub

Jelajahi direktori lengkap ToolAI yang berisi proyek AI sumber terbuka di GitHub, yang diatur berdasarkan topik, bahasa, dan lisensi.

Data repositori publik

Indeks proyek

15 Proyek
Tangkapan layar promptfoo
Agen & Multi-Agen

promptfoo

promptfoo/promptfoo

An open-source CLI and library for evaluating, testing, and red-teaming LLM applications, RAGs, and agents. It enables side-by-side model comparison and vulnerability scanning with declarative configs and CI/CD integration.

★ 23,9K⑂ 2,2KTypeScript
MITQ98
Agen & Multi-Agen

Opik: Open-Source LLM Observability, Evaluation and Agent Tracing

comet-ml/opik

Opik is an Apache-2.0-licensed platform for tracing, evaluating, debugging and monitoring LLM applications, RAG systems and multi-step agent workflows. It supports self-hosting, a hosted Comet.com option, client SDKs, a REST API, automated evaluations, prompt experimentation and production dashboards.

★ 21,1K⑂ 1,7KPython
Apache-2.0Q98
Evaluasi, Observability & Keamanan

SkillSpector

NVIDIA/SkillSpector

Security scanner for AI agent skills. Detect vulnerabilities, malicious patterns, and security risks.

★ 13,8K⑂ 1,1KPython
Apache-2.0Q98
Tangkapan layar BISHENG
Agen & Multi-Agen

BISHENG

dataelement/bisheng

BISHENG is an open-source LLM application DevOps platform designed for next-generation enterprise AI applications, offering comprehensive features like GenAI workflow orchestration, RAG, Agent management, and enterprise-grade system controls.

★ 11,8K⑂ 1,9KPython
Apache-2.0Q98
Agen & Multi-Agen

LangWatch

langwatch/langwatch

An open-core platform for evaluating, testing, tracing, and monitoring LLM applications and AI agents before release and in production.

★ 3,5K⑂ 345TypeScript
Apache-2.0Q98
Tangkapan layar TruLens: Evaluation and Tracking for LLM Experiments and AI Agents
Agen & Multi-Agen

TruLens: Evaluation and Tracking for LLM Experiments and AI Agents

truera/trulens

TruLens is an open-source, OpenTelemetry-native evaluation and tracking library for LLM applications and AI agents, enabling developers to trace every step, score performance with LLM judges, and compare app versions.

★ 3,5K⑂ 319Python
MITQ98
Tangkapan layar EvalScope
RAG & sistem pengetahuan

EvalScope

modelscope/evalscope

A streamlined and customizable framework for efficient large model (LLM, VLM, AIGC) evaluation and performance benchmarking.

★ 3,2K⑂ 440Python
Apache-2.0Q98
Agen & Multi-Agen

Claude Cookbooks: From API Examples to Evaluatable Business Assistants

anthropics/claude-cookbooks

Anthropic’s collection of Claude development examples covers classification, summarization, retrieval augmentation, tool use, image understanding, and evaluation. It is suitable for learning from specific tasks and adapting the code.

★ 52,6K⑂ 6,3KJupyter Notebook
MITQ92
Evaluasi, Observability & Keamanan

netdata

netdata/netdata

Real-time system and application monitoring with AI-assisted operational analysis.

★ 80,4K⑂ 6,6KGo
GPL-3.0Q90
Fine-tuning, Pelatihan & Data

scikit-learn

scikit-learn/scikit-learn

A Python machine learning library for classification, regression, clustering and model evaluation.

★ 67,2K⑂ 27,4KPython
BSD-3-ClauseQ90
Evaluasi, Observability & Keamanan

strix

usestrix/strix

An AI-assisted vulnerability discovery tool for authorized application security testing.

★ 60,6K⑂ 6,6KPython
Apache-2.0Q90

Halaman 1 / 2 · 15 proyek

Baru diperbarui

TypeSafe Python SDKtypesafe-ai/typesafe-sdk-python★ 218 TypeSafe JavaScript SDKtypesafe-ai/typesafe-sdk-js★ 231 TypeSafe Agent Skillstypesafe-ai/skills★ 2K cherry-studioCherryHQ/cherry-studio★ 51,5K onyxonyx-dot-app/onyx★ 32K siyuansiyuan-note/siyuan★ 46,2K

Paling banyak diberi bintang

ECCaffaan-m/ECC★ 248,8K Hermes Agentnousresearch/hermes-agent★ 227,1K tensorflowtensorflow/tensorflow★ 198,8K AutoGPTSignificant-Gravitas/AutoGPT★ 187,1K ollamaollama/ollama★ 180,2K markitdown — Document conversion and extractionmicrosoft/markitdown★ 174,9K