NVIDIA NeMo Speech
nvidia-nemo/speech
A scalable generative AI framework built for researchers and PyTorch developers working on Automatic Speech Recognition (ASR), Text-to-Speech (TTS), and Speech LLMs.
ToolAI.io · GitHub-Kanal
Durchsuchen Sie das vollständige ToolAI-Verzeichnis mit Open-Source-AI-Projekten auf GitHub, geordnet nach Thema, Sprache und Lizenz.
Daten des öffentlichen Repositorys
nvidia-nemo/speech
A scalable generative AI framework built for researchers and PyTorch developers working on Automatic Speech Recognition (ASR), Text-to-Speech (TTS), and Speech LLMs.
browseros-ai/browseros
BrowserOS is a free, open-source Chromium fork with a built-in AI agent, accompanied by BrowserOS neo, a secondary browser designed specifically for AI agents. It enables local-first browser automation, supporting multiple LLM providers and MCP-compatible tools.
langchain4j/langchain4j
An idiomatic, open-source Java library for building LLM-powered applications on the JVM, offering a unified API over popular LLM providers and vector stores with support for tool calling, MCP, agents, and RAG.
neuml/txtai
txtai is a Python-based AI framework for semantic and vector search, retrieval-augmented generation, LLM orchestration, autonomous agents and multimodal language-model workflows.
dataelement/bisheng
BISHENG is an open-source LLM application DevOps platform designed for next-generation enterprise AI applications, offering comprehensive features like GenAI workflow orchestration, RAG, Agent management, and enterprise-grade system controls.
langchain-ai/open-swe
Open SWE is an open-source framework for building internal coding agents. Built on LangGraph and Deep Agents, it provides cloud sandboxes, Slack/Linear/GitHub invocation, subagent orchestration, and automatic PR creation for engineering organizations.
xorbitsai/inference
A powerful and versatile library designed to serve language, speech recognition, and multimodal models. It allows users to swap GPT for any LLM by changing a single line of code and run models on cloud, on-prem, or locally via a unified, production-ready inference API.
oumi-ai/oumi
Oumi is a Python-based platform for preparing data, training and fine-tuning open-weight foundation models, evaluating results, running inference, and deploying models. It provides configuration recipes and a consistent CLI for local, cluster, and cloud workflows.
maximhq/bifrost
Bifrost is a Go-based AI gateway that provides a unified, OpenAI-compatible API for 23+ AI providers. It supports routing, fallbacks, load balancing, semantic caching, governance, observability, multimodal requests, plugins, and Model Context Protocol integrations.
katanemo/plano
Plano is an AI-native proxy server and data plane built in Rust that centralizes LLM routing, agent orchestration, observability, and guardrails, allowing developers to focus on core agent logic rather than infrastructure plumbing.
clearml/clearml
ClearML is an open-source MLOps/LLMOps suite providing auto-magical CI/CD to streamline AI workloads. It integrates experiment management, data management, pipeline orchestration, scheduling, and model serving into a single comprehensive platform.
plastic-labs/honcho
Memory infrastructure for building stateful AI agents that understand changing people, agents, groups, projects, and ideas over time.
Seite 3 / 7 · 75 Projekte