EmbeddedLLM
Pinned Loading
Repositories
- agentic-api Public Forked from vllm-project/agentic-api
Stateful API logic for agentic applications using vLLM
- skypilot Public Forked from skypilot-org/skypilot
SkyPilot: Run AI and batch jobs on any infra (Kubernetes or 12+ clouds). Get unified execution, cost savings, and high GPU availability via a simple interface.
- JamAIBase Public
The collaborative spreadsheet for AI. Chain cells into powerful pipelines, experiment with prompts and models, and evaluate LLM responses in real-time. Work together seamlessly to build and iterate on AI applications.
- vllm Public Forked from vllm-project/vllm
vLLM: A high-throughput and memory-efficient inference and serving engine for LLMs
- litellm Public Forked from BerriAI/litellm
Python SDK, Proxy Server (LLM Gateway) to call 100+ LLM APIs in OpenAI format - [Bedrock, Azure, OpenAI, VertexAI, Cohere, Anthropic, Sagemaker, HuggingFace, Replicate, Groq]
- Infera Public Forked from AMD-AGI/Infera
More token goodput from frontier models. A distributed, SLA-aware serving mesh — disaggregated prefill/decode, KV-aware routing, and cache offload, tuned to your production SLA.
- fastsafetensors Public Forked from foundation-model-stack/fastsafetensors
High-performance safetensors model loader
- wasm-tools Public Forked from bytecodealliance/wasm-tools
CLI and Rust libraries for low-level manipulation of WebAssembly modules
Top languages
Loading…
Most used topics
Loading…