Marketplace
By Use Case Marketplace
Compare By Use Case products across curated subcategories and trusted providers
- Products
- 15,413
- Subcategories
- 698
Selected subcategory
Model Deployment
Compare By Use Case products across curated subcategories and trusted providers
Explore By Use Case with structured category paths, practical filters, and independent product data
Showing 1–12 of 76
by Xilinx
HLS based Deep Neural Network Accelerator Library for Xilinx Ultrascale+ MPSoCs
by Soul AI Lab
SoulX-FlashTalk is the first 14B model to achieve sub-second start-up latency (0.87s) while maintaining a real-time throughput of 32 FPS on an 8xH800 node.
by Gabriel Vergnaud
The exhaustive Pattern Matching library for TypeScript, with smart type inference.
by NobodyWho
NobodyWho is an inference engine that lets you run LLMs locally and efficiently on any device.
by Salvatore Sanfilippo
DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm
by Deep Cognition and Language Research (DeCLaRe) Lab
This repository contains the dataset and the PyTorch implementations of the models from the paper Recognizing Emotion Cause in Conversations.
by Microsoft
Official inference framework for 1-bit LLMs
by ggml
Port of OpenAI's Whisper model in C/C++
by Mirai Labs
A high-performance inference engine for AI models
by SiliconFlow
OneDiff: An out-of-the-box acceleration library for diffusion models.
by SemiAnalysisAI
Open-Source Agentic Inference Benchmark | InferenceX
by tile-ai
Tile-Based Runtime for Ultra-Low-Latency LLM Inference