The Triton Inference Server provides an optimized cloud and edge inferencing solution.
Open sourceAI
100/100
needle
Automation foundation model for tiny devices: 2-bit, 8-29 MB, tool calls, structured extraction and embeddings on phones, wearables, smart homes, robots, cars a
Open sourceAI
100/100
dlib
A toolkit for making real world machine learning and data analysis applications in C++