RadarInference
Cactus
Inference · Tool
About Cactus
A hybrid edge-cloud inference engine with OpenAI-compatible APIs, custom ARM kernels, and 1-4-bit quantization, so models can run locally on mobile hardware. Native bindings cover Swift, Kotlin, Flutter, React Native, Python, and Rust. Fills the on-device slot in an otherwise cloud-only inference field.
Run text, speech, and vision models on-device on phones and wearables.
CategoryInference
TypeTool
WebsiteSimilar tools
Find more on the Radar
Browse every tool, model, and MCP server in one place.