CactusBrain
Documentation, optimized models, accounts, projects, entitlements, and automatic encrypted model delivery.
discover → integrate → shipInfrastructure for AI that stays on-device
CactusBrain gives developers SDKs and optimized models for private, offline AI on mobile and edge devices. Powered by cellm, our open-source local inference engine.
Inference runs locally. App data is not sent to CactusBrain.
YOUR APPProduct experience
CACTUSBRAIN SDKsVision · Text · more
Powered by cellmLocal inference runtime
CactusBrain brings SDKs, optimized models, and distribution together. cellm powers local execution without becoming another integration developers have to manage.
Documentation, optimized models, accounts, projects, entitlements, and automatic encrypted model delivery.
discover → integrate → shipFocused CactusBrain SDKs with task APIs, validation, and models developers can ship.
SDK → optimized model → resultThe open-source runtime that loads compatible models and executes inference locally.
CPU · Metal · local executionBoundary: CactusBrain does not proxy inference, and the SDKs do not duplicate cellm runtime logic.
Each SDK owns its workflow, validation, and developer experience. cellm remains the execution layer underneath.
On-device visual identification for apps, with matching that stays local.
On-device language model inference for private text features.
cellm is our open-source inference runtime for running optimized AI models directly on mobile and edge devices. CactusBrain SDKs add product workflows and stable APIs on top; inference stays in one inspectable engine.
Vision packages will appear only after a compatible cellm runtime adapter and model artifact are available.
Access the developer workspace and available SDKs.
Choose a target platform and connect model access to an app.
Assign an entitled model package. The SDK delivers and loads it on the device when your feature first runs.