# MiniCPM RobotManip

URL: https://interfaze.ai/models/openbmbminicpm-robotmanip

MiniCPM RobotManip by openbmb, a robotics model with multimodal capabilities. Understand and compare multimodal features, benchmarks, and capabilities.

## Comparison

| Feature | MiniCPM RobotManip | Interfaze |
| --- | --- | --- |
| Input Modalities | text, image | image, text, audio, video, document |
| Native OCR | No | Yes |
| Long Document Processing | No | Yes |
| Language Support | unknown | 162+ |
| Native Speech-to-Text | No | Yes |
| Native Object Detection | No | Yes |
| Guardrail Controls | No | Yes |
| Context Input Size | unknown | 1M |
| Tool Calling | No | Tool calling supported + built in browser, code execution and web search |

### Scaling

| Feature | MiniCPM RobotManip | Interfaze |
| --- | --- | --- |
| Scaling | Self-hosted/Provider-hosted with quantization | Unlimited |

[Try Interfaze](https://interfaze.ai/dashboard)[Read the Docs](https://interfaze.ai/docs)

View model card on [Hugging Face](https://huggingface.co/openbmb/MiniCPM-RobotManip)

MiniCPM-RobotManip is a 1.5B vision-language-action model for embodied manipulation with the following highlights:

## Benchmark Results

## Inference Example

Please Ensure transformers==5.7.0

## Acknowledgement

## License

Model weights and code are open-sourced under the [Apache-2.0](https://huggingface.co/openbmb/MiniCPM-RobotManip/blob/main/LICENSE) license.

## Want more deterministic results?

[Try Interfaze](https://interfaze.ai/dashboard)[Read the Docs](https://interfaze.ai/docs)
