# Viki Clip Models

URL: https://interfaze.ai/models/jnurikviki-clip-models

[All models](https://interfaze.ai/models)

Viki Clip Models by jnurik, a image-to-text model. Understand and compare features, benchmarks, and capabilities.

## Comparison

| Feature | Viki Clip Models | Interfaze |
| --- | --- | --- |
| Input Modalities | image | image, text, audio, video, document |
| Native OCR | No | Yes |
| Long Document Processing | No | Yes |
| Language Support | unknown | 162+ |
| Native Speech-to-Text | No | Yes |
| Native Object Detection | No | Yes |
| Guardrail Controls | No | Yes |
| Context Input Size | unknown | 1M |
| Tool Calling | No | Tool calling supported + built in browser, code execution and web search |

### Scaling

| Feature | Viki Clip Models | Interfaze |
| --- | --- | --- |
| Scaling | Self-hosted/Provider-hosted with quantization | Unlimited |

[Try Interfaze](https://interfaze.ai/dashboard)[Read the Docs](https://interfaze.ai/docs)

View model card on [Hugging Face](https://huggingface.co/jnurik/viki-clip-models)

🚀 **Live App:** [WikiLens Space](https://huggingface.co/spaces/jnurik/jnurik_img2wiki)

Этот репозиторий содержит веса и FAISS-индексы для приложения WikiLens.

-   **Base Model:** `openai/clip-vit-base-patch32`
-   **Training Data:** 120k Wikipedia photo-article pairs.
-   **Methods:** Zero-shot, Frozen Encoders, DoRA fine-tuning.

## Want more deterministic results?

[Try Interfaze](https://interfaze.ai/dashboard)[Read the Docs](https://interfaze.ai/docs)
