Jina: Jina Clip V1
jina/jina-clip-v1
Jina CLIP v1 revolutionizes multimodal AI by being the first model to excel equally in both text-to-text and text-to-image retrieval tasks. Unlike traditional CLIP models that struggle with text-only scenarios, this model achieves state-of-the-art performance across all retrieval combinations while maintaining a remarkably compact 223M parameter size. The model addresses a critical industry challenge by eliminating the need for separate models for text and image processing, reducing system complexity and computational overhead. For teams building search systems, recommendation engines, or content analysis tools, Jina CLIP v1 offers a single, efficient solution that handles both text and visual content with exceptional accuracy.
Model specifications
- Input
- text
- Output
- text
- Context
- 400 tokens
- Max output
- 10,000 tokens
- Input price
- $0.05 / 1M tokens
- Output price
- $0 / 1M tokens
- Released
- 2026-02-14
Capabilities
- Streaming
- Playground
Provider pricing, discounts and data privacy
Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.
Standard service tier
1 available provider · tier input average $0.05 / 1M tokens
Jina
Tier: Standard · Region: US
Pricing
- Input
- $0.05 / 1M tokens
No provider discount is currently published.
Data privacy and compliance
- Region
- US
- Zero data retention
- No
- Data retention
- Unknown retention
- Used for training
- Unknown
- Data collection
- Moderated
- No
- GDPR compliant
- No
- HIPAA compliant
- No
- SOC 2 certified
- No
- BYOK supported
- No
Privacy policy · Terms · Official website · Documentation · Status · Support
Frequently asked questions
- What is Jina: Jina Clip V1?
- Jina CLIP v1 revolutionizes multimodal AI by being the first model to excel equally in both text-to-text and text-to-image retrieval tasks. Unlike traditional CLIP models that struggle with text-only scenarios, this model achieves state-of-the-art performance across all retrieval combinations while maintaining a remarkably compact 223M parameter size. The model addresses a critical industry challenge by eliminating the need for separate models for text and image processing, reducing system complexity and computational overhead. For teams building search systems, recommendation engines, or content analysis tools, Jina CLIP v1 offers a single, efficient solution that handles both text and visual content with exceptional accuracy.
- How much does Jina: Jina Clip V1 cost?
- Input costs start at $0.05 / 1M tokens and output costs start at $0 / 1M tokens. Provider-level prices vary by service tier.
- What is the context length of Jina: Jina Clip V1?
- Jina: Jina Clip V1 supports a 400 token context window and up to 10,000 output tokens.
- What capabilities does Jina: Jina Clip V1 support?
- Jina: Jina Clip V1 supports Streaming, Playground.
- Which providers offer Jina: Jina Clip V1?
- Jina: Jina Clip V1 is available from Jina.
- How do providers handle data privacy for Jina: Jina Clip V1?
- 1 of 1 providers report that customer data is not used for training, and 0 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.