Image-to-Text
Transformers
Safetensors
English
cambrian_qwen
text-generation
multimodal
video-understanding
spatial-reasoning
vision-language
Eval Results (legacy)
Instructions to use nyu-visionx/Cambrian-S-7B with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Transformers
How to use nyu-visionx/Cambrian-S-7B with Transformers:
# Use a pipeline as a high-level helper # Warning: Pipeline type "image-to-text" is no longer supported in transformers v5. # You must load the model directly (see below) or downgrade to v4.x with: # 'pip install "transformers<5.0.0' from transformers import pipeline pipe = pipeline("image-to-text", model="nyu-visionx/Cambrian-S-7B")# Load model directly from transformers import AutoModelForCausalLM model = AutoModelForCausalLM.from_pretrained("nyu-visionx/Cambrian-S-7B", device_map="auto") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -140,7 +140,7 @@ language:
|
|
| 140 |
|
| 141 |
# Cambrian-S-7B
|
| 142 |
|
| 143 |
-
**[Website](https://
|
| 144 |
|
| 145 |
**Authors**: [Shusheng Yang*](https://github.com/vealocia), [Jihan Yang*](https://jihanyang.github.io/), [Pinzhi Huang†](https://pinzhihuang.github.io/), [Ellis Brown†](https://ellisbrown.github.io/), et al.
|
| 146 |
|
|
|
|
| 140 |
|
| 141 |
# Cambrian-S-7B
|
| 142 |
|
| 143 |
+
**[Website](https://cambrian-mllm.github.io/cambrian-s/)** | **[Paper](https://arxiv.org/abs/2511.04670)** | **[GitHub](https://github.com/cambrian-mllm/cambrian-s)** | **[Cambrian-S Family](https://huggingface.co/collections/nyu-visionx/cambrian-s-models)**
|
| 144 |
|
| 145 |
**Authors**: [Shusheng Yang*](https://github.com/vealocia), [Jihan Yang*](https://jihanyang.github.io/), [Pinzhi Huang†](https://pinzhihuang.github.io/), [Ellis Brown†](https://ellisbrown.github.io/), et al.
|
| 146 |
|