Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Oryx-1.5-32B-Image – AI Model by THUdyh | AlphaNeural AI
You can deploy this model and start earning money today!
THUdyh
/
Oryx-1.5-32B-Image
like
0
safetensors
llava_qwen
text-generation
conversational
en
zh
THUdyh/Oryx-Image-Data
2409.12961
Qwen/Qwen2.5-32B-Instruct
finetune
apache-2.0
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
Oryx-1.5-7B-Image
Model Summary
The Oryx-Image models are 7/32B parameter models trained based on Qwen2.5 language model with a context window of 32K tokens.
Oryx offers an on-demand solution to seamlessly and efficiently process visual inputs with arbitrary spatial sizes and temporal lengths.
Repository:
https://github.com/Oryx-mllm/Oryx
Languages:
English, Chinese
Paper:
https://arxiv.org/abs/2409.12961
Model Architecture
Architecture:
Pre-trained
Oryx-ViT
+ Qwen2.5-32B
Data:
a mixture of 4M image data
Precision:
BFloat16
Hardware & Software
Hardware:
64 * NVIDIA Tesla A100
Orchestration:
HuggingFace Trainer
Code:
Pytorch
Citation