Model Serving · Model asset
Edge Vision-Language Model
Model assetModel ServingModelsarc:EdgeVisionLanguageModel
A compact vision-language model (single-digit billions of parameters) sized to run on embedded edge GPU devices.
Responsibility. Answers questions about images on-device.
Also known as: Neva-7B
Variant of Vision-Language Model abstract
When to choose. Choose for edge deployment on embedded GPU devices where memory and power are constrained.
Relationships
deployed on structural
alternative to variability
Quantitative guidance
As stated by the sources; verify before use.
- 7 billion parameters (Ch7.5).
Classification
- Technologies
- NVIDIA Neva-7BNVIDIA Jetson AGX Orin
- Quality attributes
- Performance efficiency (ISO/IEC 25010)Flexibility (ISO/IEC 25010)
Sources
- Ch7.5: T. Nguyen, "NeMo Curator, Riva Speech AI & Multimodal Integration," in Mastering Agentic AI Systems: Guide for the NVIDIA NCP-AAI Exam, 1st ed. 2026, ch. 7.5. ISBN: 9798244538229.