Jan-v2-VL-max-Instruct-FP8

Name: Jan-v2-VL-max-Instruct-FP8
Author: janhq

license:apache-2.0

janhq

Image Model

OTHER

30B params

New

12 downloads

Early-stage

Try on Hugging Face Add to Compare

Edge AI:

Mobile

Laptop

Server

68GB+ RAM

Mobile

Laptop

Server

Quick Summary

AI model with specialized capabilities.

Device Compatibility

Mobile

4-6GB RAM

Laptop

16GB RAM

Server

GPU

Minimum Recommended

28GB+ RAM

Code Examples

Exact versions used in our evalsbashvllm

# Exact versions used in our evals
pip install vllm==0.12.0
pip install transformers==4.57.1
pip install "git+https://github.com/vllm-project/llm-compressor.git@1abfd9eb34a2941e82f47cbd595f1aab90280c80"

bashvllm

vllm serve Menlo/Jan-v2-VL-max-Instruct-FP8 \
    --host 0.0.0.0 \
    --port 1234 \
    -dp 1 \
    --enable-auto-tool-choice \
    --tool-call-parser hermes

Deploy This Model

Production-ready deployment in minutes

Together.ai

Instant API access to this model

Fastest API

Production-ready inference API. Start free, scale to millions.

Try Free API

Replicate

One-click model deployment

Easiest Setup

Run models in the cloud with simple API. No DevOps required.

Deploy Now

Disclosure: We may earn a commission from these partners. This helps keep LLMYourWay free.